3-day current streak·11-day longest streak
Hi there! 👋 Welcome to my personal coding playground! I've been fascinated by real-time audio processing and voice technology for years - it's become a hobby of mine. This is…
Hi there! 👋
Welcome to my personal coding playground! I've been fascinated by real-time audio processing and voice technology for years - it's become a hobby of mine. This is where I tinker with ideas, experiment with new approaches, and share what I've learned along the way.
My Weekend Projects & Hobby Code
Over the years, I've built these libraries in my spare time, mostly because I found the problems interesting to solve:
- RealtimeSTT - Real-time speech-to-text transcription
- RealtimeTTS - Real-time text-to-speech synthesis
- RealtimeVoiceChat - Full audio pipeline with minimized latency
What Gets Me Excited In My Free Time
- Tinkering with real-time audio processing - there's something satisfying about getting latency just right
- Playing around with speech recognition and text-to-speech as a personal challenge
- Learning about GPU acceleration through hands-on experimentation
- Contributing back to the community that has taught me so much
- The joy of solving complex technical puzzles in my own time
Current Status
I work on these projects whenever inspiration strikes and I have some time to spare. Always learning something new!
*Code shared here reflects my personal interests and experiments. Feel free to use anything that might be helpful for your own projects.*
-
RealtimeSTT ★ PINNED
A robust, efficient, low-latency speech-to-text library with advanced voice activity detection, wake word activation and instant transcription.
Python ★ 10.0k 1mo agoExplain → -
RealtimeTTS ★ PINNED
Converts text to speech in realtime
Python ★ 4.0k 1mo agoExplain → -
RealtimeVoiceChat ★ PINNED
Have a natural, spoken conversation with AI!
Python ★ 3.8k 1y agoExplain → -
stream2sentence ★ PINNED
Real-time processing and delivery of sentences from a continuous stream of characters or text chunks.
Python ★ 82 4d agoExplain → -
Linguflex ★ PINNED
Command Your World with Voice
Python ★ 811 1y agoExplain → -
LocalAIVoiceChat ★ PINNED
Local AI talk with a custom voice based on Zephyr 7B model. Uses RealtimeSTT with faster_whisper for transcription and RealtimeTTS with Coqui XTTS for synthesis.
Python ★ 726 1y agoExplain → -
AIVoiceChat
Low latency ai companion voice talk in 60 lines of code using faster_whisper and elevenlabs input streaming
Python ★ 320 1y agoExplain → -
TurnVoice
Voice Transformation for Videos. 🎤👄🎬
Python ★ 259 1y agoExplain → -
WhoSpeaks
Efficient approach to speaker diarization using voice characteristics extraction
Python ★ 109 23d agoExplain → -
LocalEmotionalAIVoiceChat
Simulates talk with an AI that can express emotions
Python ★ 88 3mo agoExplain → -
ai_cli_tools
AI at your fingertips: powerful CLI tools for speech, text, and language processing
Python ★ 22 1y agoExplain → -
WhoSpeaksLive
Private, real-time speaker diarization on hardware you control. See who is speaking as it happens, no third-party cloud required.
Python ★ 17 1d agoExplain → -
Revolver
Automatic tool creation and usage
Python ★ 10 11mo agoExplain → -
SentiMind
AI conversation with emotion
Python ★ 6 3mo agoExplain → -
vector_companion_fork ⑂
A local AI companion that uses a collection of free, open source AI models in order to create two virtual companions that will follow your computer journey wherever you go!
Python ★ 6 1y agoExplain → -
KoljaB
No description.
★ 4 1y agoExplain → -
ZipVoice-TTS ⑂
Fast and High-Quality Zero-Shot Text-to-Speech with Flow Matching
★ 3 1y agoExplain → -
remote-cli-assistant
No description.
Python ★ 2 3mo agoExplain → -
oi-fork ⑂
Oi is an open-source cli tool that works on top of codellama and generates code in any editor without extensions.
★ 2 1y agoExplain → -
github-install-assistant
Automated GitHub repository installation. Smart, LLM-driven, platform-aware CLI tool that handles complex PyTorch/CUDA environments.
Python ★ 2 3mo agoExplain → -
ReadAloud
Mark text or url and let it read out loud
Python ★ 2 1y agoExplain → -
audio.cpp ⑂
An all-in-one, pure C++ inference engine for audio models, powered by ggml. Supports TTS, STT, VAD, voice conversion, music generation, and more, with highly optimized performance. No Python dependency.
C++ ★ 1 22d agoExplain → -
privateGPT ⑂
Interact privately with your documents using the power of GPT, 100% privately, no data leaks
★ 1 2y agoExplain → -
Retrieval-based-Voice-Conversion-WebUI ⑂
Voice data <= 10 mins can also be used to train a good VC model!
★ 1 2y agoExplain → -
openai_function_call ⑂
Helper functions to create openai function calls w/ pydantic
★ 1 3y agoExplain → -
coqui-ai-TTS ⑂
🐸💬 - a deep learning toolkit for Text-to-Speech, battle-tested in research and production
Python ★ 1 1y agoExplain → -
openclaw ⑂
Your own personal AI assistant. Any OS. Any Platform. The lobster way. 🦞
★ 0 3mo agoExplain → -
OmniVoice ⑂
High-Quality Voice Cloning TTS for 600+ Languages
★ 0 3mo agoExplain → -
hermes-agent ⑂
The agent that grows with you
★ 0 3mo agoExplain → -
claw-code ⑂
Better Harness Tools, not merely storing the archive of leaked Claude Code but also make shit things done. Now rewriting in Rust.
★ 0 3mo agoExplain → -
chatterbox-streaming ⑂
Streaming and Fine-tuning for Chatterbox TTS
Python ★ 0 1y agoExplain → -
deepspeedpatcher ⑂
A graphical tool to simplify building and installing DeepSpeed 0.15.x or later on Windows systems.
★ 0 1y agoExplain → -
xtts-webui ⑂
Webui for using XTTS and for finetuning it
★ 0 2y agoExplain → -
WhisperLive ⑂
A nearly-live implementation of OpenAI's Whisper.
★ 0 2y agoExplain →
No repos match these filters.
More creators on gitmyhub
douglascrockford standardgalactic AlexTheAnalyst MorvanZhou cloudwu