jamiepine/voicebox
The open-source AI voice studio. Clone, dictate, create.
About the project
A local-first AI voice studio: clone voices, generate speech with 7 TTS engines, and dictate into any app via a global hotkey. An open-source alternative to ElevenLabs and WisprFlow running entirely on your machine.
Useful for
- Clone a voice from a short sample and synthesize speech in a chosen language
- Dictate text into any field via a global hotkey without cloud services
- Connect an MCP agent to voicebox.speak so it replies in a cloned voice
README summarized by DeepSeek V4.1 Flash. Details may be inaccurate.
Why it’s trending
- 129 stars so far today, about 193 expected by the end of the day.
- Over the last two days the pace is 2.1× that of the previous week and a half.
- The spike has held for 2 days in a row — not a one-off blip.
- GitHub Trending TypeScript today: #16, +458 stars.
- About 42 forks a day — people are taking the code.
Stars per day
Bars are daily stars, the line is the usual pace. Red marks spike days.
Numbers
- Total stars
- 55,090
- Stars in a day
- 193
- Forks
- 6,856
- Issues and pull requests
- 1,101
- Watchers
- 259
- Language
- TypeScript
- License
- MIT
- Latest release
- v0.5.0 · April 25, 2026
- Created
- January 25, 2026
- Last push
- August 9, 2026
Star trust
Growth looks organic: forks and discussion are in line with active projects, and stars arrive unevenly, the way people give them.
These are heuristics, not a verdict: we judge by the repository’s behavior, not by a list of stargazers.
Spotted in
- September 18, 2026GitHub Trending today: #12, +667 stars; GitHub Trending TypeScript today: #3, +458 stars
- September 17, 2026GitHub Trending today: #5, +665 stars; GitHub Trending TypeScript today: #1, +665 stars
- September 16, 2026GitHub Trending today: #5, +409 stars; GitHub Trending TypeScript today: #1, +409 stars
Similar projects
-
7.9
mcncarl/jianying-headless
Local automation tool for Jianying Professional on macOS: builds native drafts from a JSON plan, edits copies of existing multi-track projects, and exports MP4 via the native engine on demand. Requires installed…
-
6.2
wide-trace/open-higgsfield
Open-source self-hosted studio for image and video generation across 32 models (8 image, 24 video) in one interface with a single prompt bar, gallery and per-model settings.
-
5.0
kuhnhomeuk-cell/procedural-film
An agent skill that turns a topic into a 30-second vertical film: all visuals are drawn on canvas in vanilla JavaScript and sound is synthesised with Web Audio, with no media assets. Output is a single HTML player plus…
-
4.9
XGEN-Labs/XGEN-JING
XGEN-JING is an egocentric interactive experience model built on MiniMax-H3 that generates first-person video and audio for navigation, object interaction, and dialogue from actions, reference images, and observation…
-
4.5
inikolax/remiqora
Local GPU-powered music generation studio unifying ACE-Step 1.5 and YuE2-3B in one Vue interface with a multitrack DAW, stem separation, MIDI transcription, and LoRA fine-tuning.
-
4.4
jtydhr88/music-composition-skills
A set of 29 agent skills for Claude Code and OpenAI Codex that help compose and arrange popular music: from a brief the agent builds an ARR-SPEC with key, tempo, section map, harmony and energy curve. That spec is then…