Seismograph

What’s gaining stars on GitHub right now

OpenBMB/VoxCPM

VoxCPM2: Tokenizer-Free TTS for Multilingual Speech Generation, Creative Voice Design, and True-to-Life Cloning

Generative mediaPython#audio#deeplearning#minicpm#python#pytorch
2.7 Steady Magnitude out of 10. Star growth looks organic. Data as of September 18, 2026.
Open on Seismograph Open on GitHub

About the project

VoxCPM2 is an open-source 2B-parameter tokenizer-free TTS system that synthesizes speech in 30 languages, designs voices from text descriptions, and clones voices from short reference clips with 48kHz output.

Useful for

README summarized by DeepSeek V4.1 Flash. Details may be inaccurate.

Why it’s trending

Stars per day

0200400June 21, 2026September 18, 2026

Bars are daily stars, the line is the usual pace. Red marks spike days.

Numbers

Total stars
37,770
Stars in a day
49
Forks
4,289
Issues and pull requests
397
Watchers
168
Language
Python
License
Apache-2.0
Latest release
2.0.3 · May 11, 2026
Created
September 16, 2025
Last push
September 2, 2026

Star trust

Growth looks organic: forks and discussion are in line with active projects, and stars arrive unevenly, the way people give them.

These are heuristics, not a verdict: we judge by the repository’s behavior, not by a list of stargazers.

Spotted in

Share

README badge

Paste it into your README — the badge shows the current magnitude and links to this page.

Seismograph: 2.7
[![Seismograph](https://gitnova.dev/badge/OpenBMB/VoxCPM.svg?lang=en)](https://gitnova.dev/en/r/OpenBMB/VoxCPM)

Similar projects

  1. 7.9
    mcncarl/jianying-headless

    Local automation tool for Jianying Professional on macOS: builds native drafts from a JSON plan, edits copies of existing multi-track projects, and exports MP4 via the native engine on demand. Requires installed…

    Early signalGenerative mediaPython+751 stars in a day

  2. 6.2
    wide-trace/open-higgsfield

    Open-source self-hosted studio for image and video generation across 32 models (8 image, 24 video) in one interface with a single prompt bar, gallery and per-model settings.

    BreakoutGenerative mediaTypeScript+708 stars in a day

  3. 5.0
    kuhnhomeuk-cell/procedural-film

    An agent skill that turns a topic into a 30-second vertical film: all visuals are drawn on canvas in vanilla JavaScript and sound is synthesised with Web Audio, with no media assets. Output is a single HTML player plus…

    Early signalGenerative mediaJavaScript+44 stars in a day

  4. 4.9
    XGEN-Labs/XGEN-JING

    XGEN-JING is an egocentric interactive experience model built on MiniMax-H3 that generates first-person video and audio for navigation, object interaction, and dialogue from actions, reference images, and observation…

    Early signalGenerative mediaPython+88 stars in a day

  5. 4.8
    jamiepine/voicebox

    A local-first AI voice studio: clone voices, generate speech with 7 TTS engines, and dictate into any app via a global hotkey. An open-source alternative to ElevenLabs and WisprFlow running entirely on your machine.

    BreakoutGenerative mediaTypeScript+193 stars in a day

  6. 4.5
    inikolax/remiqora

    Local GPU-powered music generation studio unifying ACE-Step 1.5 and YuE2-3B in one Vue interface with a multitrack DAW, stem separation, MIDI transcription, and LoRA fine-tuning.

    Early signalGenerative mediaVue+57 stars in a day