Seismograph

What’s gaining stars on GitHub right now

microsoft/VibeVoice

Open-Source Frontier Voice AI

Generative mediaPython
2.2 Steady Magnitude out of 10 — how fast interest is growing, not a quality score. Star growth looks organic. Data as of October 2, 2026, 13:20 UTC.
Open on Seismograph Open on GitHub

About the project

Microsoft's family of open-source voice AI models: ASR for speech recognition (up to 60 minutes of audio in a single pass with diarization and timestamps) and TTS for long multi-speaker dialogue synthesis. Streaming and edge variants included.

Useful for

README summarized by DeepSeek V4.1 Flash. Details may be inaccurate.

Why it’s trending

Stars per day

0250500July 5, 2026October 2, 2026

Bars are daily stars, the line is the usual pace. Red marks spike days.

Numbers

Total stars
54,607
Today
6 · ≈ 16 by evening
Forks
6,147
Issues and pull requests
430
Watchers
269
Language
Python
License
MIT
Created
August 25, 2025
Last push
September 3, 2026

Star trust

Growth looks organic: forks and discussion are in line with active projects, and stars arrive unevenly, the way people give them.

These are heuristics, not a verdict: we judge by the repository’s behavior, not by a list of stargazers.

Spotted in

Share

README badge

Paste it into your README — the badge shows the current magnitude and links to this page.

Seismograph: 2.2
[![Seismograph](https://gitnova.dev/badge/microsoft/VibeVoice.svg?lang=en)](https://gitnova.dev/en/r/microsoft/VibeVoice)

Similar by description

  1. 2.5
    OpenWhispr/openwhispr

    Cross-platform desktop voice-to-text dictation app that transcribes speech locally (Whisper, NVIDIA Parakeet) or in the cloud and pastes text into any app. It also transcribes meetings, manages notes, and connects to…

    SteadyProductivity & self-hostedJavaScript+14 stars today, ≈ 28 by evening

  2. 1.5
    jankeesvw/omarchy-meeting-recorder

    Records meetings on Omarchy: mic and system audio as separate tracks, transcribed locally with whisper.cpp, with speaker labels, chapters and a player.

    CoolingProductivity & self-hostedRust+2 stars today, ≈ 3 by evening

  3. 1.5
    k2-fsa/sherpa-onnx

    A collection of offline speech processing tools built on next-gen Kaldi and onnxruntime: speech recognition and synthesis, speaker diarization and identification, VAD, speech enhancement, source separation, and keyword…

    SteadyLanguage modelsC+++3 stars today, ≈ 6 by evening

  4. 1.3
    nobodywho-ooo/nobodywho

    A Rust inference engine for running LLMs locally with bindings for Python, Kotlin, Swift, Flutter, React Native and Godot, plus speech-to-text and text-to-speech. It runs GGUF chat models offline on desktop and mobile…

    CoolingLanguage modelsRust+0 stars today, ≈ 1 by evening

  5. 2.5
    Zackriya-Solutions/meetily

    A local AI meeting assistant that records and transcribes audio in real time using Whisper/Parakeet and generates summaries via Ollama or other LLMs, without sending data to the cloud.

    SteadyAI agentsRust+15 stars today, ≈ 31 by evening

  6. 1.6
    0xShug0/audio.cpp

    A native C++ inference engine for audio models built on ggml: TTS, speech recognition, VAD, voice cloning, music generation. Runs without Python on CPU, CUDA, ROCm, Vulkan and Metal.

    SteadyGenerative mediaC+++6 stars today, ≈ 12 by evening