Seismograph

What’s gaining stars on GitHub right now

vllm-project/semantic-router

A programmable Mixture-of-Models router for heterogeneous LLM inference

Language modelsGo#mixture-of-models#semantic-router#vllm#ai-gateway#kubernetes
2.7 Steady Magnitude out of 10 — how fast interest is growing, not a quality score. Star growth looks organic. Data as of September 29, 2026, 12:30 UTC.
Open on Seismograph Open on GitHub

About the project

A programmable routing layer for Mixture-of-Models systems across heterogeneous LLM infrastructure: it selects or composes the right model path per request based on request signals, user preferences, and application policies. It helps manage quality, cost, latency, privacy, and safety without hard-coding routing logic into applications.

Useful for

README summarized by DeepSeek V4.1 Flash. Details may be inaccurate.

Why it’s trending

Stars per day

02040July 2, 2026September 29, 2026

Bars are daily stars, the line is the usual pace. Red marks spike days.

Numbers

Total stars
5,978
Today
4 · ≈ 13 by evening
Forks
989
Issues and pull requests
4,350
Watchers
63
Language
Go
License
Apache-2.0
Latest release
v0.4.0 · September 27, 2026
Created
August 26, 2025
Last push
September 29, 2026

Star trust

Growth looks organic: forks and discussion are in line with active projects, and stars arrive unevenly, the way people give them.

These are heuristics, not a verdict: we judge by the repository’s behavior, not by a list of stargazers.

Spotted in

Share

README badge

Paste it into your README — the badge shows the current magnitude and links to this page.

Seismograph: 2.7
[![Seismograph](https://gitnova.dev/badge/vllm-project/semantic-router.svg?lang=en)](https://gitnova.dev/en/r/vllm-project/semantic-router)

More in this category

  1. 8.2
    firelex/jeff

    Small fine-tuned Qwen3.5 and Gemma 4 models for zero-shot classification: given a situation description and a list of options, they return a calibrated probability for each option in a single forward pass.

    Early signalLanguage modelsPython+219 stars today, ≈ 466 by evening

  2. 7.4
    VectifyAI/PageIndex

    A vectorless RAG engine that builds a hierarchical tree index of a document and retrieves by LLM reasoning over that tree, the way a human navigates a report. Aimed at long professional PDFs such as financial, legal…

    BreakoutLanguage modelsPython+225 stars today, ≈ 486 by evening

  3. 7.2
    Niko1221/Strata

    A C++ local inference engine that runs the 125B MoE model Qwen3.8-Flash-Next on a regular PC with one NVIDIA GPU (12-24 GB) and 64 GB RAM, exposing an OpenAI/Anthropic-compatible API on localhost.

    Early signalLanguage modelsC+++144 stars today, ≈ 360 by evening

  4. 5.6
    ollaya-dev/ollaya

    A local runtime for decision models: pulls and serves classification and routing models behind a TypeSafe-compatible API. Like Ollama, but for models that return probabilities instead of text.

    Early signalLanguage modelsRust+34 stars today, ≈ 89 by evening

  5. 5.3
    PostHog/jeeves

    Jeeves is a reasoning model built on Qwen3.5-9B (LoRA + pointer head), trained with SFT and CISPO, that thinks before answering and returns calibrated probabilities for yes/no, multiple-choice and rating questions via…

    Early signalLanguage modelsPython+39 stars today, ≈ 58 by evening

  6. 4.9
    deepfates/imp

    A port of DSPy to the BEAM: declarative LM programs in Elixir with signatures, optimizers and agents running as OTP processes.

    Early signalLanguage modelsElixir+1 star today, ≈ 13 by evening