Seismograph

What’s gaining stars on GitHub right now

SemiAnalysisAI/InferenceX

Open Source Inference Research Platform Standard / 开源推理研究平台

Language modelsPython#pytorch#sglang#vllm#amd#cuda
1.4 Steady Magnitude out of 10 — how fast interest is growing, not a quality score. Star growth looks organic. Data as of September 29, 2026, 12:30 UTC.
Open on Seismograph Open on GitHub

About the project

Open platform for continuous LLM inference benchmarking: measures SGLang, vLLM, TensorRT-LLM performance on NVIDIA, AMD and TPU hardware and publishes live results on a dashboard.

Useful for

README summarized by DeepSeek V4.1 Flash. Details may be inaccurate.

Why it’s trending

Stars per day

03060July 2, 2026September 29, 2026

Bars are daily stars, the line is the usual pace. Red marks spike days.

Numbers

Total stars
1,783
Today
3 · ≈ 7 by evening
Forks
309
Issues and pull requests
3,549
Watchers
14
Language
Python
License
Apache-2.0
Latest release
tilert-v0.1.5.post2-inferencex.1 · August 19, 2026
Created
July 12, 2025
Last push
September 29, 2026

Star trust

Growth looks organic: forks and discussion are in line with active projects, and stars arrive unevenly, the way people give them.

These are heuristics, not a verdict: we judge by the repository’s behavior, not by a list of stargazers.

Spotted in

Share

README badge

Paste it into your README — the badge shows the current magnitude and links to this page.

Seismograph: 1.4
[![Seismograph](https://gitnova.dev/badge/SemiAnalysisAI/InferenceX.svg?lang=en)](https://gitnova.dev/en/r/SemiAnalysisAI/InferenceX)

More in this category

  1. 8.2
    firelex/jeff

    Small fine-tuned Qwen3.5 and Gemma 4 models for zero-shot classification: given a situation description and a list of options, they return a calibrated probability for each option in a single forward pass.

    Early signalLanguage modelsPython+219 stars today, ≈ 466 by evening

  2. 7.4
    VectifyAI/PageIndex

    A vectorless RAG engine that builds a hierarchical tree index of a document and retrieves by LLM reasoning over that tree, the way a human navigates a report. Aimed at long professional PDFs such as financial, legal…

    BreakoutLanguage modelsPython+225 stars today, ≈ 486 by evening

  3. 7.2
    Niko1221/Strata

    A C++ local inference engine that runs the 125B MoE model Qwen3.8-Flash-Next on a regular PC with one NVIDIA GPU (12-24 GB) and 64 GB RAM, exposing an OpenAI/Anthropic-compatible API on localhost.

    Early signalLanguage modelsC+++144 stars today, ≈ 360 by evening

  4. 5.6
    ollaya-dev/ollaya

    A local runtime for decision models: pulls and serves classification and routing models behind a TypeSafe-compatible API. Like Ollama, but for models that return probabilities instead of text.

    Early signalLanguage modelsRust+34 stars today, ≈ 89 by evening

  5. 5.3
    PostHog/jeeves

    Jeeves is a reasoning model built on Qwen3.5-9B (LoRA + pointer head), trained with SFT and CISPO, that thinks before answering and returns calibrated probabilities for yes/no, multiple-choice and rating questions via…

    Early signalLanguage modelsPython+39 stars today, ≈ 58 by evening

  6. 4.9
    deepfates/imp

    A port of DSPy to the BEAM: declarative LM programs in Elixir with signatures, optimizers and agents running as OTP processes.

    Early signalLanguage modelsElixir+1 star today, ≈ 13 by evening