Seismograph

What’s gaining stars on GitHub right now

ikawrakow/ik_llama.cpp

llama.cpp fork with additional SOTA quants and improved performance

Language modelsC++
1.2 Steady Magnitude out of 10 — how fast interest is growing, not a quality score. Star growth looks organic. Data as of September 29, 2026, 12:30 UTC.
Open on Seismograph Open on GitHub

About the project

A llama.cpp fork with additional SOTA quantization types and improved CPU/GPU performance for running LLMs locally. It supports new models and inference techniques that landed here before upstream.

Useful for

README summarized by DeepSeek V4.1 Flash. Details may be inaccurate.

Why it’s trending

Stars per day

01530July 2, 2026September 29, 2026

Bars are daily stars, the line is the usual pace. Red marks spike days.

Numbers

Total stars
3,271
Today
2 · ≈ 4 by evening
Forks
470
Issues and pull requests
2,343
Watchers
35
Language
C++
License
MIT
Latest release
t0002 · July 22, 2025
Created
June 27, 2024
Last push
September 29, 2026

Star trust

Growth looks organic: forks and discussion are in line with active projects, and stars arrive unevenly, the way people give them.

These are heuristics, not a verdict: we judge by the repository’s behavior, not by a list of stargazers.

Spotted in

Share

README badge

Paste it into your README — the badge shows the current magnitude and links to this page.

Seismograph: 1.2
[![Seismograph](https://gitnova.dev/badge/ikawrakow/ik_llama.cpp.svg?lang=en)](https://gitnova.dev/en/r/ikawrakow/ik_llama.cpp)

More in this category

  1. 8.2
    firelex/jeff

    Small fine-tuned Qwen3.5 and Gemma 4 models for zero-shot classification: given a situation description and a list of options, they return a calibrated probability for each option in a single forward pass.

    Early signalLanguage modelsPython+219 stars today, ≈ 466 by evening

  2. 7.4
    VectifyAI/PageIndex

    A vectorless RAG engine that builds a hierarchical tree index of a document and retrieves by LLM reasoning over that tree, the way a human navigates a report. Aimed at long professional PDFs such as financial, legal…

    BreakoutLanguage modelsPython+225 stars today, ≈ 486 by evening

  3. 7.2
    Niko1221/Strata

    A C++ local inference engine that runs the 125B MoE model Qwen3.8-Flash-Next on a regular PC with one NVIDIA GPU (12-24 GB) and 64 GB RAM, exposing an OpenAI/Anthropic-compatible API on localhost.

    Early signalLanguage modelsC+++144 stars today, ≈ 360 by evening

  4. 5.6
    ollaya-dev/ollaya

    A local runtime for decision models: pulls and serves classification and routing models behind a TypeSafe-compatible API. Like Ollama, but for models that return probabilities instead of text.

    Early signalLanguage modelsRust+34 stars today, ≈ 89 by evening

  5. 5.3
    PostHog/jeeves

    Jeeves is a reasoning model built on Qwen3.5-9B (LoRA + pointer head), trained with SFT and CISPO, that thinks before answering and returns calibrated probabilities for yes/no, multiple-choice and rating questions via…

    Early signalLanguage modelsPython+39 stars today, ≈ 58 by evening

  6. 4.9
    deepfates/imp

    A port of DSPy to the BEAM: declarative LM programs in Elixir with signatures, optimizers and agents running as OTP processes.

    Early signalLanguage modelsElixir+1 star today, ≈ 13 by evening