Seismograph

What’s gaining stars on GitHub right now

NVIDIA/TensorRT-LLM

TensorRT LLM provides users with an easy-to-use Python API to define Large Language Models (LLMs) and supports state-of-the-art optimizations to perform inference efficiently on NVIDIA GPUs. TensorRT LLM also contains components to create Python and C++ runtimes that orchestrate the inference execution in a performant way.

Language modelsPython#blackwell#cuda#moe#pytorch#llm-serving
2.0 Steady Magnitude out of 10. Star growth looks organic. Data as of September 19, 2026.
Open on Seismograph Open on GitHub

About the project

NVIDIA library for optimizing inference of large language models and visual generative models on GPUs, with a Python API, specialized kernels, and an efficient C++/Python runtime.

Useful for

README summarized by DeepSeek V4.1 Flash. Details may be inaccurate.

Why it’s trending

Stars per day

01225June 22, 2026September 19, 2026

Bars are daily stars, the line is the usual pace. Red marks spike days.

Numbers

Total stars
14,663
Stars in a day
6
Forks
2,761
Issues and pull requests
19,205
Watchers
123
Language
Python
Latest release
v1.3.0rc27 · September 18, 2026
Created
August 16, 2023
Last push
September 19, 2026

Star trust

Growth looks organic: forks and discussion are in line with active projects, and stars arrive unevenly, the way people give them.

These are heuristics, not a verdict: we judge by the repository’s behavior, not by a list of stargazers.

Spotted in

Share

README badge

Paste it into your README — the badge shows the current magnitude and links to this page.

Seismograph: 2.0
[![Seismograph](https://gitnova.dev/badge/NVIDIA/TensorRT-LLM.svg?lang=en)](https://gitnova.dev/en/r/NVIDIA/TensorRT-LLM)

Similar projects

  1. 7.7
    TheoLeeCJ/SemIf

    An open reproduction of the Jev-style semantic decision interface: reads typed option probabilities directly from a 4B model's logits without generating text. Runs on a single RTX 3090.

    BreakoutLanguage modelsPython+245 stars in a day

  2. 7.4
    TianyuCodings/NanoJev

    A nano replica of Jev built on Qwen3-0.6B that outputs probability distributions over candidates in a single forward pass with no token decoding, plus a training pipeline and game demos (maze, Snake).

    Early signalLanguage modelsPython+336 stars in a day

  3. 6.4
    Continuum-AI-Corp/OrcaBonsai-27B-Uncensored

    A tool for runtime refusal ablation in the compressed Ternary Bonsai 2 27B LLM: it projects the residual stream to remove the refusal direction without modifying or re-quantizing weights. Runs on Apple Silicon via MLX.

    Early signalLanguage modelsPython+142 stars in a day

  4. 6.1
    jaredpalmer/kev

    kev is a LoRA adapter with a small readout head on top of Qwen2.5-0.5B that answers many typed questions about a document in a single forward pass, returning calibrated probabilities instead of text.

    Early signalLanguage modelsPython+147 stars in a day

  5. 6.0
    githubnext/localjev

    A local TypeScript/Bun HTTP service implementing a Jev-compatible POST /v1/systemone API on top of an OpenAI-compatible Chat Completions endpoint (DiffusionGemma via oMLX). It acts as a bridge for typed decisions…

    Early signalLanguage modelsTypeScript+253 stars in a day

  6. 6.0
    higgsfield-ai/higgsfield

    Fault-tolerant GPU orchestration and ML framework for distributed training of models with billions to trillions of parameters, including LLMs. Manages node access, experiment queueing, and GitHub Actions integration.

    BreakoutLanguage modelsJupyter Notebook+113 stars in a day