Seismograph

What’s gaining stars on GitHub right now

Inferact/tpu-megakernels

A collection of megakernels for TPU

Language modelsPython
3.3 Early signal Magnitude out of 10 — how fast interest is growing, not a quality score. Star growth looks organic. Data as of September 25, 2026, 01:29 UTC.
Open on Seismograph Open on GitHub

About the project

A collection of fused megakernels for LLM inference on TPU, targeting Kimi K3 and Qwen3.8-27B with speculative decoding, reaching up to ~2x the decode throughput of a GB200 baseline at batch sizes 1 to 8.

Useful for

README summarized by DeepSeek V4.1 Flash. Details may be inaccurate.

Why it’s trending

Stars per day

04080September 20, 2026September 25, 2026

Bars are daily stars, the line is the usual pace. Red marks spike days.

Numbers

Total stars
95
Today
0 · ≈ 24 by evening
Forks
2
Issues and pull requests
1
Watchers
1
Language
Python
License
Apache-2.0
Created
September 23, 2026
Last push
September 25, 2026

Star trust

Growth looks organic: forks and discussion are in line with active projects, and stars arrive unevenly, the way people give them.

These are heuristics, not a verdict: we judge by the repository’s behavior, not by a list of stargazers.

Spotted in

Share

README badge

Paste it into your README — the badge shows the current magnitude and links to this page.

Seismograph: 3.3
[![Seismograph](https://gitnova.dev/badge/Inferact/tpu-megakernels.svg?lang=en)](https://gitnova.dev/en/r/Inferact/tpu-megakernels)

Similar by description

  1. 1.1
    MiaAI-Lab/DeepSeek-v4.1-Flash-EXL3-2x-DGX-Sparks

    A local EXL3 2.9 bpw checkpoint of DeepSeek-V4.1-Flash (196 GiB, 39 shards) served via an OpenAI-compatible vLLM API on a 2x NVIDIA GB10 (DGX Spark) kit with tensor-parallel 2. Includes DSpark speculative decoding and…

    CoolingLanguage modelsPython+0 stars today, ≈ 4 by evening

  2. 2.1
    MiaAI-Lab/DeepSeek-v4.1-Flash-DGX-Sparks

    Scripts and patches to serve DeepSeek-V4.1-Flash with SGLang across a 3–4 node NVIDIA DGX Spark cluster, using MXFP4/FP8, speculative decoding and an OpenAI-compatible endpoint.

    SteadyLanguage modelsPython+0 stars today, ≈ 8 by evening