Seismograph

What’s gaining stars on GitHub right now

jayleaton/glm53-tensorfold-spark

GLM-5.3-Flash (abliterated EXL3) on 2x NVIDIA DGX Spark with the TensorFold engine: 1.8x faster decode than vLLM, 4x256k concurrent threads, byte-exact speculative decoding. Work in progress.

Language modelsPython
3.1 Early signal Magnitude out of 10 — how fast interest is growing, not a quality score. Star growth looks organic. Data as of October 1, 2026, 02:51 UTC.
Open on Seismograph Open on GitHub

About the project

Experimental setup for serving the 4-bit GLM-5.3-Flash checkpoint across two NVIDIA DGX Sparks via the TensorFold engine, with an OpenAI-compatible API and a set of patches for faster decoding.

Useful for

README summarized by DeepSeek V4.1 Flash. Details may be inaccurate.

Why it’s trending

Stars per day

02550September 27, 2026October 1, 2026

Bars are daily stars, the line is the usual pace. Red marks spike days.

Numbers

Total stars
87
Today
0 · ≈ 22 by evening
Forks
5
Issues and pull requests
14
Watchers
1
Language
Python
License
Apache-2.0
Created
September 28, 2026
Last push
October 1, 2026

Star trust

Growth looks organic: forks and discussion are in line with active projects, and stars arrive unevenly, the way people give them.

These are heuristics, not a verdict: we judge by the repository’s behavior, not by a list of stargazers.

Spotted in

Share

README badge

Paste it into your README — the badge shows the current magnitude and links to this page.

Seismograph: 3.1
[![Seismograph](https://gitnova.dev/badge/jayleaton/glm53-tensorfold-spark.svg?lang=en)](https://gitnova.dev/en/r/jayleaton/glm53-tensorfold-spark)

Similar by description

  1. 5.3
    ashhart/TensorFold

    LLM inference server for Apple Silicon (MLX) and NVIDIA GPUs with an OpenAI-compatible API and exact speculative decoding. Supports Nemotron, Qwen3.8, GLM-5.3 and Gemma 4 families with per-family kernels and draft…

    Early signalLanguage modelsPython+0 stars today, ≈ 48 by evening

  2. 3.8
    MiaAI-Lab/Qwen3.8-Flash-Next-Single-DGX-Spark-TensorFold

    Scripts to serve Qwen3.8 Flash Next on a single NVIDIA DGX Spark via an OpenAI-compatible API, with image and video input and a 262,144-token context.

    Early signalLanguage modelsShell+0 stars today, ≈ 19 by evening