# jayleaton/glm53-tensorfold-spark > Experimental setup for serving the 4-bit GLM-5.3-Flash checkpoint across two NVIDIA DGX Sparks via the TensorFold engine, with an OpenAI-compatible API and a set of patches for faster decoding. - Magnitude: 3.1 out of 10 — Early signal - Stars: 87 total · +0 stars today, ≈ 20 by evening - Star trust: star growth looks organic - Category: Language models · Language: Python · License: Apache-2.0 · Created: 2026-09-28 · Last push: 2026-10-01 - GitHub: https://github.com/jayleaton/glm53-tensorfold-spark · Page: https://gitnova.dev/en/r/jayleaton/glm53-tensorfold-spark ## Useful for - Stand up an OpenAI-compatible GLM-5.3-Flash server on a pair of DGX Sparks - Benchmark TensorFold vs vLLM decode speed on your own prompts - Enable tool-calling and structured-output patches via config/prod.env ## Why it’s here - 0 stars so far today, about 20 expected by the end of the day. - The spike has held for 3 days in a row — not a one-off blip. - The repository is 3 days old and already has 87 stars. With less than two weeks of history, there's no usual pace to compare the spike against yet. - Top new repositories this week: #155. ## Star trust Star growth looks organic. Star-trust labels are heuristics based on the repository’s behavior, not a check of every stargazer. ## Numbers - Forks: 5 - Issues and pull requests: 14 - Watchers: 1 - Average over the last week: 27 per day - Usual pace: too little history (under two weeks) - Stars in the last hour (measured): 1 ## Stars per day, last 5 days (oldest → newest, today is partial) 2026-09-27 … 2026-10-01: 0, 10, 49, 28, 0 ## Spotted in now - Top new repositories this week: #155 ## Similar by description 1. **ashhart/TensorFold** — 5.3 · Early signal · Language models · Python · +0 stars today, ≈ 48 by evening LLM inference server for Apple Silicon (MLX) and NVIDIA GPUs with an OpenAI-compatible API and exact speculative decoding. Supports Nemotron, Qwen3.8, GLM-5.3 and Gemma 4 families with per-family kernels and draft models. Full card: https://gitnova.dev/en/r/ashhart/TensorFold.md 2. **MiaAI-Lab/Qwen3.8-Flash-Next-Single-DGX-Spark-TensorFold** — 3.8 · Early signal · Language models · Shell · +0 stars today, ≈ 17 by evening Scripts to serve Qwen3.8 Flash Next on a single NVIDIA DGX Spark via an OpenAI-compatible API, with image and video input and a 262,144-token context. Full card: https://gitnova.dev/en/r/MiaAI-Lab/Qwen3.8-Flash-Next-Single-DGX-Spark-TensorFold.md --- Magnitude (0–10) measures how fast and how unusually interest in a repository is growing right now. It is not a quality score. Days are UTC. “So far today” is a fact; “expected by the end of the day” is a forecast. Summaries and use cases are written by an LLM (DeepSeek V4.1 Flash) from the README and may be inaccurate: verify specific claims (benchmarks, speed, hardware) in the repository itself. Data as of 2026-10-01 03:34 UTC, updated every 30 minutes.