# MiaAI-Lab/GLM-5.3-Flash-EXL3-2x-DGX-Sparks-TensorFold > Scripts to serve GLM-5.3-Flash on two NVIDIA DGX Sparks via TensorFold with an OpenAI-compatible API, 1M-token context, and image/video input. - Magnitude: 5.5 out of 10 — Early signal - Stars: 91 total · +98 stars today, ≈ 131 by evening - Star trust: star growth looks organic - Category: Language models · Language: Shell · License: Apache-2.0 · Created: 2026-09-30 · Last push: 2026-10-01 - GitHub: https://github.com/MiaAI-Lab/GLM-5.3-Flash-EXL3-2x-DGX-Sparks-TensorFold · Homepage: https://x.com/MiaAI_lab · Page: https://gitnova.dev/en/r/MiaAI-Lab/GLM-5.3-Flash-EXL3-2x-DGX-Sparks-TensorFold ## Useful for - Deploy a local OpenAI-compatible GLM-5.3-Flash endpoint on two DGX Sparks - Serve requests with up to 1M-token context and image input via the API - Benchmark prefill and decode throughput on your own DGX Spark pair ## Why it’s here - 98 stars so far today, about 131 expected by the end of the day. - The repository is 1 day old and already has 91 stars. With less than two weeks of history, there's no usual pace to compare the spike against yet. - Top new repositories this week: #161. - About 71 forks a day — people are taking the code. - Recent forks include notable developers: @d3vilbug (302 followers). ## Star trust Star growth looks organic. Star-trust labels are heuristics based on the repository’s behavior, not a check of every stargazer. ## Numbers - Forks: 9 - Issues and pull requests: 10 - Watchers: 3 - Average over the last week: 66 per day - Usual pace: too little history (under two weeks) - Stars in the last hour (measured): 12 ## Stars per day, last 5 days (oldest → newest, today is partial) 2026-09-27 … 2026-10-01: 0, 0, 0, 0, 98 ## Spotted in now - Top new repositories this week: #161 ## Similar by description 1. **jayleaton/glm53-tensorfold-spark** — 2.9 · Early signal · Language models · Python · +5 stars today, ≈ 10 by evening Experimental setup for serving the 4-bit GLM-5.3-Flash checkpoint across two NVIDIA DGX Sparks via the TensorFold engine, with an OpenAI-compatible API and a set of patches for faster decoding. Full card: https://gitnova.dev/en/r/jayleaton/glm53-tensorfold-spark.md 2. **MiaAI-Lab/Qwen3.8-Flash-Next-Single-DGX-Spark-TensorFold** — 3.6 · Early signal · Language models · Shell · +6 stars today, ≈ 11 by evening Scripts to serve Qwen3.8 Flash Next on a single NVIDIA DGX Spark via an OpenAI-compatible API, with image and video input and a 262,144-token context. Full card: https://gitnova.dev/en/r/MiaAI-Lab/Qwen3.8-Flash-Next-Single-DGX-Spark-TensorFold.md 3. **architectds/collabosm** — 0.6 · Steady · Language models · Python · +1 star today, ≈ 2 by evening A set of scripts to run Qwen3.8-Flash-Next (125B MoE) on a single Colab A100-80GB High-RAM with an OpenAI-compatible endpoint and measured throughput. Full card: https://gitnova.dev/en/r/architectds/collabosm.md 4. **mmastrac/djev-spark** — 0.0 · Steady · Language models · HTML · +0 stars today, ≈ 0 by evening Container recipe for running DiffusionGemma 26B-A4B (NVFP4) on a DGX Spark with vLLM and a structured-decision server exposing Jev's /v1/systemone API. Full card: https://gitnova.dev/en/r/mmastrac/djev-spark.md --- Magnitude (0–10) measures how fast and how unusually interest in a repository is growing right now. It is not a quality score. Days are UTC. “So far today” is a fact; “expected by the end of the day” is a forecast. Summaries and use cases are written by an LLM (DeepSeek V4.1 Flash) from the README and may be inaccurate: verify specific claims (benchmarks, speed, hardware) in the repository itself. Data as of 2026-10-01 15:54 UTC, updated every 30 minutes.