MiaAI-Lab/Qwen3.8-Flash-Next-Single-DGX-Spark-TensorFold
Qwen3.8 Flash Next on one DGX Spark (TensorFold)
About the project
Scripts to serve Qwen3.8 Flash Next on a single NVIDIA DGX Spark via an OpenAI-compatible API, with image and video input and a 262,144-token context.
Useful for
- Serve Qwen3.8 Flash Next locally on a DGX Spark behind an OpenAI-compatible API
- Send images or videos in chat messages via image_url/video_url content parts
- Point any OpenAI client at base_url http://:8888/v1
README summarized by DeepSeek V4.1 Flash. Details may be inaccurate.
Why it’s trending
- 0 stars so far today, about 61 expected by the end of the day.
- The spike has held for 2 days in a row — not a one-off blip.
- The repository is 1 day old and already has 94 stars. With less than two weeks of history, there's no usual pace to compare the spike against yet.
- Top new repositories this week: #179.
- About 10 forks a day — people are taking the code.
- Recent forks include notable developers: @olexale (972 followers).
Stars per day
Bars are daily stars, the line is the usual pace. Red marks spike days.
Numbers
- Total stars
- 94
- Today
- 0 · ≈ 61 by evening
- Forks
- 6
- Issues and pull requests
- 6
- Watchers
- 1
- Language
- Shell
- Created
- September 29, 2026
- Last push
- September 29, 2026
Star trust
Growth looks organic: forks and discussion are in line with active projects, and stars arrive unevenly, the way people give them.
These are heuristics, not a verdict: we judge by the repository’s behavior, not by a list of stargazers.
Spotted in
- September 30, 2026Top new repositories this week: #179
Similar by description
-
7.2
Niko1221/Strata
A C++ local inference engine that runs the 125B MoE model Qwen3.8-Flash-Next on a regular PC with one NVIDIA GPU (12-24 GB) and 64 GB RAM, exposing an OpenAI/Anthropic-compatible API on localhost.
-
2.0
PrismML-Eng/Bonsai-demo
Demo scripts for running the ternary Bonsai 2 27B model (5.9 GB) locally via a llama.cpp fork or MLX: chat, vision, tool calling and 262K context on a laptop or single GPU.
-
0.7
architectds/collabosm
A set of scripts to run Qwen3.8-Flash-Next (125B MoE) on a single Colab A100-80GB High-RAM with an OpenAI-compatible endpoint and measured throughput.