# General-Instinct/InstinctFlash > A high-performance serving runtime for robotics models (VLA, WAM, policies) with FP8 acceleration and optimized kernels on Jetson Thor and RTX 4090/5090. - Magnitude: 2.3 out of 10 — Steady - Stars: 26 total · +13 stars today, ≈ 15 by evening - Star trust: star growth looks organic - Category: Hardware, IoT & robotics · Language: C++ · License: AGPL-3.0 · Created: 2026-05-28 · Last push: 2026-09-18 - GitHub: https://github.com/General-Instinct/InstinctFlash · Homepage: https://general-instinct.com/ · Page: https://gitnova.dev/en/r/General-Instinct/InstinctFlash ## Useful for - Serve a fine-tuned robotics checkpoint for inference via `instinctflash serve` - Expose robot action predictions over the openpi WebSocket protocol to existing clients - Speed up pi0.5 or LingBot-VA inference with FP8 on Jetson Thor ## Why it’s here - 13 stars so far today, about 15 expected by the end of the day. The usual pace is 0 per day, so that's 15× as much. - Hacker News: “Show HN: InstinctFlash – Run 5B world-action models in real time on Jetson Thor” — 16 points, 4 h ago. ## Star trust Star growth looks organic. Star-trust labels are heuristics based on the repository’s behavior, not a check of every stargazer. ## Numbers - Forks: 5 - Issues and pull requests: 11 - Watchers: 0 - Average over the last week: 2 per day - Usual pace: 0 per day - Stars in the last hour (measured): 4 - Latest release: thor-2026-09-15 (2026-09-15) ## Stars per day, last 30 days (oldest → newest, today is partial) 2026-08-24 … 2026-09-22: 0, 0, 0, 0, 0, 0, 0, 0, 0, 0, 0, 0, 0, 0, 0, 0, 0, 0, 0, 0, 0, 1, 0, 0, 1, 1, 0, 0, 0, 13 ## Hacker News - Show HN: InstinctFlash – Run 5B world-action models in real time on Jetson Thor — 16 points, 1 comments: https://news.ycombinator.com/item?id=49802789 ## Spotted in now - Spotted on Hacker News ## Similar by description 1. **joykraft/vla_pi0** — 0.0 · Steady · Hardware, IoT & robotics · Python · +0 stars today, ≈ 0 by evening Project for the Episode1 single-arm robot: teleoperation data collection via LeRobot, LoRA fine-tuning of the π₀ (pi0) VLA model on OpenPI/JAX, and deployment of the policy to a real robot over WebSocket. Full card: https://gitnova.dev/en/r/joykraft/vla_pi0.md 2. **NVIDIA/TensorRT-LLM** — 1.5 · Steady · Language models · Python · +8 stars today, ≈ 10 by evening NVIDIA library for optimizing inference of large language models and visual generative models on GPUs, with a Python API, specialized kernels, and an efficient C++/Python runtime. Full card: https://gitnova.dev/en/r/NVIDIA/TensorRT-LLM.md 3. **microsoft/onnxruntime** — 1.3 · Steady · Language models · C++ · +4 stars today, ≈ 5 by evening Cross-platform accelerator for ML inference and training. Runs models from PyTorch, TensorFlow, scikit-learn and others via the ONNX format with graph optimizations and hardware acceleration. Full card: https://gitnova.dev/en/r/microsoft/onnxruntime.md 4. **Binaire-0101/FRZi-inference** — 0.0 · Steady · Language models · +0 stars today, ≈ 0 by evening FRZi-inference is an inference engine for running machine learning models. It is designed to execute pre-trained models and obtain predictions. Full card: https://gitnova.dev/en/r/Binaire-0101/FRZi-inference.md --- Magnitude (0–10) measures how fast and how unusually interest in a repository is growing right now. It is not a quality score. Days are UTC. “So far today” is a fact; “expected by the end of the day” is a forecast. Summaries and use cases are written by an LLM (DeepSeek V4.1 Flash) from the README and may be inaccurate: verify specific claims (benchmarks, speed, hardware) in the repository itself. Data as of 2026-09-22 19:29 UTC, updated every 30 minutes.