Seismograph

What’s gaining stars on GitHub right now

vllm-project/vllm-ascend

Community maintained hardware plugin for vLLM on Ascend

Language modelsC++#ascend#inference#llm#llm-serving#llmops
1.2 Steady Magnitude out of 10. Star growth looks organic. Data as of September 14, 2026.
Open on Seismograph Open on GitHub

About the project

A hardware plugin for vLLM that runs LLM inference on Ascend NPUs (Atlas A2/A3). It enables deployment of Transformer, MoE, embedding and multimodal models on Huawei Ascend hardware.

Useful for

README summarized by DeepSeek V4.1 Flash. Details may be inaccurate.

Why it’s trending

Stars per day

0815June 17, 2026September 14, 2026

Bars are daily stars, the line is the usual pace. Red marks spike days.

Numbers

Total stars
2,811
Stars in a day
2
Forks
2,238
Issues and pull requests
16,450
Watchers
74
Language
C++
License
Apache-2.0
Latest release
v0.26.0rc1 · September 3, 2026
Created
January 29, 2025
Last push
September 13, 2026

Star trust

Growth looks organic: forks and discussion are in line with active projects, and stars arrive unevenly, the way people give them.

These are heuristics, not a verdict: we judge by the repository’s behavior, not by a list of stargazers.

Spotted in

Share

README badge

Paste it into your README — the badge shows the current magnitude and links to this page.

Seismograph: 1.2
[![Seismograph](https://gitnova.dev/badge/vllm-project/vllm-ascend.svg?lang=en)](https://gitnova.dev/en/r/vllm-project/vllm-ascend)

Similar projects

  1. 6.1
    JustVugg/colibri

    A pure-C, zero-dependency inference engine for running large MoE models (744B–2.8T parameters) on consumer hardware by treating VRAM, RAM, and storage as a single multitier hierarchy and streaming experts from disk.

    BreakoutLanguage modelsC+1,048 stars in a day

  2. 5.7
    asgeirtj/system_prompts_leaks

    A collection of extracted system prompts from Anthropic, OpenAI, Google, xAI and others — the hidden instructions chatbots receive before a user's first message.

    BreakoutLanguage modelsJavaScript+428 stars in a day

  3. 5.1
    kennethwolters/litelm

    A lightweight litellm alternative: routes LLM calls across 19 providers and translates message formats in ~2,900 lines with two dependencies, without proxy, caching, or cost tracking.

    Early signalLanguage modelsPython+9 stars in a day

  4. 5.0
    MiaAI-Lab/DeepSeek-v4.1-Flash-EXL3-2x-DGX-Sparks

    A local EXL3 2.9 bpw checkpoint of DeepSeek-V4.1-Flash (196 GiB, 39 shards) served via an OpenAI-compatible vLLM API on a 2x NVIDIA GB10 (DGX Spark) kit with tensor-parallel 2. Includes DSpark speculative decoding and…

    Early signalLanguage modelsPython+81 stars in a day

  5. 4.8
    unclecode/crawl4ai

    Open-source web crawler and scraper that turns pages into clean, LLM-ready Markdown for RAG, agents and data pipelines. Runs via Python API, CLI and Docker with no API keys.

    BreakoutLanguage modelsPython+411 stars in a day

  6. 4.5
    Edge0-AI/Edge0

    An open-source streaming MoE inference framework: expert weights are offloaded from SSD on demand while a trained prerouter predicts routing ahead of time. Runs on Apple Silicon via MLX and ships with two ready-to-run…

    CoolingLanguage modelsPython+118 stars in a day