# deepseek-ai/DeepEP-Ascend > A high-performance communication library for ML training and inference on Huawei Ascend NPUs: expert-parallel all-to-all for MoE dispatch/combine, plus PP, CP/DP and Engram. Public buffer APIs are aligned with NVIDIA DeepEP. - Magnitude: 1.1 out of 10 — Quiet - Stars: 118 total · +0 stars today, ≈ 0 by evening - Star trust: star growth looks organic - Category: DevOps & infrastructure · Language: C++ · Created: 2026-09-30 · Last push: 2026-09-30 - GitHub: https://github.com/deepseek-ai/DeepEP-Ascend · Page: https://gitnova.dev/en/r/deepseek-ai/DeepEP-Ascend ## Useful for - Run MoE training with expert-parallel dispatch/combine on an Ascend NPU cluster - Speed up MoE inference with FP8 dispatch and deferred epilogues - Port code from NVIDIA DeepEP to Ascend without changing the buffer API ## Why it’s here - The repository is 0 days old and already has 118 stars. With less than two weeks of history, there's no usual pace to compare the spike against yet. - Top new repositories this week: #118. - About 28 forks a day — people are taking the code. - Recent forks include notable developers: @gmh5225 (1,277 followers). ## Star trust Star growth looks organic. Star-trust labels are heuristics based on the repository’s behavior, not a check of every stargazer. ## Numbers - Forks: 6 - Issues and pull requests: 0 - Watchers: 0 - Average over the last week: 0 per day - Usual pace: too little history (under two weeks) - Stars in the last hour (measured): 16 ## Stars per day, last 4 days (oldest → newest, today is partial) 2026-09-27 … 2026-09-30: 0, 0, 121, 0 ## Spotted in now - Top new repositories this week: #118 ## Similar by description 1. **vllm-project/vllm-ascend** — 1.8 · Steady · Language models · Python · +0 stars today, ≈ 5 by evening A hardware plugin for vLLM that runs LLM inference on Ascend NPUs (Atlas A2/A3). It enables deployment of Transformer, MoE, embedding and multimodal models on Huawei Ascend hardware. Full card: https://gitnova.dev/en/r/vllm-project/vllm-ascend.md 2. **deepseek-ai/DeepGEMM-Ascend** — 5.7 · Early signal · Language models · C++ · +0 stars today, ≈ 94 by evening A port of DeepGEMM to Huawei Ascend: GEMM kernels (BF16, FP8, FP4, MQA logits, MegaMoE) with an API compatible with DeepGEMM, targeting peak NPU performance. Full card: https://gitnova.dev/en/r/deepseek-ai/DeepGEMM-Ascend.md 3. **microsoft/onnxruntime** — 1.7 · Steady · Language models · C++ · +0 stars today, ≈ 7 by evening Cross-platform accelerator for ML inference and training. Runs models from PyTorch, TensorFlow, scikit-learn and others via the ONNX format with graph optimizations and hardware acceleration. Full card: https://gitnova.dev/en/r/microsoft/onnxruntime.md --- Magnitude (0–10) measures how fast and how unusually interest in a repository is growing right now. It is not a quality score. Days are UTC. “So far today” is a fact; “expected by the end of the day” is a forecast. Summaries and use cases are written by an LLM (DeepSeek V4.1 Flash) from the README and may be inaccurate: verify specific claims (benchmarks, speed, hardware) in the repository itself. Data as of 2026-09-30 06:54 UTC, updated every 30 minutes.