# raullenchai/Rapid-MLX > An open-source (Apache 2.0) LLM inference server for Apple Silicon built on MLX, exposing OpenAI- and Anthropic-compatible APIs and focused on reliable tool calling for coding agents. Claimed up to 4× faster than mlx-lm on the same weights. - Magnitude: 1.5 out of 10 — Steady - Stars: 3,913 total · +1 star measured 2026-10-06, 11:46–14:22 UTC, ≈ 4 by evening - Star trust: star growth looks organic - Category: Language models · Language: Python · Created: 2026-02-25 · Last push: 2026-10-06 - GitHub: https://github.com/raullenchai/Rapid-MLX · Homepage: https://rapidmlx.com · Page: https://gitnova.dev/en/r/raullenchai/Rapid-MLX ## Useful for - Point Claude Code or another agent at the local Anthropic-compatible endpoint - Run a local OpenAI-compatible server on a Mac for chat and tool calling without the cloud - Benchmark decode speed against mlx-lm or Ollama on your own M-series Mac ## Why it’s here - Star-counter measurements on 2026-10-06 (UTC), 11:46–14:22: 3912 → 3913 stars (+1). This is the change over that interval. - Estimated end-of-day forecast: about +4 stars, using observed gains and the previous day. - GitHub Trending Python today: #8, +16 stars. ## Star trust Star growth looks organic. Star-trust labels are heuristics based on the repository’s behavior, not a check of every stargazer. ## Numbers - Forks: 429 - Issues and pull requests: 4,154 - Watchers: 65 - Average over the last week: 8 per day - Usual pace: 8 per day - Stars in the last hour (measured): 0 - Latest release: v0.15.6 (2026-10-04) ## Stars per day, last 30 days (oldest → newest, today is partial) 2026-09-07 … 2026-10-06: 17, 12, 11, 8, 5, 6, 6, 6, 10, 8, 8, 17, 7, 5, 9, 12, 7, 10, 7, 5, 8, 4, 6, 2, 7, 15, 10, 4, 15, 1 ## Spotted in now - GitHub Trending Python today: #8, +16 stars ## Similar by description 1. **incoai/splash** — 2.0 · Steady · Language models · C++ · +15 stars measured 2026-10-06, 00:22–14:18 UTC, ≈ 26 by evening A local LLM inference engine for Apple silicon, specialized per model: serves OpenAI- and Anthropic-compatible APIs to coding agents on a single Mac. Full card: https://gitnova.dev/en/r/incoai/splash.md --- Magnitude (0–10) measures how fast and how unusually interest in a repository is growing right now. It is not a quality score. Days are UTC. “So far today” is a fact; “expected by the end of the day” is a forecast. Summaries and use cases are written by an LLM (DeepSeek V4.1 Flash) from the README and may be inaccurate: verify specific claims (benchmarks, speed, hardware) in the repository itself. Data as of 2026-10-06 14:31 UTC, updated every 30 minutes.