# ArtificialAnalysis/aa-agentperf-local > A tool from Artificial Analysis that measures how fast a local LLM server serves an AI agent by replaying recorded agent conversations and reporting throughput and latency, without judging output quality. - Magnitude: 2.8 out of 10 — Early signal - Stars: 72 total · +0 stars today, ≈ 8 by evening - Star trust: star growth looks organic - Category: Language models · Language: Python · License: Apache-2.0 · Created: 2026-09-26 · Last push: 2026-10-01 - GitHub: https://github.com/ArtificialAnalysis/aa-agentperf-local · Homepage: https://artificialanalysis.ai · Page: https://gitnova.dev/en/r/ArtificialAnalysis/aa-agentperf-local ## Useful for - Benchmark throughput and latency of your own llama.cpp, vLLM, SGLang or LM Studio server on recorded agent trajectories - Compare speed across models and hardware via managed-run with pinned recipes and SHA-256 verification - Check server setup with the quick synthetic aa-mini-v1 replay at 8192 context ## Why it’s here - 0 stars so far today, about 8 expected by the end of the day. - The spike has held for 3 days in a row — not a one-off blip. - The repository is 6 days old and already has 72 stars. With less than two weeks of history, there's no usual pace to compare the spike against yet. - Top new repositories this week: #159. - Hacker News: “Gemini 4 Argon (High): Intelligence, Performance and Price Analysis” — 111 points, 1 day ago. ## Star trust Star growth looks organic. Star-trust labels are heuristics based on the repository’s behavior, not a check of every stargazer. ## Numbers - Forks: 1 - Issues and pull requests: 33 - Watchers: 0 - Average over the last week: 11 per day - Usual pace: too little history (under two weeks) - Stars in the last hour (measured): 1 - Latest release: v0.3.3 (2026-10-01) ## Stars per day, last 13 days (oldest → newest, today is partial) 2026-09-20 … 2026-10-02: 0, 0, 0, 0, 0, 0, 0, 1, 20, 11, 30, 10, 0 ## Hacker News - Gemini 4 Argon (High): Intelligence, Performance and Price Analysis — 111 points, 61 comments: https://news.ycombinator.com/item?id=49914236 - GPT-6.1 Sol replaces GPT-6 Sol after just 7 days, with near-Astra intelligence — 80 points, 101 comments: https://news.ycombinator.com/item?id=49906669 - Sonnet 5.5 scores just behind Opus 5.5 on Artificial Analysis Intelligence Index — 11 points, 6 comments: https://news.ycombinator.com/item?id=49885725 - Gemini 4 Argon: Google is back among the top three labs in intelligence achieved — 6 points, 0 comments: https://news.ycombinator.com/item?id=49916067 ## Spotted in now - Top new repositories this week: #159 ## More in this category 1. **Niko1221/Strata** — 8.5 · Breakout · Language models · C++ · +0 stars today, ≈ 1,074 by evening A C++ local inference engine that runs the 125B MoE model Qwen3.8-Flash-Next on a regular PC with one NVIDIA GPU (12-24 GB) and 64 GB RAM, exposing an OpenAI/Anthropic-compatible API on localhost. Full card: https://gitnova.dev/en/r/Niko1221/Strata.md 2. **MiaAI-Lab/GLM-5.3-Flash-EXL3-2x-DGX-Sparks-TensorFold** — 6.4 · Early signal · Language models · Shell · +0 stars today, ≈ 108 by evening Scripts to serve GLM-5.3-Flash on two NVIDIA DGX Sparks via TensorFold with an OpenAI-compatible API, 1M-token context, and image/video input. Full card: https://gitnova.dev/en/r/MiaAI-Lab/GLM-5.3-Flash-EXL3-2x-DGX-Sparks-TensorFold.md 3. **ninjahawk/livenerf** — 5.8 · Early signal · Language models · Python · +0 stars today, ≈ 140 by evening A benchmark for tracking whether a frontier model quietly degrades after release: it runs a frozen question panel daily and statistically measures drift in accuracy and token counts against the launch-week baseline. Full card: https://gitnova.dev/en/r/ninjahawk/livenerf.md 4. **VectifyAI/PageIndex** — 5.7 · Peaking · Language models · Python · +0 stars today, ≈ 153 by evening A vectorless RAG engine that builds a hierarchical tree index of a document and retrieves by LLM reasoning over that tree, the way a human navigates a report. Aimed at long professional PDFs such as financial, legal and technical documents. Full card: https://gitnova.dev/en/r/VectifyAI/PageIndex.md 5. **magnitudedev/magnitude** — 5.2 · Breakout · Language models · Rust · +0 stars today, ≈ 273 by evening Open source inference engine for consumer hardware that profiles your machine, recommends the best local models, then downloads, tunes, and runs them, with one-click connection to AI agents. Full card: https://gitnova.dev/en/r/magnitudedev/magnitude.md --- Magnitude (0–10) measures how fast and how unusually interest in a repository is growing right now. It is not a quality score. Days are UTC. “So far today” is a fact; “expected by the end of the day” is a forecast. Summaries and use cases are written by an LLM (DeepSeek V4.1 Flash) from the README and may be inaccurate: verify specific claims (benchmarks, speed, hardware) in the repository itself. Data as of 2026-10-02 03:05 UTC, updated every 30 minutes.