Seismograph

What’s gaining stars on GitHub right now

ArtificialAnalysis/aa-agentperf-local

Benchmark local LLM serving by replaying real agent trajectories

Language modelsPython#ai-agents#artificial-analysis#benchmark#inference#llm
2.6 Steady Magnitude out of 10 — how fast interest is growing, not a quality score. Star growth looks organic. Data as of October 2, 2026, 01:39 UTC.
Open on Seismograph Open on GitHub

About the project

A tool from Artificial Analysis that measures how fast a local LLM server serves an AI agent by replaying recorded agent conversations and reporting throughput and latency, without judging output quality.

Useful for

README summarized by DeepSeek V4.1 Flash. Details may be inaccurate.

Why it’s trending

Stars per day

01530September 20, 2026October 2, 2026

Bars are daily stars, the line is the usual pace. Red marks spike days.

Numbers

Total stars
70
Today
0 · ≈ 7 by evening
Forks
1
Issues and pull requests
33
Watchers
0
Language
Python
License
Apache-2.0
Latest release
v0.3.3 · October 1, 2026
Created
September 26, 2026
Last push
October 1, 2026

Star trust

Growth looks organic: forks and discussion are in line with active projects, and stars arrive unevenly, the way people give them.

These are heuristics, not a verdict: we judge by the repository’s behavior, not by a list of stargazers.

Hacker News discussions

Spotted in

Share

README badge

Paste it into your README — the badge shows the current magnitude and links to this page.

Seismograph: 2.6
[![Seismograph](https://gitnova.dev/badge/ArtificialAnalysis/aa-agentperf-local.svg?lang=en)](https://gitnova.dev/en/r/ArtificialAnalysis/aa-agentperf-local)

More in this category

  1. 8.6
    Niko1221/Strata

    A C++ local inference engine that runs the 125B MoE model Qwen3.8-Flash-Next on a regular PC with one NVIDIA GPU (12-24 GB) and 64 GB RAM, exposing an OpenAI/Anthropic-compatible API on localhost.

    BreakoutLanguage modelsC+++0 stars today, ≈ 1,175 by evening

  2. 6.4
    MiaAI-Lab/GLM-5.3-Flash-EXL3-2x-DGX-Sparks-TensorFold

    Scripts to serve GLM-5.3-Flash on two NVIDIA DGX Sparks via TensorFold with an OpenAI-compatible API, 1M-token context, and image/video input.

    Early signalLanguage modelsShell+0 stars today, ≈ 116 by evening

  3. 5.8
    ninjahawk/livenerf

    A benchmark for tracking whether a frontier model quietly degrades after release: it runs a frozen question panel daily and statistically measures drift in accuracy and token counts against the launch-week baseline.

    Early signalLanguage modelsPython+0 stars today, ≈ 153 by evening

  4. 5.7
    VectifyAI/PageIndex

    A vectorless RAG engine that builds a hierarchical tree index of a document and retrieves by LLM reasoning over that tree, the way a human navigates a report. Aimed at long professional PDFs such as financial, legal…

    PeakingLanguage modelsPython+0 stars today, ≈ 167 by evening

  5. 5.3
    magnitudedev/magnitude

    Open source inference engine for consumer hardware that profiles your machine, recommends the best local models, then downloads, tunes, and runs them, with one-click connection to AI agents.

    BreakoutLanguage modelsRust+0 stars today, ≈ 299 by evening

  6. 4.8
    ashhart/TensorFold

    LLM inference server for Apple Silicon (MLX) and NVIDIA GPUs with an OpenAI-compatible API and exact speculative decoding. Supports Nemotron, Qwen3.8, GLM-5.3 and Gemma 4 families with per-family kernels and draft…

    Early signalLanguage modelsPython+0 stars today, ≈ 79 by evening