Seismograph

What’s gaining stars on GitHub right now

SeerRay-Lab/Xiaomi-OCR-0

Language modelsPython
More actions
My radar Open on Seismograph

About the project

A unified 0.8B vision-language model for document parsing and OCR-centric understanding: outputs structured Markdown, extracts fields as JSON, and answers questions about document images.

Useful for
  • Parse PDF pages or scans into structured Markdown with tables and formulas
  • Extract requested fields from documents as JSON via KIE
  • Ask questions about document image content via VQA

README summarized by DeepSeek V4.1 Flash. Details may be inaccurate.

0.9 Steady Magnitude out of 10 — how fast interest is growing, not a quality score.

Star growth looks organic. Data as of October 3, 2026, 01:56 UTC.

The star-growth assessment does not verify whether the project is safe to run.

Stars per day

02040September 27, 2026October 3, 2026

Bars show daily stars; the line is a moving average of the available history. Short histories do not yet establish a reliable usual pace. Red marks spike days.

Why it’s trending

  • 0 stars so far today, about 4 expected by the end of the day.
  • The repository is 5 days old and already has 56 stars. With less than two weeks of history, there's no usual pace to compare the spike against yet.
  • Top new repositories this week: #191.

Explore related projects

  1. PaddlePaddle/PaddleOCR

    OCR toolkit and document AI engine that turns PDFs and images into structured, LLM-ready data (JSON/Markdown), with support for 100+ languages, tables, formulas and seals.

    Similar descriptions and use cases
  2. run-llama/liteparse

    A local Rust document parser that extracts text with bounding boxes from PDF, DOCX, XLSX, PPTX and images, with OCR via Tesseract or HTTP servers. Runs without cloud or LLMs, with bindings for Python, Node.js, WASM and…

    Similar descriptions and use cases
  3. beatrizalmeidaf/papero-pdf-text-extractor

    Open-source API and Python library for extracting PDF structure: reading order, tables, formulas, figures and block positions. Runs on CPU without ML models, files are never stored.

    Similar descriptions and use cases

Three GitHub discoveries every day

What they do, why they are gaining interest, and what to check before using them.

Share on X Telegram

Numbers

Total stars
56
Today
0 · ≈ 4 by evening
Forks
0
Issues and pull requests
0
Watchers
1
Language
Python
License
Apache-2.0
Created
September 28, 2026
Last push
September 30, 2026

Star trust

Growth looks organic: forks and discussion are in line with active projects, and stars arrive unevenly, the way people give them.

These are heuristics, not a verdict: we judge by the repository’s behavior, not by a list of stargazers.

Spotted in
README badge

Paste it into your README — the badge shows the current magnitude and links to this page.

Seismograph: 0.9
[![Seismograph](https://gitnova.dev/badge/SeerRay-Lab/Xiaomi-OCR-0.svg?lang=en)](https://gitnova.dev/en/r/SeerRay-Lab/Xiaomi-OCR-0)