SeerRay-Lab/Xiaomi-OCR-0
More actions
About the project
A unified 0.8B vision-language model for document parsing and OCR-centric understanding: outputs structured Markdown, extracts fields as JSON, and answers questions about document images.
Useful for
- Parse PDF pages or scans into structured Markdown with tables and formulas
- Extract requested fields from documents as JSON via KIE
- Ask questions about document image content via VQA
README summarized by DeepSeek V4.1 Flash. Details may be inaccurate.
Star growth looks organic. Data as of October 3, 2026, 01:56 UTC.
The star-growth assessment does not verify whether the project is safe to run.
Stars per day
Bars show daily stars; the line is a moving average of the available history. Short histories do not yet establish a reliable usual pace. Red marks spike days.
Why it’s trending
- 0 stars so far today, about 4 expected by the end of the day.
- The repository is 5 days old and already has 56 stars. With less than two weeks of history, there's no usual pace to compare the spike against yet.
- Top new repositories this week: #191.
Three GitHub discoveries every day
What they do, why they are gaining interest, and what to check before using them.
Numbers
- Total stars
- 56
- Today
- 0 · ≈ 4 by evening
- Forks
- 0
- Issues and pull requests
- 0
- Watchers
- 1
- Language
- Python
- License
- Apache-2.0
- Created
- September 28, 2026
- Last push
- September 30, 2026
Star trust
Growth looks organic: forks and discussion are in line with active projects, and stars arrive unevenly, the way people give them.
These are heuristics, not a verdict: we judge by the repository’s behavior, not by a list of stargazers.
Spotted in
- October 3, 2026Top new repositories this week: #184