# SeerRay-Lab/Xiaomi-OCR-0 > A unified 0.8B vision-language model for document parsing and OCR-centric understanding: outputs structured Markdown, extracts fields as JSON, and answers questions about document images. - Magnitude: 0.9 out of 10 — Steady - Stars: 56 total · +0 stars today, ≈ 4 by evening - Star trust: star growth looks organic - Category: Language models · Language: Python · License: Apache-2.0 · Created: 2026-09-28 · Last push: 2026-09-30 - GitHub: https://github.com/SeerRay-Lab/Xiaomi-OCR-0 · Homepage: https://huggingface.co/spaces/SeerRay-Lab/Xiaomi-OCR-0 · Page: https://gitnova.dev/en/r/SeerRay-Lab/Xiaomi-OCR-0 ## Useful for - Parse PDF pages or scans into structured Markdown with tables and formulas - Extract requested fields from documents as JSON via KIE - Ask questions about document image content via VQA ## Why it’s here - 0 stars so far today, about 4 expected by the end of the day. - The repository is 5 days old and already has 56 stars. With less than two weeks of history, there's no usual pace to compare the spike against yet. - Top new repositories this week: #192. ## Star trust Star growth looks organic. Star-trust labels are heuristics based on the repository’s behavior, not a check of every stargazer. ## Numbers - Forks: 0 - Issues and pull requests: 0 - Watchers: 1 - Average over the last week: 10 per day - Usual pace: too little history (under two weeks) - Stars in the last hour (measured): 0 ## Stars per day, last 7 days (oldest → newest, today is partial) 2026-09-27 … 2026-10-03: 0, 0, 16, 32, 3, 5, 0 ## Spotted in now - Top new repositories this week: #192 ## Similar by description 1. **PaddlePaddle/PaddleOCR** — 2.0 · Cooling · Language models · Python · +0 stars today, ≈ 23 by evening OCR toolkit and document AI engine that turns PDFs and images into structured, LLM-ready data (JSON/Markdown), with support for 100+ languages, tables, formulas and seals. Full card: https://gitnova.dev/en/r/PaddlePaddle/PaddleOCR.md 2. **run-llama/liteparse** — 1.1 · Cooling · Data & analytics · Rust · +0 stars today, ≈ 7 by evening A local Rust document parser that extracts text with bounding boxes from PDF, DOCX, XLSX, PPTX and images, with OCR via Tesseract or HTTP servers. Runs without cloud or LLMs, with bindings for Python, Node.js, WASM and a CLI. Full card: https://gitnova.dev/en/r/run-llama/liteparse.md 3. **beatrizalmeidaf/papero-pdf-text-extractor** — 4.8 · Early signal · Data & analytics · Python · +0 stars today, ≈ 27 by evening Open-source API and Python library for extracting PDF structure: reading order, tables, formulas, figures and block positions. Runs on CPU without ML models, files are never stored. Full card: https://gitnova.dev/en/r/beatrizalmeidaf/papero-pdf-text-extractor.md 4. **mikefarah/yq** — 2.0 · Quiet · Developer tools · Go · +0 stars today, ≈ 2 by evening A portable command-line processor for YAML, JSON, XML, CSV, TOML, HCL and properties using jq-like syntax. It lets you read, filter and modify structured config files directly from the shell. Full card: https://gitnova.dev/en/r/mikefarah/yq.md 5. **StarTrail-org/PixelRAG** — 1.1 · Steady · Language models · Python · +0 stars today, ≈ 4 by evening PixelRAG is a visual RAG system that renders web pages, PDFs and images into screenshot tiles and retrieves over the images directly, preserving tables, charts and layout lost when parsing to text. Full card: https://gitnova.dev/en/r/StarTrail-org/PixelRAG.md --- Magnitude (0–10) measures how fast and how unusually interest in a repository is growing right now. It is not a quality score. Days are UTC. “So far today” is a fact; “expected by the end of the day” is a forecast. Summaries and use cases are written by an LLM (DeepSeek V4.1 Flash) from the README and may be inaccurate: verify specific claims (benchmarks, speed, hardware) in the repository itself. Data as of 2026-10-03 03:19 UTC, updated every 30 minutes.