datalab-to/chandra
More actions
OCR model that handles complex tables, forms, handwriting with full layout.
About the project
Chandra OCR 2 is an OCR model that converts images and PDFs into structured HTML, Markdown, or JSON while preserving layout, tables, math, and handwriting.
Useful for
- Convert scanned PDFs to Markdown while preserving tables and layout
- Extract data from forms and handwritten documents into JSON
- Run local OCR via CLI or the Streamlit app
README summarized by DeepSeek V4.1 Flash. Details may be inaccurate.
Star growth looks organic. Data as of October 3, 2026, 14:03 UTC.
The star-growth assessment does not verify whether the project is safe to run.
Stars per day
Bars show daily stars; the line is a moving average of the available history. Short histories do not yet establish a reliable usual pace. Red marks spike days.
Why it’s trending
- 4 stars so far today, about 10 expected by the end of the day.
- Over the last two days the pace is 2.9× that of the previous week and a half.
- The spike has held for 2 days in a row — not a one-off blip.
- GitHub Trending Python today: #5, +20 stars.
Three GitHub discoveries in each edition
What they do, why they are gaining interest, and what to check before using them.
Editions at 10:00, 16:00, 22:00 MSK.
Numbers
- Total stars
- 12,391
- Today
- 4 · ≈ 10 by evening
- Forks
- 1,250
- Issues and pull requests
- 117
- Watchers
- 89
- Language
- Python
- License
- Apache-2.0
- Latest release
- v0.2.0 · March 18, 2026
- Created
- October 8, 2025
- Last push
- June 26, 2026
Star trust
Growth looks organic: forks and discussion are in line with active projects, and stars arrive unevenly, the way people give them.
These are heuristics, not a verdict: we judge by the repository’s behavior, not by a list of stargazers.
Spotted in
- October 3, 2026GitHub Trending Python today: #5, +20 stars