Seismograph

What’s gaining stars on GitHub right now

datalab-to/chandra

Language modelsPython#ai#ocr
More actions
My radar Open on Seismograph

OCR model that handles complex tables, forms, handwriting with full layout.

About the project

Chandra OCR 2 is an OCR model that converts images and PDFs into structured HTML, Markdown, or JSON while preserving layout, tables, math, and handwriting.

Useful for
  • Convert scanned PDFs to Markdown while preserving tables and layout
  • Extract data from forms and handwritten documents into JSON
  • Run local OCR via CLI or the Streamlit app

README summarized by DeepSeek V4.1 Flash. Details may be inaccurate.

2.1 Steady Magnitude out of 10 — how fast interest is growing, not a quality score.

Star growth looks organic. Data as of October 3, 2026, 14:03 UTC.

The star-growth assessment does not verify whether the project is safe to run.

Stars per day

050100July 6, 2026October 3, 2026

Bars show daily stars; the line is a moving average of the available history. Short histories do not yet establish a reliable usual pace. Red marks spike days.

Why it’s trending

  • 4 stars so far today, about 10 expected by the end of the day.
  • Over the last two days the pace is 2.9× that of the previous week and a half.
  • The spike has held for 2 days in a row — not a one-off blip.
  • GitHub Trending Python today: #5, +20 stars.

Explore related projects

  1. PaddlePaddle/PaddleOCR

    OCR toolkit and document AI engine that turns PDFs and images into structured, LLM-ready data (JSON/Markdown), with support for 100+ languages, tables, formulas and seals.

    Shared topics: ocr
  2. SeerRay-Lab/Xiaomi-OCR-0

    A unified 0.8B vision-language model for document parsing and OCR-centric understanding: outputs structured Markdown, extracts fields as JSON, and answers questions about document images.

    Similar descriptions and use cases
  3. beatrizalmeidaf/papero-pdf-text-extractor

    Open-source API and Python library for extracting PDF structure: reading order, tables, formulas, figures and block positions. Runs on CPU without ML models, files are never stored.

    Similar descriptions and use cases

Three GitHub discoveries in each edition

What they do, why they are gaining interest, and what to check before using them.

Editions at 10:00, 16:00, 22:00 MSK.

Share on X Telegram

Numbers

Total stars
12,391
Today
4 · ≈ 10 by evening
Forks
1,250
Issues and pull requests
117
Watchers
89
Language
Python
License
Apache-2.0
Latest release
v0.2.0 · March 18, 2026
Created
October 8, 2025
Last push
June 26, 2026

Star trust

Growth looks organic: forks and discussion are in line with active projects, and stars arrive unevenly, the way people give them.

These are heuristics, not a verdict: we judge by the repository’s behavior, not by a list of stargazers.

Spotted in
README badge

Paste it into your README — the badge shows the current magnitude and links to this page.

Seismograph: 2.1
[![Seismograph](https://gitnova.dev/badge/datalab-to/chandra.svg?lang=en)](https://gitnova.dev/en/r/datalab-to/chandra)