Top 17 Jupyter Notebook OCR Projects
-
deep-text-recognition-benchmark
Text recognition (optical character recognition) with deep learning methods, ICCV 2019
-
SaaSHub
SaaSHub - Software Alternatives and Reviews. SaaSHub helps you find the best software and product alternatives
-
Pix2Text
An Open-Source Python3 tool with SMALL models for recognizing layouts, tables, math formulas (LaTeX), and text in images, converting them into Markdown format. A free alternative to Mathpix, empowering seamless conversion of visual content into text-based representations. 80+ languages are supported.
-
-
-
document-ai-samples
Sample applications and demos for Document AI, the end-to-end document processing platform on Google Cloud
-
deep-text-recognition-benchmark
PyTorch code of my ICDAR 2021 paper Vision Transformer for Fast and Efficient Scene Text Recognition (ViTSTR) (by roatienza)
-
Project mention: Why multimodal AI needs typed artifacts instead of ad-hoc URLs | news.ycombinator.com | 2026-01-06
-
Multi-Type-TD-TSR
Extracting Tables from Document Images using a Multi-stage Pipeline for Table Detection and Table Structure Recognition
-
-
-
Calliar
A dataset for online Arabic calligraphy. A collection of 2500 annotated calligraphic styles.
-
tutorials
Git Repo for Articles on Ergo Sum blog and the youtube channel https://www.youtube.com/channel/UCiie9CN--dazA7iT2sry5FA (by rogerfitz)
-
Documents-Parsing-Lab
Jupyter notebooks testing different OCR models for document parsing (Dolphin, MonkeyOCR, Marker, Nanonets, ...)
I’m excited to share a new project I’ve been working on: Documents-Parsing-Lab
-
-
docutron
Docutron Toolkit: detection and segmentation analysis for legal data extraction over documents.
-
OnlineHTR
Online Handwritten Text Recognition (HTR) system implemented with PyTorch. Based on https://doi.org/10.1007/s10032-020-00350-4.
-
OCR-evaluation
This project is a practical, beginner-friendly guide for users with datasets of scanned images, archival documents, or photos of text who want to extract accurate text using OCR — no deep technical setup required. All workflows run in Google Colab, so you don’t need to install anything locally.
Jupyter Notebook OCR discussion
Jupyter Notebook OCR related posts
Index
What are some of the best open-source OCR projects in Jupyter Notebook? This list will help you:
| # | Project | Stars |
|---|---|---|
| 1 | deep-text-recognition-benchmark | 3,933 |
| 2 | Pix2Text | 3,214 |
| 3 | tarsier | 1,761 |
| 4 | PyMuPDF-Utilities | 722 |
| 5 | document-ai-samples | 327 |
| 6 | deep-text-recognition-benchmark | 312 |
| 7 | vlmrun-cookbook | 310 |
| 8 | Multi-Type-TD-TSR | 286 |
| 9 | videocr-PaddleOCR | 231 |
| 10 | ocrpy | 225 |
| 11 | Calliar | 157 |
| 12 | tutorials | 89 |
| 13 | Documents-Parsing-Lab | 82 |
| 14 | Easter2 | 79 |
| 15 | docutron | 28 |
| 16 | OnlineHTR | 25 |
| 17 | OCR-evaluation | 1 |