Papyrus - Z29K
ARTICLES GEMSHELF
FR|EN

// OCR

8 entries tagged with "ocr"

Chandra

Chandra

OCR model that turns images and PDFs into Markdown, HTML or JSON while keeping layout. Handles tables, math, handwriting and forms across 90+ languages.

2026.07.20READ
Capso

Capso

Native macOS screenshot and screen recording app in the vein of CleanShot X. Annotation, OCR, webcam recording, a video editor and capture history built in.

2026.07.15READ
PaddleOCR

PaddleOCR

Open-source OCR toolkit that turns images and PDFs into structured, LLM-ready data (Markdown, JSON). Handles 100+ languages, tables, formulas and charts.

2026.06.14READ
Receipt OCR

Receipt OCR

Python OCR engine for receipts: raw text extraction via Tesseract and structured data via LLMs (OpenAI, Gemini, Groq). Ships with a CLI and FastAPI.

2026.06.05READ
Falcon-Perception

Falcon-Perception

PyTorch inference engine for vision-language models: detection, segmentation and OCR driven by natural language queries, with paged KV cache and CUDA graphs.

2026.06.01READ
MarkItDown

MarkItDown

Microsoft's Python library and CLI to convert PDF, Word, Excel, images, audio and more into clean Markdown, ready for LLMs and text analysis pipelines.

2026.04.23READ
Stirling PDF

Stirling PDF

Self-hosted open-source PDF toolkit with 50+ built-in tools. Merge, split, OCR, sign, convert and compress PDFs with REST API and workflow automation.

2026.04.05READ
Tesseract.js

Tesseract.js

JavaScript OCR library that runs in the browser and Node.js. Recognizes text in 100+ languages via WebAssembly with no server-side processing needed.

2026.03.31READ

// NAVIGATION

ARTICLESGEMSHELF

// DIRECTORY

THOUGHTS
z29k

Digital Partner
Strategy • Organization • Execution

© 2026 Papyrus - Z29K - z29k.fr - Build with Zola