惯性聚合 高效追踪和阅读你感兴趣的博客、新闻、科技资讯
阅读原文 在惯性聚合中打开

推荐订阅源

D
Docker
V
V2EX
OSCHINA 社区最新新闻
OSCHINA 社区最新新闻
云风的 BLOG
云风的 BLOG
Blog — PlanetScale
Blog — PlanetScale
Recent Announcements
Recent Announcements
Last Week in AI
Last Week in AI
博客园 - Franky
Microsoft Security Blog
Microsoft Security Blog
Hugging Face - Blog
Hugging Face - Blog
H
Hackread – Cybersecurity News, Data Breaches, AI and More
Vercel News
Vercel News
MyScale Blog
MyScale Blog
大猫的无限游戏
大猫的无限游戏
罗磊的独立博客
H
Help Net Security
月光博客
月光博客
Martin Fowler
Martin Fowler
博客园 - 【当耐特】
宝玉的分享
宝玉的分享
P
Proofpoint News Feed
GbyAI
GbyAI
腾讯CDC
freeCodeCamp Programming Tutorials: Python, JavaScript, Git & More

Show HN

GitHub - astefanutti/shaderbang: Shebang for Shaders Show HN: Generate Claude Code Workflows using Spec Driven Development approach Show HN: AI agents for UK GDAD PCF roles and their skills The Two Pillars: Mixer Mode and Meta-Software in the Reorganization of Software Work After AI GitHub - JaiCode08/teleport-env What 1,000+ Harness Experiments Taught Me About Self-Improving Agents Show HN: Liiists, a Markdown-first, iOS and CLI list app SwiperTab – Get this Extension for 🦊 Firefox (en-US) GitHub - kouhxp/fftext: Summarize, explain, fact-check, or translate any text, URL, or file. No GPU. No cloud. One command GitHub - sweetpad-dev/sweetpad: Develop Swift/iOS projects using VSCode GitHub - dogmaticdev/IRON: IRON a.k.a. Intermediate Representation Object Notation is a Interpreter/Database that is used to create Programming Languages. GitHub - sjhalani7/vaen: Package your AI coding harness into a portable .agent file, and share it across repos, teams, & the community without ever having to copy-paste instructions, skills, MCP config, or secrets. Show HN: Gandalf the Grader Show HN: Citadeld – replay any CI failure locally from a single file GitHub - tdortman/cuSBF: High-Performance GPU Super Bloom Filter coral-ai/claude-code-token-xray at main · Coral-Bricks-AI/coral-ai GitHub - ulyssestenn/funes: Funes is a Git-based framework for LLM-managed knowledge work: an AI Librarian ingests raw sources, builds an interlinked Markdown knowledge base, and uses it to produce cited reports, analyses, and other outputs. GitHub - ThatXliner/gah: Git Add Hunk, built for agents to use GitHub - harmont-dev/harmont-cli: Command-line client for the Harmont CI platform GitHub - brooksmcmillin/mcp-authflow: OAuth 2.0 Authorization Server framework for MCP servers GitHub - javaid-codes/audit-supply-chain-agents GitHub - amorey/gochan: A small library of common channel architectures for Go, inspired by Rust GitHub - arifozgun/OpenGem: Free, Open-Source AI API Gateway with Gemini, OpenAI & Anthropic Compatibility in 1 file GitHub - Pranesh950/BioPetals: 🌸 Run BIOxAI models at home, BitTorrent-style. Fine-tuning and inference up to 10x faster than offloading GitHub - cnguyen14/bounty-doctor: Diagnose a GitHub bounty issue before you waste hours: detects honeypot scam repos, AI-bot attempt swarms, and stale contests. Show HN: CoreMCP – MCP Server for On-Prem DBs Show HN: KittyHTML – Render HTML/CSS as an inline image in your terminal GitHub - bingud/filemat: Web-based file manager Show HN: TruthLens – Free multi-signal deepfake image detector GitHub - apexlocal-jz/claude-usage-tray: Windows system-tray app showing your Claude Code rate-limit usage at a glance. Zero deps, ~300 lines of PowerShell. Cross-IDE (works regardless of VS Code, Cursor, plain terminal).
OCR.chat — Free OCR + chat with your PDF or image — text,...
nadermx · 2026-06-27 · via Show HN

OCR that reads your documents — then lets you chat with them

Drop an image or PDF for clean text, tables and math in 100+ languages — then ask questions and get answers cited to the page. No signup to try.

⚡ Fast

No languages match your search.

Fast handles printed text. Premium AI reads handwriting, math, tables & complex layouts.

📄

Drop a file, click to browse, or paste a screenshot

PNG · JPG · WEBP · TIFF · multi-page PDF

🔒 Your files are processed privately and deleted automatically.

💬

Then chat with your document

Once the text is extracted, ask questions and get answers grounded in your document — each one cited back to the page. Summaries, totals, dates, clauses: just ask.

Chat with a PDF

📊

Real tables, not mush

Tables come out as real Markdown/CSV — never misaligned text. Nothing is silently dropped.

👀

See it side by side

Review the extracted text next to your original, with low-confidence spans flagged before you trust them.

⬇️

One paste, every format

Copy or download as TXT, Markdown, DOCX, searchable PDF, JSON, or LaTeX — all from one screen.

🌍

100+ languages

Latin, CJK, Arabic, Cyrillic, Indic and more. Foreign words are transcribed, never auto-translated.

Math → LaTeX

Equations and formulas convert to clean LaTeX you can paste straight into your paper.

🔌

Simple API

One POST request returns clean Markdown or JSON. Metered per page, no surprises.

A tool for every document

Create a free account

Unlock the Premium AI engine for handwriting, math and tables, save your OCR history, and get more pages every month — free to start, no card required.

Sign up free See plans

How it works

1

Upload or paste

Drop an image or PDF, click to browse, or paste a screenshot. Multi-page PDFs and batches welcome.

2

The right engine reads it

A fast engine handles printed text; the premium AI engine tackles handwriting, complex layouts, tables and math — in 100+ languages.

3

Review & export

Check the text side by side with your original, then copy or download as TXT, Markdown, Word, searchable PDF, CSV or JSON.

Why OCR.chat

Feature OCR.chat Typical OCR tools
Try with no signupOften paywalled
Real tables (Markdown/CSV)Misaligned text
Handwriting & math (LaTeX)Rare
100+ languages, no auto-translateLimited
TXT · MD · DOCX · PDF · CSV · JSONOne or two
Side-by-side review + confidence
Simple REST APIEnterprise only

Frequently asked questions

Yes — you can extract text with no signup at all. A free account adds more pages each month, and paid plans unlock unlimited pages, the premium AI engine, batch processing, and the API.

PNG, JPG, WEBP, GIF, BMP, TIFF, and multi-page PDF, in over 100 languages including CJK, Arabic, Cyrillic and Indic scripts.

The fast engine is excellent on printed documents; the premium AI engine handles handwriting, complex layouts, tables, and equations (math to LaTeX, tables to Markdown/CSV), with low-confidence spans flagged.

Plain text, Markdown, Word (DOCX), searchable PDF, CSV, and JSON, plus copy to clipboard.

Files are processed for OCR only and deleted automatically. We never sell, share, or train on your documents.

Yes — one POST request returns clean text, Markdown, or JSON, metered per page.

Most images and short PDFs return in a few seconds; digital PDFs with a text layer are read instantly.

Yes — every page of a multi-page PDF is processed and combined; paid plans add higher page limits and batch processing.

No — OCR.chat runs in your browser, nothing to install. Developers can also use the REST API.

Real tables (Markdown/CSV) not misaligned text, nothing silently dropped, side-by-side review with confidence, handwriting and math, and every export format — without a signup wall.

Yes — the result screen shows the extracted text beside your original with low-confidence spans flagged so you can verify before exporting.

Start free with no signup; paid plans start at $5/month for more pages, the premium AI engine, batch and API, and page packs never expire.