惯性聚合 高效追踪和阅读你感兴趣的博客、新闻、科技资讯
阅读原文 在惯性聚合中打开

推荐订阅源

月光博客
月光博客
Martin Fowler
Martin Fowler
Last Week in AI
Last Week in AI
罗磊的独立博客
阮一峰的网络日志
阮一峰的网络日志
博客园 - 【当耐特】
奇客Solidot–传递最新科技情报
奇客Solidot–传递最新科技情报
博客园 - 三生石上(FineUI控件)
S
SegmentFault 最新的问题
V
Visual Studio Blog
Hugging Face - Blog
Hugging Face - Blog
雷峰网
雷峰网
博客园_首页
人人都是产品经理
人人都是产品经理
量子位
美团技术团队
The Cloudflare Blog
小众软件
小众软件
WordPress大学
WordPress大学
有赞技术团队
有赞技术团队
M
MIT News - Artificial intelligence
Microsoft Security Blog
Microsoft Security Blog
D
DataBreaches.Net
博客园 - Franky

Hacker News: Show HN

PurrrrrFocus: Pomodoro Timer App - App Store Workflow Engine — Multi-Step Orchestration for Bun RapidPhoto: Pro Photo Editor App - App Store GitHub - DheerG/swarms: Achieve extraordinary results with claude code across a variety of tasks SPICE simulation → oscilloscope → verification with Claude Code — Lucas Gerads Show HN: VCoding – A 5 MB native Windows IDE with no dynamic dependencies Show HN: LLMs don't hallucinate because they're bad at math, it's the format GitHub - Agent-FM/agentfm-core: AgentFM is a peer-to-peer network that turns everyday computers into a decentralized AI supercomputer. AgentFM lets you run massive AI workloads directly across a global mesh of idle CPUs and GPUs. Show HN: Tracking Top US Science Olympiad Alumni over Last 25 Years GitHub - Potarix/agent-hub: One place to talk to all your agents Show HN: Runtime security for AI agents(injection,tool abuse, data exfiltration) GitHub - dubeyKartikay/lazyspotify: Terminal Spotify client for macOS and Linux GitHub - the-banana-tool/king-louie: Easy to use GUI Personal AI Assistant. Win/Linux/Mac. Show HN I made my vacation rental bookable by AI agents–no Airbnb, 0% commission GitHub - basteez/jsf-autoreload: maven plugin to enable hot reload on jsf projects uvm32/hosts/host-gdbstub at main · ringtailsoftware/uvm32 GitHub - labsai/EDDI: Config-driven engine that turns JSON into production-grade AI agents. Multi-agent orchestration, 12+ LLM providers, MCP/A2A protocols, RAG, persistent memory, and enterprise compliance (EU AI Act, GDPR, HIPAA). Built on Quarkus. GitHub - glitchnsec/fortyone-oss: AI Executive Assistant Platform Quickstart | Alien GitHub - muxshed/shed: One stream in, or many. Every destination, simultaneously. No cloud middleman, no per-channel fees, no limits. GitHub - ocrbase-hq/ocrbase: 📄 PDF/IMG ->.MD/JSON Document OCR API for PaddleOCR and GLMOCR. Self-hostable. GitHub - impactjo/home-memory: MCP server that lets your AI assistant remember everything about your home. GitHub - Sets88/dbcls: DbCls is a powerful terminal database client that supports various databases GitHub - neptun2000/heor-agent-mcp GitHub - SeanFDZ/macmind: Single-layer transformer in HyperTalk for the classic Macintosh RollQuation: Math Puzzles - Apps on Google Play GitHub - dropbox/witchcraft Show HN: Agent-cache – Multi-tier LLM/tool/session caching for Valkey and Redis GitHub - opentalon/opentalon: OpenTalon is an open-source platform built from the ground up in Go as a robust alternative to OpenClaw LinkedIn™ 职位抓取工具 - Chrome 应用商店
OCR.chat — Free OCR + chat with your PDF or image — text,...
nadermx · 2026-06-27 · via Hacker News: Show HN

OCR that reads your documents — then lets you chat with them

Drop an image or PDF for clean text, tables and math in 100+ languages — then ask questions and get answers cited to the page. No signup to try.

⚡ Fast

No languages match your search.

Fast handles printed text. Premium AI reads handwriting, math, tables & complex layouts.

📄

Drop a file, click to browse, or paste a screenshot

PNG · JPG · WEBP · TIFF · multi-page PDF

🔒 Your files are processed privately and deleted automatically.

💬

Then chat with your document

Once the text is extracted, ask questions and get answers grounded in your document — each one cited back to the page. Summaries, totals, dates, clauses: just ask.

Chat with a PDF

📊

Real tables, not mush

Tables come out as real Markdown/CSV — never misaligned text. Nothing is silently dropped.

👀

See it side by side

Review the extracted text next to your original, with low-confidence spans flagged before you trust them.

⬇️

One paste, every format

Copy or download as TXT, Markdown, DOCX, searchable PDF, JSON, or LaTeX — all from one screen.

🌍

100+ languages

Latin, CJK, Arabic, Cyrillic, Indic and more. Foreign words are transcribed, never auto-translated.

Math → LaTeX

Equations and formulas convert to clean LaTeX you can paste straight into your paper.

🔌

Simple API

One POST request returns clean Markdown or JSON. Metered per page, no surprises.

A tool for every document

Create a free account

Unlock the Premium AI engine for handwriting, math and tables, save your OCR history, and get more pages every month — free to start, no card required.

Sign up free See plans

How it works

1

Upload or paste

Drop an image or PDF, click to browse, or paste a screenshot. Multi-page PDFs and batches welcome.

2

The right engine reads it

A fast engine handles printed text; the premium AI engine tackles handwriting, complex layouts, tables and math — in 100+ languages.

3

Review & export

Check the text side by side with your original, then copy or download as TXT, Markdown, Word, searchable PDF, CSV or JSON.

Why OCR.chat

Feature OCR.chat Typical OCR tools
Try with no signupOften paywalled
Real tables (Markdown/CSV)Misaligned text
Handwriting & math (LaTeX)Rare
100+ languages, no auto-translateLimited
TXT · MD · DOCX · PDF · CSV · JSONOne or two
Side-by-side review + confidence
Simple REST APIEnterprise only

Frequently asked questions

Yes — you can extract text with no signup at all. A free account adds more pages each month, and paid plans unlock unlimited pages, the premium AI engine, batch processing, and the API.

PNG, JPG, WEBP, GIF, BMP, TIFF, and multi-page PDF, in over 100 languages including CJK, Arabic, Cyrillic and Indic scripts.

The fast engine is excellent on printed documents; the premium AI engine handles handwriting, complex layouts, tables, and equations (math to LaTeX, tables to Markdown/CSV), with low-confidence spans flagged.

Plain text, Markdown, Word (DOCX), searchable PDF, CSV, and JSON, plus copy to clipboard.

Files are processed for OCR only and deleted automatically. We never sell, share, or train on your documents.

Yes — one POST request returns clean text, Markdown, or JSON, metered per page.

Most images and short PDFs return in a few seconds; digital PDFs with a text layer are read instantly.

Yes — every page of a multi-page PDF is processed and combined; paid plans add higher page limits and batch processing.

No — OCR.chat runs in your browser, nothing to install. Developers can also use the REST API.

Real tables (Markdown/CSV) not misaligned text, nothing silently dropped, side-by-side review with confidence, handwriting and math, and every export format — without a signup wall.

Yes — the result screen shows the extracted text beside your original with low-confidence spans flagged so you can verify before exporting.

Start free with no signup; paid plans start at $5/month for more pages, the premium AI engine, batch and API, and page packs never expire.