Choose one image
Import a screenshot, scan, product image or photo, then rotate or enable gentle scan enhancement only when needed.
Your image and recognized text are not uploaded. The current unified models cover Simplified and Traditional Chinese, English and major Latin-script languages; Professional and Ultimate also cover Japanese.
These facts describe the current product, not unshipped roadmap work.
Reviewed by the DataDance product and engineering team · Published · Updated
Confirm important settings before export; the processing location is always disclosed.
Import a screenshot, scan, product image or photo, then rotate or enable gentle scan enhancement only when needed.
Fast is the mobile tier; Professional is the desktop default and adds Japanese; Ultimate targets difficult small text.
Click a polygon to find its line. Copy or TXT keeps visual rows and separates detected table columns with tabs; Markdown emits real tables where structure is stable.
Fast minimises download and memory for clear screenshots and mobile devices. It covers Chinese, English, German, Spanish, Portuguese, French and other supported Latin-script languages. Professional balances photographed text and scans and adds Japanese. Ultimate is desktop-only and spends more download, memory and time on difficult small text. Every tier pins the model name, byte size and SHA-256 so the engine cannot silently change underneath a result.
The source image is preserved, with manual rotation and optional gentle scan enhancement instead of forced binarisation. After recognition, geometry clusters text into visual rows, reconstructs reading order and detects stable table columns. Copy and TXT use tabs between columns, Markdown emits table syntax, and JSON retains corrected text, polygons, confidence and layout rows. The engine runs in its own Worker; changing tiers, clearing or exiting terminates it to reduce retained memory.
Fast supports Simplified Chinese, Traditional Chinese, English, German, Spanish, Portuguese and French. Professional and Ultimate add Japanese. The current unified PP-OCRv6 models do not include Korean recognition, so the Korean interface states that boundary instead of silently returning unreliable output.
Clear print, screenshots, scans, labels and mixed layouts are the primary scope. Handwriting, formula semantics and arbitrary merged-cell tables are not claimed. Low-confidence output stays visible for proofreading instead of being silently rewritten.
Runtime and model files download to your browser, but the source image and recognized text are never uploaded. Mobile exposes only the Fast tier and limits oversized input to reduce the risk of WeChat or Safari closing under memory pressure.
Updated 2026-08-25 · Live capability detection inside the tool is authoritative
These visible answers match the current product behaviour and structured data.
No. Decoding, detection, recognition, layout reconstruction and proofreading all run in the current browser.
Use Fast on mobile, Professional by default on desktop or for Japanese, and Ultimate only for difficult small text or many low-confidence lines.
Fast covers Chinese, English, German, Spanish, Portuguese and French. Professional and Ultimate add Japanese. The current unified model does not recognize Korean.
ImgIng reconstructs visual rows from text geometry. Copy and TXT insert tabs between detected columns, Markdown emits table syntax when consecutive rows share a stable column count, and JSON preserves geometry for further processing.
Each task page documents real settings, limits and format advice—not keyword-swapped duplicates.