OCR Lab — PP-OCRv6 text recognition in your browser
Read text out of a screenshot with Baidu's PP-OCRv6 detection and recognition models running entirely in your browser through onnxruntime-web — the image is never uploaded. The 6 MB ONNX package ships with the app and is cached in your browser after the first load, so there is nothing to bring and nothing to configure. Per-line confidence, bounding boxes, a diagnostics panel and text/CSV/JSON/Markdown export are all free and local, and work signed out; an optional metered pass repairs reading order, corrects flagged character confusions and extracts structured fields, cross-checked against what the recogniser actually produced so it cannot quietly invent an invoice total. You can also supply your own fine-tuned ONNX package if you have one. Models are PP-OCRv6 from PaddlePaddle/PaddleOCR (Apache-2.0), redistributed unmodified with their NOTICE and licence; the browser-side pipeline follows the approach popularised by the PaddleOCR community write-ups, with the detection channel count, BGR normalisation and already-softmaxed recognition head re-derived from the published packages.
Details
gpt-terra Every public app is built from a security-scanned skill and must pass a clean scan — skill and frontend — before it can be listed. Have a skill of your own? Turn it into an app — or read the step-by-step walkthrough.