OCR Lab — PP-OCRv6 text recognition in your browser

ocr-lab.skillsafe.ai

Clean

Read text out of a screenshot with Baidu's PP-OCRv6 detection and recognition models running entirely in your browser through onnxruntime-web — the image is never uploaded. The 6 MB ONNX package ships with the app and is cached in your browser after the first load, so there is nothing to bring and nothing to configure. Per-line confidence, bounding boxes, a diagnostics panel and text/CSV/JSON/Markdown export are all free and local, and work signed out; an optional metered pass repairs reading order, corrects flagged character confusions and extracts structured fields, cross-checked against what the recogniser actually produced so it cannot quietly invent an invoice total. You can also supply your own fine-tuned ONNX package if you have one. Models are PP-OCRv6 from PaddlePaddle/PaddleOCR (Apache-2.0), redistributed unmodified with their NOTICE and licence; the browser-side pipeline follows the approach popularised by the PaddleOCR community write-ups, with the detection channel count, BGR normalisation and already-softmaxed recognition head re-derived from the published packages.

Share

Details

PricingUsage-based + 10% creator margin
Billed model rate$2.75 in / $16.50 out per 1M tokens
Creator margin+10%
Effective rate$3.00 in / $18.00 out per 1M tokens
Security scanClean — skill and frontend scanned
Model gpt-terra
Created2026-08-13
Updated2026-08-18

Every public app is built from a security-scanned skill and must pass a clean scan — skill and frontend — before it can be listed. Have a skill of your own? Turn it into an app — or read the step-by-step walkthrough.