免费 · PDF
OCR 可搜索 PDF.
为扫描 PDF 添加本地 OCR 文本层,以便搜索和选择。检查结果并保留原始页面。目前不支持阿拉伯语导出。
20 MiB source · 200 source pages · up to 20 selected pages per run. All original pages stay in the exported PDF. Documents stay in your browser; OCR engine and selected models load from this site. Existing digital signatures become invalid. No PDF/A certification or secure redaction.
Printed text works best; handwriting is unreliable. Arabic/RTL and supplementary Unicode characters are unsupported in this PDF export because the engine does not preserve them reliably. Arabic OCR remains available in PDF to Word. Pages with any existing text are skipped, including pages mixing digital text and scanned areas.
使用方法: OCR 可搜索 PDF
输入数据或选择文件,调整设置并生成结果。复制或下载之前,请检查结果。
限制与支持格式
输入内容在您的设备上处理。货币换算仅发送货币代码,以获取带日期的参考汇率。
英语技术说明
20 MiB source, 200 source pages, 20 selected pages per run and 500,000 review characters. Arabic/RTL and supplementary Unicode export unsupported. Printed text only; handwriting is unreliable. Pages with existing text are skipped. No PDF/A, signature preservation or secure redaction.
Ten recognition models are available: English, French, Spanish, German, Portuguese, Italian, Dutch, Hindi, simplified Chinese and Japanese. Native invisible text layers are aligned to original visible page geometry, including CropBox and rotation. All source pages remain in the output; excluded and unselected pages receive no layer. Any existing nonempty text skips the whole page, including mixed digital/scanned pages. Text layers are read-only: rerun OCR or exclude incorrect pages; use PDF-to-Word for manually corrected text. Extra rotation and contrast affect recognition only; no automatic deskew. Generated fixture checks verify searchable text, approximate word selection bounds and unchanged pixels with PDF.js; general accuracy and native viewer behavior need review. Existing digital signatures become invalid. All processing is local.
Recognize a scanned reference number, keep a digital page unchanged and find the reference in the exported PDF.