Extract English Text from Scanned PDF
Run OCR on up to five PDF pages and download recognized English text in page order.
Recognize printed English text in a PNG or JPEG with temporary server-side Tesseract OCR.
Selecting a file does not upload it. Upload and convert sends it to temporary server storage. One file at a time; 10 MiB maximum.
Supported formats: .png, .jpg, .jpeg
OCR limit: five PDF pages, four megapixels per page and sixteen megapixels total. OCR results require human review.
The isolated OCR worker checks the image content, recognizes printed words and returns a UTF-8 text file. Selecting an image does not upload it; processing starts only after you choose Upload and convert.
One PNG or JPEG, at most four megapixels and 6000 pixels per side. Animated and multipage images are unsupported. Text output does not preserve fonts or layout. Blank images produce an empty text result. English printed text only. Maximum 10 MiB, five PDF pages, four megapixels per rendered page and sixteen megapixels total. Encrypted, malformed and active-content PDFs are rejected. Handwriting, unusual layouts and poor scans can produce missing or incorrect words. Review the result; no accuracy percentage is promised.
Your file is uploaded only when you choose to convert. An isolated worker processes it without network access. Uploaded files expire after one hour; completed outputs expire one hour after completion. Failed, cancelled or deleted jobs become eligible for immediate cleanup. Service outages can delay physical deletion; a previously issued download link can remain valid briefly. Clear files to request early deletion. Downloads saved on your device remain there.
PDF processing is limited to 200 pages; HTML and reverse conversions have the stricter limits listed above. Complex files may reach time, memory or scratch limits sooner. No malware-free, perfect-fidelity or certified-accessibility claim is made.
Handwriting is not supported reliably. Use clear printed English text and check names, numbers and punctuation.
It applies grayscale contrast adjustment, light noise filtering, orientation detection and small-angle deskew. You can turn it off; cleanup can alter fine details.
Run OCR on up to five PDF pages and download recognized English text in page order.
Create a PDF with scanned page images and a searchable English OCR text layer.