Convert PDF Paragraphs to Text

via Freelancer ·

Budget / Salary₹12,500–37,500
TypeFreelance project
LocationRemote
Posted1 hour ago
I have a batch of PDFs that contain full-length paragraphs. Some pages let you highlight the words, while other pages are only images of the text, so a mix of simple copy-and-paste and OCR will be necessary.

Your task is to pull every paragraph out of each file and deliver clean, spell-checked text in plain .txt or .docx format. I don’t need charts, tables, or original layout—just the words in the order they appear, ready for re-formatting on my end. Accuracy is more important than speed; please proof-read the OCR output so it matches the source exactly.

Deliverables
• One text file per original PDF, named identically to the source
• All paragraphs extracted and combined in reading order
• Obvious OCR artefacts corrected (e.g., “l” vs “1”, broken hyphens)

I’ll review by running a side-by-side comparison of random pages; if the character match rate is essentially perfect, the milestone is approved. If you already work with tools such as Adobe Acrobat Pro, ABBYY FineReader, or Tesseract, you’ll be up and running quickly.
data entry proofreading editing pdf ocr data extraction abbyy finereader adobe acrobat
Apply on Freelancer →

Project sourced from Freelancer.com. Applications happen directly on the original platform — we never collect your data.