Accurate PDF Text Extraction
Budget / Salary$30–250
TypeFreelance project
LocationRemote
Posted1 hour ago
I have three large, text-heavy PDFs whose existing OCR is unreliable. In total there are about 1900 pages in English . I need each file converted into clean, searchable text I can drop straight into AI workflows, so absolute accuracy is more important than preserving fancy layouts.
Scope of work
• Run high-quality OCR (ABBYY FineReader, Adobe Acrobat Pro, Tesseract, or a comparable engine).
• Proof and correct the output so obvious recognition errors, missing characters, and split words are gone.
• Deliver either a single .txt file or a simple .docx per PDF—plain text only, no special formatting required.
Acceptance criteria
1. Every page accounted for; no missing paragraphs.
2. Error rate low enough that a quick read shows no garbled words or stray symbols.
3. File names match the originals for easy reference.
If you have an efficient workflow for large volumes of text and can turn this around quickly, let’s get started.
I own the PDFs and have paid for them. I also own the books.
Scope of work
• Run high-quality OCR (ABBYY FineReader, Adobe Acrobat Pro, Tesseract, or a comparable engine).
• Proof and correct the output so obvious recognition errors, missing characters, and split words are gone.
• Deliver either a single .txt file or a simple .docx per PDF—plain text only, no special formatting required.
Acceptance criteria
1. Every page accounted for; no missing paragraphs.
2. Error rate low enough that a quick read shows no garbled words or stray symbols.
3. File names match the originals for easy reference.
If you have an efficient workflow for large volumes of text and can turn this around quickly, let’s get started.
I own the PDFs and have paid for them. I also own the books.
Apply on Freelancer →
Project sourced from Freelancer.com. Applications happen directly on the original platform — we never collect your data.