Comprehensive PDF Data Extraction
Budget / SalaryHourly project
TypeFreelance project
LocationRemote
Posted57 minutes ago
I have a multi-page PDF that combines mixed-format text, embedded images, and several complex tables. I need every element moved into a single Excel workbook so I can work with the data directly.
Here is what I’m after:
• All paragraphs, column text, and captions captured accurately, preserving their original line breaks and order.
• Each table recreated on its own Excel sheet, laid out exactly as it appears in the PDF—same headers, footers, merged cells, and alignment.
• Every image exported at full resolution and placed in a dedicated sheet, with a clear reference to the page it came from.
Please use whatever tools you prefer—Power Query, Python (pdfplumber, tabula-py, camelot), or Adobe Acrobat scripting—as long as the final file opens cleanly in Excel 365 without broken links or missing characters.
I’ll consider the job complete when the workbook mirrors the source document page-for-page, ready for immediate analysis with no manual cleanup required.
Here is what I’m after:
• All paragraphs, column text, and captions captured accurately, preserving their original line breaks and order.
• Each table recreated on its own Excel sheet, laid out exactly as it appears in the PDF—same headers, footers, merged cells, and alignment.
• Every image exported at full resolution and placed in a dedicated sheet, with a clear reference to the page it came from.
Please use whatever tools you prefer—Power Query, Python (pdfplumber, tabula-py, camelot), or Adobe Acrobat scripting—as long as the final file opens cleanly in Excel 365 without broken links or missing characters.
I’ll consider the job complete when the workbook mirrors the source document page-for-page, ready for immediate analysis with no manual cleanup required.
Apply on Freelancer →
Project sourced from Freelancer.com. Applications happen directly on the original platform — we never collect your data.