PDF Numerical Data Extraction
Budget / SalaryHourly project
TypeFreelance project
LocationRemote
Posted44 minutes ago
I have a collection of PDF documents that contain numbers scattered across tables, embedded charts, and interactive form fields. I need every one of those figures captured accurately and transferred into a structured spreadsheet or database of your choice—Excel or Google Sheets is fine as long as the final file remains easy for me to filter, sort, and run basic calculations on.
Here is what success looks like for me: every numeric value that appears in a table, chart, or form field inside each PDF ends up in the corresponding row and column of the output file, with the original hierarchy (document → page → table/chart/form) clearly preserved so I can trace any figure back to its source page without opening the PDF again. Totals and subtotals must add up, and any discrepancies should be flagged in a notes column.
I will supply the PDFs as soon as we start. You return:
• A clean, well-labeled spreadsheet containing all extracted numbers
• A brief log noting pages that were unclear or required manual interpretation
Accuracy is more important than speed, but I’d like a realistic turnaround estimate once you’ve seen the sample file. If you already have scripts or OCR tools that help with bulk extraction, feel free to use them; just be sure to proof-read the results before delivery.
Here is what success looks like for me: every numeric value that appears in a table, chart, or form field inside each PDF ends up in the corresponding row and column of the output file, with the original hierarchy (document → page → table/chart/form) clearly preserved so I can trace any figure back to its source page without opening the PDF again. Totals and subtotals must add up, and any discrepancies should be flagged in a notes column.
I will supply the PDFs as soon as we start. You return:
• A clean, well-labeled spreadsheet containing all extracted numbers
• A brief log noting pages that were unclear or required manual interpretation
Accuracy is more important than speed, but I’d like a realistic turnaround estimate once you’ve seen the sample file. If you already have scripts or OCR tools that help with bulk extraction, feel free to use them; just be sure to proof-read the results before delivery.
Apply on Freelancer →
Project sourced from Freelancer.com. Applications happen directly on the original platform — we never collect your data.