PDF Text & Table Extraction

via Freelancer ·

Budget / SalaryHourly project
TypeFreelance project
LocationRemote
Posted2 hours ago
I have a collection of structured PDFs that contain two main elements I need captured: the running text and a handful of simple, clearly-delineated tables. Your task is to pull both elements out accurately and place them into a single Excel workbook, setting up one sheet that lists the extracted text in reading order and a separate sheet that reproduces each simple table row-for-row.

Accuracy is key—line breaks in the text should match the source, and table layouts must stay intact without losing column alignment. Feel free to use Python (pdfplumber, PDFMiner, Camelot, Tabula, or another reliable library), Adobe tools, or any extraction workflow you trust, as long as I receive:

• Excel file (.xlsx) with “Text” and “Tables” sheets populated
• No stray headers, footers, or page numbers unless they appear inside the body copy
• Consistent column formatting for all tables

Let me know the approach and turnaround you can commit to, and I’ll send a sample PDF so you can confirm everything parses cleanly before moving on to the full batch.
data processing excel pdf data extraction data analysis data management
Apply on Freelancer →

Project sourced from Freelancer.com. Applications happen directly on the original platform — we never collect your data.