PDF Text Extraction Task
Budget / SalaryHourly project
TypeFreelance project
LocationRemote
Posted1 hour ago
I have a batch of PDFs and I need every bit of readable text lifted out of them and delivered as clean, plain-text files—no bolding, italics, or structural markup, just the raw words exactly as they appear. Accuracy is more important than speed, but I’d like to keep the turnaround reasonable, so let me know how many pages you can comfortably handle per day.
If the PDFs contain any scanned pages, feel free to apply OCR with Adobe Acrobat, ABBYY FineReader, Tesseract, or a similar tool, as long as the output stays faithful to the source wording and spelling.
Deliverables:
• One .txt file per source PDF, named to match the original
• All text preserved in original reading order
Before we wrap up, I’ll spot-check a few pages against the originals; once everything matches, the job is complete. Let me know your experience with text extraction and your estimated timeline so we can get started.
If the PDFs contain any scanned pages, feel free to apply OCR with Adobe Acrobat, ABBYY FineReader, Tesseract, or a similar tool, as long as the output stays faithful to the source wording and spelling.
Deliverables:
• One .txt file per source PDF, named to match the original
• All text preserved in original reading order
Before we wrap up, I’ll spot-check a few pages against the originals; once everything matches, the job is complete. Let me know your experience with text extraction and your estimated timeline so we can get started.
Apply on Freelancer →
Project sourced from Freelancer.com. Applications happen directly on the original platform — we never collect your data.