Archive Oral History Transcripts

via Freelancer ·

Budget / SalaryHourly project
TypeFreelance project
LocationRemote
Posted1 hour ago
I have located a series of publicly accessible oral-history transcripts that display only through a FlipViewer interface; there is no built-in download option. I need every single page of every transcript captured manually, saved as lossless PNG images, and placed into a well-ordered private archive for my ongoing research.
The transcripts are displayed online using a “FlipViewer” interface. They cannot simply be downloaded as PDFs. The work therefore involves opening each available transcript and systematically capturing every page, then organising the resulting images according to a strict file-naming and folder structure.
This is primarily a manual archival/data-collection task. Accuracy, completeness, consistency, and careful file organisation are more important than speed.

The work:

For each oral-history interview assigned to you, you will:
1. Open the interview record
2. Identify each reel for which an online transcript is available.
3. Open the transcript in the FlipViewer.
4. Capture every page of the transcript, at a resolution high enough for the text to remain clearly readable.
5. Save every transcript page as an individual PNG image at the highest available/native resolution. Do not resize, recompress, crop out page numbers or margins, enhance, OCR, or convert the images.
6. Check that no pages have been missed, duplicated, cropped, or captured out of sequence.
7. Organise the files according to the naming system I provide.
8. Record completion and basic metadata in a spreadsheet.

For example, a file might be named:
Smith_003860_001_p001.png
where:
Smith = surname of interviewee
003860 = interview accession number
001 = reel number
p001 = page number

Files will be sorted into folders named for each interviewee.

You will also maintain a simple Google Sheet recording information such as:

- Interviewee
- Accession number
- Reel number
- Whether an online transcript exists
- Number of transcript pages
- Whether capture is complete
- Any missing, unreadable, duplicated, or problematic pages
- Notes on unusual cases

I will provide the master list of interviews and accessions.

Important: The work must be complete and exact. I need the transcript pages as archival research material, so missing even one page is a problem.

Please:
- capture the entire visible transcript page, without cutting off text or page numbers;
- use the highest practical image quality;
- preserve the original page order;
- do not alter, enhance, rewrite, summarise, or OCR the transcript unless specifically asked;
- do not substitute an automated web scrape for manual verification;
- do not share or redistribute the material;
- keep the material only for the duration of the project and delete your local copies after I confirm receipt.

This material is being collected solely for private academic research.

Pilot first
I would like to begin with a small paid pilot using a few oral-history interviews.

The pilot will allow us to establish:
- the best capture method;
- average pages captured per hour;
- the appropriate file structure;
- quality-control procedures; and
- the likely cost of completing the larger collection.

If the pilot is successful, there will potentially be a substantial amount of follow-on work.

Who I am looking for
You do not need specialist historical knowledge.
You do need to be:
- extremely careful and methodical;
- comfortable working with repetitive archival material;
- good at following naming conventions exactly;
- experienced with file and folder organisation;
- comfortable with Google Sheets;
- able to notice missing pages, duplicates, and other inconsistencies;
- reliable over a potentially large project.

Previous experience with archival digitisation, document capture, libraries, data entry, research assistance, or large-scale image/file organisation would be an advantage.

When applying, please answer the following:
1. Have you done similar archival, document-capture, data-entry, or large-scale file-organisation work before?
2. What method would you use to capture a long FlipViewer document page by page while ensuring that no pages are skipped or duplicated?
3. How would you quality-check your own work?
4. Are you willing to complete a small paid pilot before we agree on the larger project?

Please quote either your hourly rate or the price you would charge for the pilot.
Please do not send a generic proposal. I am particularly interested in your answer to questions 2 and 3.

Deliverable: a zipped directory tree containing the interviewee folders filled with correctly named PNG pages, ready for immediate offline study. If this first batch goes smoothly there will be additional collections to process, so meticulous work here can turn into repeat projects.
project management data entry image processing google sheets data collection data management
Apply on Freelancer →

Project sourced from Freelancer.com. Applications happen directly on the original platform — we never collect your data.