I will extract tables from text based pdfs into clean CSV files
Python Data Automation and Web Scraping Specialist
About this Gig
I extract tables from text-based PDFs into CSV files for spreadsheet import.
Please send a representative sample through Fiverr before ordering so I can confirm the layout, pages, tables and scope. Only send documents you own or are authorized to process.
Deliverables: CSV output, source file/page/table references, and a checklist of agreed counts and key cells checked against your source. Complex layouts may require manual review.
Not included: scanned or image-only PDFs, OCR, native Excel (.xlsx) or JSON output, password unlocking, or automatic cross-page table merging.
For a reusable CLI, the supported layout and batch scope must be agreed on a sample. CSV import settings matter for leading zeros and formula-like values.
FAQ
Can you process scanned or image-only PDFs?
This service covers text-based PDFs only. OCR is not included. Please provide a representative sample so I can confirm whether the tables can be extracted.
Can the reusable CLI handle future files?
The CLI can process PDF folders. Layout compatibility and output checks must be agreed on a representative sample; arbitrary layouts and thousands of files are not pre-validated.
How are files handled for this service?
Please send only files you are authorized to share. Before processing, we will agree on the working copies, delivery files and retention period. This extraction tool does not automatically delete files after delivery. Do not send credentials or unrelated personal information.
