I will convert your PDF tables into clean excel with a reusable python script

Uzbekistan

I speak English, Russian, Uzbek

Python Automation for Excel, PDF and Data Reports

I build small Python tools that take a boring, repeating job off your desk. Most of my work looks like this: you have data in one shape - PDFs, a CRM export, a folder of Excel files, a supplier's web...
About this Gig

Your PDF has a table. You need it in Excel, and not once but every week.


Most sellers retype it by hand and charge per page. I write you a Python script instead: you run it yourself, as often as you like, on as many files as you like.


What makes it hold up on real documents: tables that break across pages with the header repeating (the repeats are dropped), subtotal lines sitting in the middle of the document (separated from real line items), numbers written as 1 234 567,89 or 1,234,567.89 (both read correctly), and long item names wrapping inside a cell (kept together).


It also checks its own work: the sum of the extracted rows is compared against the total printed in the PDF, and quantity x price is recomputed for every line. If those disagree you see it in the output, instead of finding out a month later.


You get the clean Excel file, the script, and a short README so you can run it yourself.


Scanned PDFs (images rather than text) need OCR. Message me with a sample first.

Convert from:

PDF

Convert to:

XLS, XLSX