I will extract a public business directory into a clean spreadsheet

United States

I speak English

Business data extraction and spreadsheet cleanup, with sources on every row

I turn public business data into clean spreadsheets: extraction from directories and rosters that allow it, and cleanup and deduplication of the files you already have. Every row I deliver carries its...
About this Gig

You have a public online business directory (a chamber of commerce member list, an industry association directory, a licensing board roster, a trade group member page) and need it as a spreadsheet instead of a page. I read a public directory you point me to and return the listings as clean, deduplicated rows for outreach, market research, or a CRM import.


Scope: only directories publicly accessible without a login, allowed under robots.txt and terms for automated reading. A locked, paywalled, or terms-restricted source is declined before work starts, at no charge.


Only business-level facts are collected: name, category, address, and public contact details where listed. No personal data about private individuals.


Every row carries three provenance columns: source_url (the exact page or file the row came from), retrieved_at (the date I read it), and source_last_updated (the date the source itself states, blank when it states none; blank is a fact, not the same as retrieved_at).


CSV at every tier, XLSX added at Standard and Premium. Extraction is automated, with my review before delivery. The sample image in the gallery shows real, fetched rows, not invented ones.

Technology:

Python

•

Excel

•

Beautiful soup

•

Pandas

Information type:

Listings

•

Websites

Technique:

Automated