// 06 · Capability //
Data scraping
Structured data from public web sources, collected on a schedule, cleaned and delivered where you need it. We work within the terms and the law of each source and will say no to jobs that do not.
Scraping is easy to start and hard to keep running. The first script takes an afternoon; the version that still works in six months, survives a layout change and tells you when it breaks is the actual product.
What we build
- Collectors for listings, reviews, maps, prices and public records
- Pipelines that deduplicate, normalise and enrich the results
- Exports to CSV, JSON, PostgreSQL or an API your team calls
- Monitoring that notices when a source changes shape
Where we draw the line
We check each source's terms, robots directives and the applicable law before we build, and we write that assessment down for you. We rate-limit, we identify ourselves where the source expects it, and we do not build around measures a site has put in place to stop automated access. If a job does not pass that check, we will tell you and suggest an alternative source.
// Worked example //
Before
Manual collection from Google Maps and Search, slow and error-prone
After
Hourly structured exports to CSV and JSON with no manual step
Source: Google Maps data extraction case study, carried forward from the previous site. Figures are being re-verified with the client.
// FAQ //