Scraping engineer
Python, Scrapy, Playwright and Crawlee. Builds and maintains scrapers end to end: discovery, extraction, pagination, retries, storage. Typically one to two new sites a week alongside maintenance of what is already running.
Hire the team
One engineer or a squad, by the month. They work in your repo and your tracker, report daily, and ship scrapers that keep running.
Roles
Python, Scrapy, Playwright and Crawlee. Builds and maintains scrapers end to end: discovery, extraction, pagination, retries, storage. Typically one to two new sites a week alongside maintenance of what is already running.
Cloudflare, Akamai, DataDome, PerimeterX, Kasada. TLS and browser fingerprinting, mobile and private API reverse engineering, login and session flows. The person you bring in for the sites that keep returning 403.
Everything around the scraper: delivery pipelines to S3, BigQuery and Google Sheets, scheduling and monitoring, internal dashboards, and the small APIs that put scraped data in front of your team.
What you get
Onboarding
You grant repo, tracker and chat access. A 30-minute call covers the sites, the fields, and how you want work reported.
The engineer opens a pull request for the first site in your repo, following your conventions, with a sample output attached.
The first scraper runs end to end and lands data where you asked for it. You review rows, not slides.
Daily updates, a weekly report, and a backlog you control. New sites go in the tracker; the engineer picks them up.
Questions
Our standard day covers the whole European morning and the early US East Coast morning. For US teams we shift the day later so there are three to four shared hours; tell us what you need and it is set before day one.
You do, the same way you manage your own team: tickets in your tracker, reviews on your pull requests. On our side a lead checks in weekly, reads the weekly report before it reaches you, and steps in if something stalls. You never have to manage us managing them.
You do. Code is written in your repository under your licence from the first commit, and scraped data is delivered to your storage. We keep nothing after the engagement ends except what you ask us to keep running.
Either way. Most teams prefer the engineer to run on their own proxy and cloud accounts so the bills and access stay with them. If you do not have accounts yet, we run on ours and add the usage at cost to the monthly invoice, with receipts.
Yes. Send yours or use ours; we sign before the kickoff call and before any access is shared. Engineers assigned to your project sign the same terms individually.
Book a 15-minute call or send us the sites and what you need. We propose an engineer and a start date, usually within a week, and send a one-page agreement with the first invoice. Day one begins when access is granted.
A 15-minute call is enough to match you with the right person and a start date.
Hi! Send me the site you need data from and I'll get an engineer to look at it.
Chat on WhatsApp Or get a quote