Research Engineer - Web Crawlers
Recruiter Fit Breakdown & Candid Summary
ElevenLabs is seeking a Research Engineer to design and operate large-scale web crawlers dedicated to gathering world-class training data for frontier AI models.
The ideal candidate brings deep hands-on experience in building distributed systems, handling web-scale data extraction, and turning messy web sources into clean datasets.
You will operate in a high-velocity, low-bureaucracy environment alongside top-tier talent.
Engineers without a strong background in distributed crawling infrastructure or those who prefer structured, rigid environments should not apply.
Role Responsibilities
- 1Build and operate large-scale, distributed web crawlers that discover, fetch, and extract data across billions of pages.
- 2Solve hard crawling problems such as messy HTML content extraction, web-scale deduplication, freshness strategies, and rate-limit handling.
- 3Design targeted crawling pipelines to source high-value media like audio, video, and multilingual content for training-ready datasets.
- 4Create monitoring, exploration, and request tooling and infrastructure for researchers to interact with newly crawled data.
Skills Matrix
Must-Have Skills
Nice-to-Have Skills
Will an ATS filter reject your resume for this role?
Check your keyword match score, missing must-have skills, and formatting flags — before you apply.
No signups. Free check in 3 seconds.
Ready to submit your application?
Apply directly on ElevenLabs's official job portal.