Notifications

Loading notifications...
Elevenlabs Cover

Research Engineer - Web Crawlers

Elevenlabs
Remote Full Time Negotiable 12 days ago

About the Job

We launched in January 2023 with the first human-like AI voice model. Today, we serve millions of users and thousands of businesses - from fast-growing startups to large enterprises like Deutsche Telekom and Meta. Our investors are some of the world's most prominent, including Andreessen Horowitz, ICONIQ Growth and ...

We have expanded from voice into three main platforms:

ElevenAgents enables businesses to deliver seamless and intelligent customer experiences, with the integrations, testing, monitoring, and reliability necessary to deploy voice and chat agents at scale.

Key Responsibilities

We are looking for a Research Engineer to join the research team at ElevenLabs, focused on large-scale web crawling for our frontier AI models. The quality of our models is bounded by the quality and scale of the data behind them, and you will own the crawling systems that source world-class data from the open web. You will thrive in this role if you enjoy:
Building and operating large-scale, distributed web crawlers that discover, fetch, and extract data across billions of pages reliably and efficiently.
Solving hard crawling problems such as content extraction from messy HTML, deduplication at web scale, freshness and recrawl strategies, and politeness and rate-limit handling.
Designing targeted crawling pipelines that find high-value data sources, including audio, video, and multilingual content, and turn them into clean training-ready datasets.
Creating tooling and infrastructure that lets researchers request, monitor, and explore newly crawled web data quickly and reliably.
 

Required Skills & Abilities

We do not require any formal certifications or degrees. Instead, we are seeking enthusiastic engineers who can showcase solving impressively hard problems with artifacts such as past projects, designs, or GitHub contributions. Ideally, you bring:
Hands-on experience building and scaling web crawlers or scraping systems, ideally in support of machine learning training data.
Strong engineering skills in distributed systems at scale (e.g., Kubernetes, queue-based architectures, or custom pipelines processing billions of documents).
The capacity to autonomously evaluate the quality, coverage, and compliance of crawled data, and to build the tooling to measure it.
 

Apply now

Please let Elevenlabs know you found this job on Job Vista. This helps us grow!