Research Engineer - Data Infrastructure

ElevenLabs · United Kingdom · Engineering

Posted 2026-08-28

Apply for this role →

ABOUT THE ROLE

We are looking for a Research Engineer to join the research team at ElevenLabs, focused on the data infrastructure that powers our frontier AI models. The quality of our models is bounded by the quality and scale of the data behind them, and you will own the systems that make world-class data possible. You will thrive in this role if you enjoy:

- Building large-scale data pipelines for collecting, processing, filtering, and transforming datasets used to train state-of-the-art models.

- Training models used in our data processing pipelines, such as classifiers, quality filters, and labeling models.

- Designing data curation strategies such as deduplication, quality scoring, labeling, and augmentation that measurably improve model performance.

- Creating tooling and infrastructure that lets researchers explore and train on massive datasets quickly and reliably.

REQUIREMENTS

We do not require any formal certifications or degrees. Instead, we are seeking enthusiastic engineers who can showcase solving impressively hard problems with artifacts such as past projects, designs, or GitHub contributions. Ideally, you bring:

- Experience building data-intensive systems, ideally in support of machine learning training pipelines.

- Strong engineering skills in distributed data processing at scale (e.g., Kubernetes, or custom pipelines over large datasets).

- The capacity to autonomously evaluate how data quality, composition, and curation affect model outcomes, and to build the tooling to measure it.

Bonus: Experience building or operating web crawlers.

LOCATION

This role is remote and can be executed globally. If you prefer, you can work from our offices in London, New York, San Francisco, and Warsaw.

#LI-Remote

Apply for this role →

← Back to all jobs