APPLY TO SPEEDRUN
← Twelve Labs
Twelve Labs · Hiring

Senior ML Research Scientist, Perception Models

Seoul, South KoreaHybridFull Time

Who we are

Video is 90% of the world's data. Most of it is invisible to machines.

TwelveLabs builds the intelligence layer to change that. Our multimodal AI models understand video the way humans do — across sight, sound, and motion — and power production-scale AI workloads across media, entertainment, sports, security, and government.

We have raised more than $210 million from NEA, Radical Ventures, Amazon, NVIDIA, Snowflake, Databricks, Index Ventures, NAVER Ventures, Korea Investment Partners, Quadrille Capital, Red Bull Ventures, and AI pioneers including Fei-Fei Li, Silvio Savarese, and Alexandr Wang.

We are a global company, headquartered in San Francisco with offices in Seoul, New York, and London, and employees around the world. We believe the differences in our cultural, educational, and life experiences make our products stronger. Building technology that understands the world in all its complexity requires people who see it from every angle. We are looking for individuals who are driven by hard problems and want their work to matter. Come build it with us!

About the Team

TwelveLabs builds multimodal foundation models that understand what happens in video and how visual information is organized across time and space. The Perception Models team develops general-purpose representations that work across full videos, clips, regions, and individual entities.

Rather than building a collection of narrow, task-specific models, we use signals from capabilities such as detection, segmentation, and tracking to strengthen shared video representations.

About the Role

As a Senior ML Research Scientist, you will lead research at the intersection of spatiotemporal modeling and multimodal representation learning. You will define open-ended research problems, validate ideas through rigorous large-scale experiments, and turn promising results into production capabilities.

This role is ideal for a hands-on researcher with deep expertise in at least one of these areas and a strong interest in connecting them.

In this role, you will

Even if you don't check every box, we encourage you to apply.

If you're a zero-to-one achiever, a ferocious learner, and a kind team player who motivates others, you'll find a home at TwelveLabs.

You may be a good fit if you have

We evaluate based on relevant technical skills and sustained industry impact. This role is typically a strong fit for engineers with an MS and deep industry experience who have evolved from individual contributor to technical leader in production ML environments.

Preferred Qualifications

What makes this role unique

This is an opportunity to define how a foundation model understands the same video across multiple levels of granularity, not simply improve a single task-specific benchmark.

Others

Benefits and Perks

Interested in This Role?

Apply at Twelve Labs

You'll head to Twelve Labs's own careers page.