Jobs and Careers
BY

Research Scientist - Seed Multimodal Interaction and World Model - Reinforcement Learning Focus

ByteDance
San Jose, United Statesfull_timeVerifiedPosted 14 Feb 2026
💰 $438,000/yr($187,040/yr$438,000/yr)

About the role

About Seed Team
Established in 2023, the ByteDance Seed team is dedicated to discovering new approaches to general intelligence, and pursuing the edge of intelligence. Our research spans large language models, speech, vision, world models, AI infrastructure, next-generation interfaces and more.
With a long-term vision and determination in AI, the ByteDance Seed team remains committed to foundational research. We aim to become a world-class AI research team that drives real technological progress and delivers societal benefits.
With labs across China, Singapore, and the U.S., our team has already released industry-leading general-purpose large models and advanced multimodal capabilities, powering over 50 real-world applications — including Doubao, Coze, and Jimeng.

- Design and implement reinforcement learning (RL) training systems for large-scale multimodal foundation models
- Develop unified modeling frameworks that integrate video, audio, and language, with a focus on visual latent reasoning
- Explore Reinforcement Learning-based approaches to bridge understanding and generation for multimodal visual reasoning
- Collaborate with researchers to evaluate models on tasks involving world modeling, reasoning, and instruction-conditioned generation

The base salary range for this position in the selected city is $187040 - $438000 annually.

Apply for this role

Generate a tailored application kit with a matched cover letter, interview prep, and CV highlights — in under 60 seconds.

Apply Now →Generate Application Kit

Free account required — sign up in 30s

Company

ByteDance

View company profile →