Source description
About the role
About the team
The Seed Multimodal Interaction and World Model team is dedicated to developing models that have human-level multimodal understanding and interaction capabilities. The team is working to advance the exploration and development of multimodal assistant products.
Responsibilities
- Develop multimodal foundation models integrating vision, language, audio, and environment signals. - Design and optimize world models for reasoning, planning, and interaction. - Build training pipelines including data curation, alignment, and reinforcement learning. - Improve agent capabilities such as perception, memory, decision-making, and tool use. - Explore next-generation interaction paradigms between humans and intelligent systems.
More at ByteDance
Related open roles
Student Researcher (Vision Foundation Model - Seed) - 2027 Start (BS/MS)
San Francisco Bay Area
Structured Data Fusion Large Model Researcher-Risk Control-Soaring Star Talent Program
Singapore
Research Scientist, Vision Foundation Model
San Francisco Bay Area
Senior Research Scientist - DPU & AI Infra
Seattle
Research Scientist, Generative AI Graduate (Intelligent Creation) - Global Frontier Tech Recruitment Program - 2027 Start (PhD)
San Francisco Bay Area
Research Scientist - AI Compute & DPU - Global Frontier Tech Recruitment Program - 2027 Start (PhD)
San Francisco Bay Area