Source description
About the role
About the team
The mission of the Seed Speech team is to enrich interactive and creative processes through the application of multimodal speech technologies. The team focuses on the forefront of research and product development in speech and audio, music, natural language understanding, and multimodal deep learning.
Responsibilities
- Develop and scale speech foundation models for understanding and generation tasks. - Design training pipelines including data construction, instruction tuning, and model alignment. - Improve core capabilities such as speech recognition, synthesis, reasoning, and robustness. - Optimize model architectures, training efficiency, and system performance. - Explore natural and interactive interfaces for speech-based systems.
More at ByteDance
Related open roles
Student Researcher (Vision Foundation Model - Seed) - 2027 Start (BS/MS)
San Francisco Bay Area
Structured Data Fusion Large Model Researcher-Risk Control-Soaring Star Talent Program
Singapore
Research Scientist, Vision Foundation Model
San Francisco Bay Area
Senior Research Scientist - DPU & AI Infra
Seattle
Research Scientist, Generative AI Graduate (Intelligent Creation) - Global Frontier Tech Recruitment Program - 2027 Start (PhD)
San Francisco Bay Area
Research Scientist - AI Compute & DPU - Global Frontier Tech Recruitment Program - 2027 Start (PhD)
San Francisco Bay Area