Source description
About the role
About the team
The Vision-Applied Research team focuses on applied research in Generative AI and CV/Multimodal Understanding, and delivering intelligent solutions to TikTok, and Lemon8, enabling users to make and share creative content in a much easier way. The team has research groups dedicated to generative models for content creation, image generation, video synthesis, intelligent image/video editing, and virtual humans.
We are seeking an experienced Multimodal Model Training and Inference Optimization Engineer with expertise in optimizing AI model training and inference, including distributed training/inference and acceleration. The ideal candidate will work at the cutting edge of AI efficiency, enhancing the performance, scalability, and deployment of large-scale generative AI models.
Responsibilities
- Optimize large model training pipelines to improve efficiency, speed, and scalability. - Develop and improve distributed training strategies such as data parallelism, model parallelism, pipeline parallelism and communication to accelerate model training. - Benchmark and profile deep learning models to identify performance bottlenecks and optimize computational resources.
More at ByteDance
Related open roles
Student Researcher (Vision Foundation Model - Seed) - 2027 Start (BS/MS)
San Francisco Bay Area
Structured Data Fusion Large Model Researcher-Risk Control-Soaring Star Talent Program
Singapore
Research Scientist, Vision Foundation Model
San Francisco Bay Area
Senior Research Scientist - DPU & AI Infra
Seattle
Research Scientist, Generative AI Graduate (Intelligent Creation) - Global Frontier Tech Recruitment Program - 2027 Start (PhD)
San Francisco Bay Area
Research Scientist - AI Compute & DPU - Global Frontier Tech Recruitment Program - 2027 Start (PhD)
San Francisco Bay Area