Source description
About the role
About the Team
The Seed Infrastructures team oversees the distributed training, reinforcement learning framework, high-performance inference, and heterogeneous hardware compilation technologies for AI foundation models.
Responsibilities
- Design, develop, and optimize high-performance inference systems for large-scale LLMs and VLMs, covering inference engines, serving frameworks, and end-to-end deployment pipelines. 2. Build state-of-the-art model inference engines through advanced performance optimization techniques such as compiler-level optimizations, parallel computing, graph fusion, efficient CUDA kernel development, low-precision computation, streaming inference, speculative decoding, and high-concurrency request optimization. 3. Collaborate closely with other research teams to identify performance bottlenecks, conduct in-depth performance analysis, and optimize large models; contribute to the development of model toolchains and the broader technical ecosystem.
More at ByteDance
Related open roles
Senior Cloud Acceleration Engineer – DPU & AI Infra
Seattle
MLLM Algorithm Engineer Graduate (Search) - 2026 Start (BS/MS)
Singapore
Software Engineer (Treasury) - Global Payments
Singapore
Backend Engineer - Large Model Knowledge System/Content understanding/RAG (ByteDance Singapore)
Singapore
Student Researcher (AI Foundation Model Infrastructure - Seed) - 2027 Start (BS/MS)
Seattle
Research Intern (RDMA/High Speed Network) - 2026 Start (PhD)
San Francisco Bay Area