Padmi
Morph logo
Morph

LLM inference optimization · code-generation models

Senior Machine Learning Infrastructure Engineer

San Francisco Bay Area$130k–$185k/yrPosted 7 days ago
InfrastructureSeniorFull Time
Apply at Morph

Opens the source posting on ycombinator.com

Source description

About the role

View original

Goal: 99.99% uptime

We serve custom inference stacks that have irregular GPU load.

We're looking for people that have done genuinely amazing work in infrastructure that are interested in a challenge, working with both traditional infrastructure such as load balancers, NLB, etc., as well as very different infrastructure around inference engines and GPU loads.

This is a role that will inherently require deep experience with inference engines.

Contributions to vLLM, SGLang, trtllm, or inference frameworks a plus.

Every role at Morph comes with unlimited tokens on claude code/codex

More at Morph

Related open roles

View all roles