Source description
About the role
We are looking for a Senior Engineer Serverless Inference to design, build, and optimize scalable AI inference platforms. The ideal candidate should have experience with distributed systems, cloud-native architectures, Kubernetes, and Go/Golang, with a focus on building reliable, high-performance backend services. Key Responsibilities Design and develop scalable serverless inference and backend services. Build and maintain distributed systems with high availability and reliability. Collaborate with platform, infrastructure, and product teams to deliver production-ready AI solutions. Improve system performance through observability, automation, and capacity planning. Contribute to architecture, reliability, and operational excellence. Required Skills 5+ years of experience in backend or distributed systems engineering. Strong hands-on experience with Go/Golang and Kubernetes . Experience with cloud-native architectures , microservices , and distributed systems . Knowledge of SRE practices , observability, incident management, and performance optimization. Familiarity with AI/ML inference technologies, GPU infrastructure, or LLM serving frameworks (vLLM, Triton, TensorRT-LLM) is an added advantage.
More at Hiringeye Solutions