Source description
About the role
Job Title: MLOps / Cloud Engineer Experience: 4-7 Years Location: Hyderabad Employment Type: Full-Time Overview: Responsible for deployment, CI/CD, scaling, GPU cluster orchestration, and monitoring. Responsibilities: Deploy GPU workloads using Kubernetes Configure GPU scheduling Deploy Triton Inference Server and NVIDIA TAO toolkit Implement CI/CD for ML models Set up monitoring: GPU utilization Model latency Throughput Build rollback and A/B testing pipelines Automate model retraining pipelines Skills: Kubernetes Docker GPU scheduling Helm CI/CD tools Preferred: Experience managing NVIDIA GPU clusters Experience in distributed inference systems Compensation: Up to 1,500,000.00 per year Advantages: - Provident Fund Work Location: In person .
More at VirtualMaze Softsys