Source description
About the role
Build inference APIs,microservices,and CI / CD pipelines for AI applications. Monitor model quality,latency,cost,drift,and hallucinations in production. Fine-tune and enhance models using parameter-efficient techniques where required. Optimize inference performance using caching,batching,quantization,and prompt optimization. Bachelors or masters degree in relevant stream
More at HCLTech