Padmi

Artificial Intelligence Engineer

BangalorePosted 30 days ago
Software engineeringSeniorFull Time
Apply at leadsoc technologies pvt ltd

Opens the source posting on foundit.in

Source description

About the role

View original

Job Title: AI Engineer – Model Optimization & Acceleration Location: Bangalore, India Role Overview Seeking an AI Engineer to optimize and deploy ML models across heterogeneous platforms (CPU, GPU, NPU). Work on scalable, production-ready AI systems across domains like robotics, healthcare, and automotive. Key Responsibilities Optimize diverse models: generative (LLMs, diffusion), vision (classification, detection, segmentation), multi-modal, and speech Port models across frameworks (e.g., PyTorch → ONNX → runtimes) Deploy on hardware accelerators (GPU/NPU) and optimize performance Improve inference latency, throughput, and memory (batching, caching, parallelism, fusion) Apply quantization and model compression (FP32 → lower precision) Profile and debug system and model performance Required Skills Strong in PyTorch (or similar), ONNX (or equivalent) Proficient in Python and C++ Experience with GPU/hardware acceleration (CUDA/ROCm or similar) Solid understanding of deep learning models (transformers, CNNs) Knowledge of optimization, quantization, and performance tuning Good to Have Edge AI or embedded deployment Generative or multi-modal AI systems Distributed inference or streaming pipelines Experience 7–10 years

One address, no account. We’ll tell you when matching roles go live.

More at leadsoc technologies pvt ltd

Related open roles

View all roles