Source description
About the role
Design, develop, and deploy production-ready AI/ML solutions for enterprise applications Build and optimize Large Language Model (LLM) applications, AI agents, and Retrieval-Augmented Generation (RAG) pipelines Fine-tune, evaluate, and optimize open-source and commercial AI models for specific business use cases Develop scalable APIs and AI services using Python frameworks such as FastAPI or Flask Build intelligent automation workflows using AI, vector databases, and orchestration frameworks Design and implement machine learning pipelines for data preprocessing, feature engineering, model training, validation, and deployment Collaborate with Product Managers, Architects, UI/UX Designers, and DevOps teams to deliver AI-powered products Optimize model inference performance, latency, and infrastructure costs Mentor junior AI engineers and participate in architecture discussions, technical reviews, and code reviews Research and evaluate emerging AI technologies and recommend adoption where appropriate Ensure AI applications follow security, privacy, and responsible AI best practices Requirements 4+ years in Artificial Intelligence, Machine Learning, and production software development with at least 2 years of hands-on experience in Generative AI and LLM-based applications Strong experience with Machine Learning, Deep Learning, NLP, and Generative AI Hands-on experience with LLMs such as GPT, Claude, Llama, Qwen, Mistral, or Gemma Experience developing AI Agents and multi-agent workflows Strong understanding of prompt engineering, structured outputs, function calling, and tool integration Experience implementing Retrieval-Augmented Generation (RAG) architectures Knowledge of embedding models and semantic search Disclaimer : This job posting has been aggregated from external source. Role details, content, and availability are subject to change. Applicants are advised to confirm the latest information directly on the company website before applying.
More at Moon Technolabs