Source description
About the role
As an AI + Backend Engineer at our company in Gurugram, you will be responsible for designing, developing, and optimizing Large Language Models (LLMs) based AI systems for text generation, information retrieval, and agent-based applications. You will work on fine-tuning and implementing LLM models for tasks like entity recognition, summarization, and question-answering. It will be your duty to architect and develop backend services and APIs that integrate AI models with scalable infrastructure. Moreover, you will collaborate with ML researchers and engineers to enhance model efficiency and deployment strategies. Your role will also involve monitoring, troubleshooting, and optimizing model and agent inference performance in real-world applications, ensuring robust data pipelines for preprocessing, model training, and inference workflows. Key Responsibilities: - Design, develop, and optimize LLM-based AI systems for various applications. - Fine-tune and implement LLM models for tasks like entity recognition and summarization. - Architect and develop backend services and APIs integrating AI models. - Collaborate with ML researchers to enhance model efficiency and deployment strategies. - Monitor and optimize model and agent inference performance in real-world applications. - Ensure robust data pipelines for preprocessing, model training, and inference workflows. Qualifications Required: - 7-8+ years of experience in AI/ML and backend development, with a focus on NLP and LLMs. - Proficiency in Python, with experience in libraries like PyTorch, TensorFlow, Hugging Face Transformers. - Expertise in LLM fine-tuning, prompt engineering, and agent-based frameworks. - Experience with vector databases and embeddings-based retrieval. - Strong backend development skills using FastAPI, Flask, or Django. - Experience in designing and deploying scalable AI architectures in cloud environments. - Knowledge of containerization and model deployment techniques. - Familiarity with CI/CD pipelines, model versioning, and MLOps best practices. - Excellent problem-solving skills, logical thinking, and the ability to work in an agile environment. - Strong communication and collaboration skills to work across different teams. Please note that the job description did not contain any additional details about the company. As an AI + Backend Engineer at our company in Gurugram, you will be responsible for designing, developing, and optimizing Large Language Models (LLMs) based AI systems for text generation, information retrieval, and agent-based applications. You will work on fine-tuning and implementing LLM models for tasks like entity recognition, summarization, and question-answering. It will be your duty to architect and develop backend services and APIs that integrate AI models with scalable infrastructure. Moreover, you will collaborate with ML researchers and engineers to enhance model efficiency and deployment strategies. Your role will also involve monitoring, troubleshooting, and optimizing model and agent inference performance in real-world applications, ensuring robust data pipelines for preprocessing, model training, and inference workflows. Key Responsibilities: - Design, develop, and optimize LLM-based AI systems for various applications. - Fine-tune and implement LLM models for tasks like entity recognition and summarization. - Architect and develop backend services and APIs integrating AI models. - Collaborate with ML researchers to enhance model efficiency and deployment strategies. - Monitor and optimize model and agent inference performance in real-world applications. - Ensure robust data pipelines for preprocessing, model training, and inference workflows. Qualifications Required: - 7-8+ years of experience in AI/ML and backend development, with a focus on NLP and LLMs. - Proficiency in Python, with experience in libraries like PyTorch, TensorFlow, Hugging Face Transformers. - Expertise in LLM fine-tuning, prompt engineering, and agent-based frameworks. - Experience with vector databases and embeddings-based retrieval. - Strong backend development skills using FastAPI, Flask, or Django. - Experience in designing and deploying scalable AI architectures in cloud environments. - Knowledge of containerization and model deployment techniques. - Familiarity with CI/CD pipelines, model versioning, and MLOps best practices. - Excellent problem-solving skills, logical thinking, and the ability to work in an agile environment. - Strong communication and collaboration skills to work across different teams. Please note that the job description did not contain any additional details about the company.
More at StatusNeo Technology Consulting Pvt. Ltd