Padmi

AI/ML Engineer Generative AI (LLMs + Multimodal)

Delhi NCRPosted 3 months ago
Software engineeringJuniorFull Time; Regular
Apply at 4 Way Technologies

Opens the source posting on shine.com

Source description

About the role

View original

You will be joining 4 Way Technologies, an IT Company known for developing AI Applications using cutting-edge technologies. The company has its own product - Chatreal.ai, an AI companion App that is building next-gen AI companions powered by large language models and real-time image generation. Role Overview: As an AI Engineer at Chatreal.ai, you will be responsible for fine-tuning and optimizing LLMs (DeepSeek + in-house models), building RAG pipelines & vector search systems, improving response quality, personality consistency & memory systems, designing evaluation pipelines, building ComfyUI-based image generation workflows, integrating Stable Diffusion, LoRA, ControlNet, inpainting/outpainting, optimizing GPU usage, VRAM efficiency & inference latency, and scaling AI systems for real-world production traffic. Key Responsibilities: - Fine-tuning and optimizing LLMs (DeepSeek + in-house models) - Building RAG pipelines & vector search systems - Improving response quality, personality consistency & memory systems - Designing evaluation pipelines (quality, hallucination, engagement) - Building ComfyUI-based image generation workflows - Integrating Stable Diffusion, LoRA, ControlNet, inpainting/outpainting - Optimizing GPU usage, VRAM efficiency & inference latency - Scaling AI systems for real-world production traffic Qualifications Required: - 2+ years hands-on AI/ML experience - Strong Python + PyTorch skills - Experience fine-tuning LLMs (PEFT / LoRA / quantization) - Experience with RAG + vector databases - Hands-on experience with diffusion pipelines (Stable Diffusion preferred) - Production mindset with skills in monitoring, versioning, and cost optimization - Comfortable working in a fast-paced startup environment In addition, experience with AI companion or conversational AI products, ComfyUI workflows, GPU infra optimization, and real-time streaming generation systems would be considered a bonus. Joining ChatReal will offer you the opportunity to work on cutting-edge multimodal AI projects, have high ownership, fast iteration, direct impact on architecture and roadmap, and work closely with product & engineering leadership. You will be joining 4 Way Technologies, an IT Company known for developing AI Applications using cutting-edge technologies. The company has its own product - Chatreal.ai, an AI companion App that is building next-gen AI companions powered by large language models and real-time image generation. Role Overview: As an AI Engineer at Chatreal.ai, you will be responsible for fine-tuning and optimizing LLMs (DeepSeek + in-house models), building RAG pipelines & vector search systems, improving response quality, personality consistency & memory systems, designing evaluation pipelines, building ComfyUI-based image generation workflows, integrating Stable Diffusion, LoRA, ControlNet, inpainting/outpainting, optimizing GPU usage, VRAM efficiency & inference latency, and scaling AI systems for real-world production traffic. Key Responsibilities: - Fine-tuning and optimizing LLMs (DeepSeek + in-house models) - Building RAG pipelines & vector search systems - Improving response quality, personality consistency & memory systems - Designing evaluation pipelines (quality, hallucination, engagement) - Building ComfyUI-based image generation workflows - Integrating Stable Diffusion, LoRA, ControlNet, inpainting/outpainting - Optimizing GPU usage, VRAM efficiency & inference latency - Scaling AI systems for real-world production traffic Qualifications Required: - 2+ years hands-on AI/ML experience - Strong Python + PyTorch skills - Experience fine-tuning LLMs (PEFT / LoRA / quantization) - Experience with RAG + vector databases - Hands-on experience with diffusion pipelines (Stable Diffusion preferred) - Production mindset with skills in monitoring, versioning, and cost optimization - Comfortable working in a fast-paced startup environment In addition, experience with AI companion or conversational AI products, ComfyUI workflows, GPU infra optimization, and real-time streaming generation systems would be considered a bonus. Joining ChatReal will offer you the opportunity to work on cutting-edge multimodal AI projects, have high ownership, fast iteration, direct impact on architecture and roadmap, and work closely with product & engineering leadership.

One address, no account. We’ll tell you when matching roles go live.

More at 4 Way Technologies

Related open roles

View all roles