Padmi

AI/ML Engineer Generative AI (LLMs + Multimodal) Delhi NCR Only

Delhi NCRPosted 2 months ago
Software engineeringJuniorFull Time; Regular
Apply at 4 Way Technologies

Opens the source posting on shine.com

Source description

About the role

View original

About Company 4 Way Technologies is an IT Company well known for developing AI Applications using cutting edge technologies. We have our own product - Chatreal.ai, which is an AI companion App. About Chatreal.AI ChatReal is building next-gen AI companions powered by large language models and real-time image generation. Were looking for a hands-on AI Engineer who can take models from research to production at scale. If youve shipped LLM systems, optimized inference costs, and worked with diffusion/image pipelines we want to talk. What Youll Work On Fine-tuning and optimizing LLMs (DeepSeek + in-house models) Building RAG pipelines & vector search systems Improving response quality, personality consistency & memory systems Designing evaluation pipelines (quality, hallucination, engagement) Building ComfyUI-based image generation workflows Integrating Stable Diffusion, LoRA, ControlNet, inpainting/outpainting Optimizing GPU usage, VRAM efficiency & inference latency Scaling AI systems for real-world production traffic What Were Looking For 2+ years hands-on AI/ML experience Strong Python + PyTorch Experience fine-tuning LLMs (PEFT / LoRA / quantization) Experience with RAG + vector databases Hands-on with diffusion pipelines (Stable Diffusion preferred) Production mindset monitoring, versioning, cost optimization Comfortable in a fast-paced startup environment Bonus Experience with AI companion or conversational AI products Experience with ComfyUI workflows GPU infra optimization experience Real-time streaming generation systems Why ChatReal Build cutting-edge multimodal AI High ownership, fast iteration Direct impact on architecture and roadmap Work closely with product & engineering leadership About Company 4 Way Technologies is an IT Company well known for developing AI Applications using cutting edge technologies. We have our own product - Chatreal.ai, which is an AI companion App. About Chatreal.AI ChatReal is building next-gen AI companions powered by large language models and real-time image generation. Were looking for a hands-on AI Engineer who can take models from research to production at scale. If youve shipped LLM systems, optimized inference costs, and worked with diffusion/image pipelines we want to talk. What Youll Work On Fine-tuning and optimizing LLMs (DeepSeek + in-house models) Building RAG pipelines & vector search systems Improving response quality, personality consistency & memory systems Designing evaluation pipelines (quality, hallucination, engagement) Building ComfyUI-based image generation workflows Integrating Stable Diffusion, LoRA, ControlNet, inpainting/outpainting Optimizing GPU usage, VRAM efficiency & inference latency Scaling AI systems for real-world production traffic What Were Looking For 2+ years hands-on AI/ML experience Strong Python + PyTorch Experience fine-tuning LLMs (PEFT / LoRA / quantization) Experience with RAG + vector databases Hands-on with diffusion pipelines (Stable Diffusion preferred) Production mindset monitoring, versioning, cost optimization Comfortable in a fast-paced startup environment Bonus Experience with AI companion or conversational AI products Experience with ComfyUI workflows GPU infra optimization experience Real-time streaming generation systems Why ChatReal Build cutting-edge multimodal AI High ownership, fast iteration Direct impact on architecture and roadmap Work closely with product & engineering leadership

One address, no account. We’ll tell you when matching roles go live.

More at 4 Way Technologies

Related open roles

View all roles