Source description
About the role
Role description: This role is with one of our prominent portfolio companies. About Us A San Francisco based startup building next-generation Voice AI products that redefine how humans interact with machines from smart voice assistants to automated customer conversations, voice-driven tools, and more. We are at the intersection of speech technology, large language models, and real-time systems, backed by leading investors and supported by domain experts. We re now building our founding engineering team in India to shape the core product experience. What Youll Do Build and deploy real-time voice-based AI applications using ASR (Automatic Speech Recognition), TTS (Text-to-Speech), and LLMs. Work on latency-sensitive systems to enable near real-time conversations. Design and implement prompt-chaining, memory, and tool integration for LLM-powered voice agents. Set up and manage scalable infra for voice/audio processing and AI model serving. Work closely with the founding team on product shaping, roadmap planning, and technical strategy. Continuously experiment with and evaluate new models, APIs, and speech/LLM techniques. Who You Are 3-8 years of experience in AI/ML, deep learning, or backend-heavy engineering roles. Solid hands-on experience with speech technologies (ASR, TTS, diarization, etc.). Comfortable working with Python, PyTorch, Hugging Face, OpenAI, or similar frameworks. Experience deploying real-time systems (Docker, Kubernetes, AWS/GCP). Strong problem-solving skills with a product-first mindset. Self-starter who thrives in high-ownership, fast-paced environments. Excellent written and verbal communication
More at Elevation Capital