Padmi

Principal Voice AI Scientist

HyderabadPosted 3 months ago
Computer ResearchStaff+Full Time; Regular
Apply at YAL

Opens the source posting on shine.com

Source description

About the role

View original

In this role as a Senior Speech Scientist at YAL, your primary responsibility will be to design, build, and enhance speech and language models to improve product experiences. You will tackle complex challenges across the speech processing stack, collaborate with a skilled team, and drive meaningful impact through thorough research and engineering efforts. Responsibilities: - Research, develop, and optimize models for speech translation, ASR, TTS, voice activity detection, speaker identification, speech enhancement, or related areas - Design and conduct experiments, analyze outcomes, and iterate swiftly to enhance system performance - Establish robust data pipelines for training, evaluation, and deployment of speech models - Collaborate with engineering teams to integrate speech models into production systems, focusing on latency, accuracy, and scalability - Stay updated with the latest research, evaluate new techniques, and suggest their adoption where suitable - Contribute to technical documentation, internal knowledge sharing, and peer code reviews - Engage in the hiring process and provide guidance to junior team members Qualifications: - Master's degree in Speech Processing, Electrical Engineering, Computer Science, Computational Linguistics, or a related field - 4+ years of practical experience in developing speech or audio ML models in an industry environment - Profound understanding of signal processing basics, acoustic modeling, and language modeling - Proficiency in Python and at least one deep learning framework (preferably PyTorch) - Experience in training and fine-tuning large neural network architectures (Transformers, Conformers, RNN-T, etc.) - Strong software engineering skills and proficiency in working with production codebases - Advanced analytical and problem-solving abilities with a focus on experimental rigor - Effective written and verbal communication skills Nice to Have: - Publications in leading speech or ML conferences - Familiarity with streaming/real-time speech processing systems - Knowledge of model compression, quantization, and on-device deployment - Experience working with multilingual or low-resource speech data Join YAL in crafting a future where privacy-first communication, intelligent discovery, and cutting-edge AI innovation converge under the vision "Where AI Meets Integrity". In this role as a Senior Speech Scientist at YAL, your primary responsibility will be to design, build, and enhance speech and language models to improve product experiences. You will tackle complex challenges across the speech processing stack, collaborate with a skilled team, and drive meaningful impact through thorough research and engineering efforts. Responsibilities: - Research, develop, and optimize models for speech translation, ASR, TTS, voice activity detection, speaker identification, speech enhancement, or related areas - Design and conduct experiments, analyze outcomes, and iterate swiftly to enhance system performance - Establish robust data pipelines for training, evaluation, and deployment of speech models - Collaborate with engineering teams to integrate speech models into production systems, focusing on latency, accuracy, and scalability - Stay updated with the latest research, evaluate new techniques, and suggest their adoption where suitable - Contribute to technical documentation, internal knowledge sharing, and peer code reviews - Engage in the hiring process and provide guidance to junior team members Qualifications: - Master's degree in Speech Processing, Electrical Engineering, Computer Science, Computational Linguistics, or a related field - 4+ years of practical experience in developing speech or audio ML models in an industry environment - Profound understanding of signal processing basics, acoustic modeling, and language modeling - Proficiency in Python and at least one deep learning framework (preferably PyTorch) - Experience in training and fine-tuning large neural network architectures (Transformers, Conformers, RNN-T, etc.) - Strong software engineering skills and proficiency in working with production codebases - Advanced analytical and problem-solving abilities with a focus on experimental rigor - Effective written and verbal communication skills Nice to Have: - Publications in leading speech or ML conferences - Familiarity with streaming/real-time speech processing systems - Knowledge of model compression, quantization, and on-device deployment - Experience working with multilingual or low-resource speech data Join YAL in crafting a future where privacy-first communication, intelligent discovery, and cutting-edge AI innovation converge under the vision "Where AI Meets Integrity".

One address, no account. We’ll tell you when matching roles go live.

More at YAL

Related open roles

View all roles