Source description
About the role
What you ll do Design, build and train robust deep learning speech systems that power Agara s realtime conversations across languages, accents, noise environments, and voice interfaces. You will work closely with a team of machine learning scientists and engineers to train and produce speech models. Applied research and experimentation will be an integral part of your role. We require that you have B.Tech, M.Tech or Ph.D. in Computer Science or equivalent work experience Relevant work experience of 6 years Enthusiastic to work in a fast-paced environment, along with machine learning scientists and engineers Comfortable with at least one deep learning framework (PyTorch / Tensorflow) Previous extensive experience building real-world speech systems such as speech recognition, text to speech, speaker diarization Passionate about innovation and pushing the state of art in speech Excellent programming skills in Python Bonus if you have Publications in top-tier journals/conferences such as ICASSP, NeurIPS, PAMI, InterSpeech. Previous experience building voice conversational AI systems Experience working with textual NLP systems
More at Agara Labs