M
BengaluruApplied SciencesTop payGCC
Apply on Microsoft →Research Microsoft before you apply
Check ratings, real-employee reviews, verified pay, and interview difficulty.
Your work will impact millions—enabling next-generation human-machine experiences for diverse markets, with a keen focus on India. In this strategic role, you will have significant ownership on execution, technical directions and drive innovation in speech recognition, AOAI customisation, and generative AI. You'll collaborate with top scientists and engineers and influence cross-functional teams to scale model quality, and deliver breakthrough technologies for voice agents, video translation, and call centre analytics. Based in Hyderabad/Bengaluru, this on-site role offers opportunities to grow, and shape the future of multimodal interaction for Indian and global audiences. Drive execution and set technical directions in multilingual speech model, speech LLMs, model customization and impact accuracy, latency, and compute. Build novel data engineering solutions to synthesize complex speech scenarios and finetune models. Build data analysis metrics and solutions to understand the model results, identify gaps, and guide solutions. Mentor and influence peers, sharing expertise and fostering a growth-oriented inclusive team culture. BS/MS/PhD Degree in CS/EE or related fields with strong focus in speech recognition systems, machine learning, and AI technology innovations. 6+ years of experience in speech or machine learning in academic or industrial setting, or 6+ years' experience in software development skills and aptitude for software design, coding and quality. Demonstration of excellent problem-solving skills in speech and machine learning areas. Proven track record of delivering impactful results and high-quality solutions in complex technical environments. Strong programming skills in Python, C++ or similar languages, with experience in large-scale data processing and distributed computing. Effective communication skills, both verbal and written. Experience with speech/audio processing, multilingual model development, or voice agent technologies. Familiarity with Azure, cloud-based AI platforms, or enterprise-scale deployment of speech solutions. Contributions to open-source projects, patents, or publications in top-tier conferences/journals. Demonstrated leadership in driving technical direction, influencing cross-functional teams, and mentoring peers.