Lead hands-on research and development across ASR, MT, TTS, and speech-to-speech translation for real-time voice products
Design, train, and optimize large-scale ASR models for multilingual accuracy, robustness, and ultra-low-latency streaming
Improve cascaded translation pipelines end to end: segmentation, ASR→MT interfaces, streaming MT inference, and incremental decoding
Develop and refine real-time TTS models with natural prosody, stable speaker characteristics, and fast inference
Build and experiment with end-to-end and LLM-based speech-to-speech translation systems, including streaming and one-shot approaches
Own the full lifecycle of model delivery: prototyping, ablations, training, evaluation, optimization, and production deployment
Work closely with engineering teams to integrate models into real-time systems, ensuring reliability, uptime, and quality at scale
Drive improvements in inference efficiency, model serving, voice UX, and robustness to real-world acoustic conditions
Establish strong practices for evaluation, reproducibility, monitoring, and continuous model improvement in production
Mentor researchers and engineers, promote hands-on collaboration, and raise the bar for model quality and operational excellence