Bland AI
Machine Learning Researcher, Audio
San Francisco
Sponsorship not specified$160k-$250kDetected 92 days ago
Machine LearningLLMsResearch
About the role
- As a Machine Learning Researcher at Bland, you'll be working on foundational research and development across the core components of our voice stack: speech-to-text, large language models, neural audio codecs, and text-to-speech.
- Your work will define how our agents understand, reason, and speak in real time at enterprise scale.
- You will take ideas from theory to large-scale training to production inference systems serving millions of calls per day.
Responsibilities
- Build and Scale Next-Generation TTS Systems
- Design and train large scale text-to-speech models capable of expressive, controllable, human-sounding output.
- Develop neural audio codec-based TTS architectures for efficient, high-fidelity generation.
- Optimize for real-time, low-latency inference in production.
- Build and fine-tune large scale ASR systems robust to accents, noise, telephony artifacts, and code switching.
- Research and implement neural audio codecs that achieve extreme compression with minimal perceptual loss.
- Design codec architectures that enable downstream generative modeling and controllable synthesis.
- Develop Scalable Training Pipelines
Compensation
- Every tool you need to succeed
- Beautiful office in Jackson Square, SF with rooftop views
- If you are energized by building and scaling TTS models, pioneering neural audio codecs, and pushing the boundaries of speech-to-text systems, we would love to hear from you.
- $160,000 to $250,000
Benefits
- Healthcare, dental, vision, all the good stuff
- Meaningful equity in a fast-growing company
Apply directly at Bland AI →Create a free account for alerts like thisView Bland AI immigration profile
This listing is sourced directly from Bland AI's careers page and normalized into a canonical job model.