Bland AI

Bland AI

Machine Learning Researcher, Audio

San Francisco

Sponsorship not specified$160k-$250kDetected 92 days ago
Machine LearningLLMsResearch

About the role

  • As a Machine Learning Researcher at Bland, you'll be working on foundational research and development across the core components of our voice stack: speech-to-text, large language models, neural audio codecs, and text-to-speech.
  • Your work will define how our agents understand, reason, and speak in real time at enterprise scale.
  • You will take ideas from theory to large-scale training to production inference systems serving millions of calls per day.

Responsibilities

  • Build and Scale Next-Generation TTS Systems
  • Design and train large scale text-to-speech models capable of expressive, controllable, human-sounding output.
  • Develop neural audio codec-based TTS architectures for efficient, high-fidelity generation.
  • Optimize for real-time, low-latency inference in production.
  • Build and fine-tune large scale ASR systems robust to accents, noise, telephony artifacts, and code switching.
  • Research and implement neural audio codecs that achieve extreme compression with minimal perceptual loss.
  • Design codec architectures that enable downstream generative modeling and controllable synthesis.
  • Develop Scalable Training Pipelines

Compensation

  • Every tool you need to succeed
  • Beautiful office in Jackson Square, SF with rooftop views
  • If you are energized by building and scaling TTS models, pioneering neural audio codecs, and pushing the boundaries of speech-to-text systems, we would love to hear from you.
  • $160,000 to $250,000

Benefits

  • Healthcare, dental, vision, all the good stuff
  • Meaningful equity in a fast-growing company

This listing is sourced directly from Bland AI's careers page and normalized into a canonical job model.