Vailexa

Vailexa

Deep Learning Scientist, Speech Synthesis

Santa Clara, California

Sponsorship not specifiedDetected 9 hours ago
PythonGitMachine LearningDeep LearningPyTorchNLPLLMsEmbedded SystemsElectrical EngineeringSignal ProcessingCommunicationCollaboration

About the role

  • If you're someone who wants to learn fast, take ownership, and grow beyond limits, you'll feel right at home here.

Responsibilities

  • Maintain and enhance text to speech evaluation systems
  • Develop and refine training datasets for speech models
  • Collaborate with cross functional teams to deliver new product features
  • Participate in code development, design reviews, and test planning

Requirements

  • Master's degree or PhD in Computer Science, Electrical Engineering, Artificial Intelligence, Applied Mathematics, Linguistics, or Computational Linguistics or equivalent experience
  • Minimum of 5 years of relevant experience
  • Hands on experience with speech technologies such as speech synthesis and voice cloning
  • Experience training speech models
  • Knowledge of speech signal processing techniques including FFT, MFCC, and mel spectrograms
  • Familiarity with version control tools such as Git, Gerrit, or GitLab

Nice to have

  • Fluency in one or more languages such as Spanish, Mandarin, German, Japanese, Russian, French, Arabic, Hindi, Korean, Italian, or Portuguese
  • Experience with multilingual or code switched text to speech systems
  • Experience with voice cloning and cross lingual voice cloning
  • Knowledge of text normalization and inverse text normalization using neural networks or WFST
  • Experience working with grapheme to phoneme systems for multiple languages
  • Interest in linguistics, phonetics, and language technologies
  • Strong C plus plus programming skills
  • Familiarity with GPU technologies such as CUDA, cuDNN, or TensorRT

This listing is sourced directly from Vailexa's careers page and normalized into a canonical job model.