Vailexa
Deep Learning Scientist, Speech Synthesis
Santa Clara, California
Sponsorship not specifiedDetected 9 hours ago
PythonGitMachine LearningDeep LearningPyTorchNLPLLMsEmbedded SystemsElectrical EngineeringSignal ProcessingCommunicationCollaboration
About the role
- If you're someone who wants to learn fast, take ownership, and grow beyond limits, you'll feel right at home here.
Responsibilities
- Maintain and enhance text to speech evaluation systems
- Develop and refine training datasets for speech models
- Collaborate with cross functional teams to deliver new product features
- Participate in code development, design reviews, and test planning
Requirements
- Master's degree or PhD in Computer Science, Electrical Engineering, Artificial Intelligence, Applied Mathematics, Linguistics, or Computational Linguistics or equivalent experience
- Minimum of 5 years of relevant experience
- Hands on experience with speech technologies such as speech synthesis and voice cloning
- Experience training speech models
- Knowledge of speech signal processing techniques including FFT, MFCC, and mel spectrograms
- Familiarity with version control tools such as Git, Gerrit, or GitLab
Nice to have
- Fluency in one or more languages such as Spanish, Mandarin, German, Japanese, Russian, French, Arabic, Hindi, Korean, Italian, or Portuguese
- Experience with multilingual or code switched text to speech systems
- Experience with voice cloning and cross lingual voice cloning
- Knowledge of text normalization and inverse text normalization using neural networks or WFST
- Experience working with grapheme to phoneme systems for multiple languages
- Interest in linguistics, phonetics, and language technologies
- Strong C plus plus programming skills
- Familiarity with GPU technologies such as CUDA, cuDNN, or TensorRT
Apply directly at Vailexa →Create a free account for alerts like thisView Vailexa immigration profile
This listing is sourced directly from Vailexa's careers page and normalized into a canonical job model.