Deep Learning Scientist, Speech Synthesis

Vailexa
  • Santa Clara, California
    17 days ago

    Job Description

    Build More Than Just a Career. Build Your Future.

    At Vailexa, we’re not just hiring — we’re building thinkers, creators, and future leaders.

    We believe in giving people the space to grow, the freedom to think, and the opportunity to create real impact from day one. If you’re someone who wants to learn fast, take ownership, and grow beyond limits, you’ll feel right at home here.

    Key Responsibilities

    • Train speech synthesis mel spectrogram and vocoder models
    • Measure and benchmark model performance across use cases
    • Maintain and enhance text to speech evaluation systems
    • Analyze model accuracy and bias and recommend improvements
    • Improve processes related to speech data preparation, augmentation, and filtering
    • Develop and refine training datasets for speech models
    • Characterize performance and quality metrics across different platforms
    • Collaborate with cross functional teams to deliver new product features
    • Participate in code development, design reviews, and test planning
    • Identify issues, propose solutions, and contribute to continuous innovation

    Required Qualifications

    • Master’s degree or PhD in Computer Science, Electrical Engineering, Artificial Intelligence, Applied Mathematics, Linguistics, or Computational Linguistics or equivalent experience
    • Minimum of 5 years of relevant experience
    • Strong programming skills in Python
    • Solid understanding of programming fundamentals and software design
    • Deep knowledge of machine learning and deep learning techniques including CNN, RNN, LSTM, and Transformers
    • Experience applying deep learning to speech synthesis, large language models, and speech to speech translation
    • Hands on experience with speech technologies such as speech synthesis and voice cloning
    • Experience training speech models
    • Proficiency with PyTorch deep learning frameworks
    • Knowledge of speech signal processing techniques including FFT, MFCC, and mel spectrograms
    • Familiarity with version control tools such as Git, Gerrit, or GitLab
    • Strong collaboration and communication skills in a matrixed environment

    Preferred Qualifications

    • Fluency in one or more languages such as Spanish, Mandarin, German, Japanese, Russian, French, Arabic, Hindi, Korean, Italian, or Portuguese
    • Experience with multilingual or code switched text to speech systems
    • Experience with voice cloning and cross lingual voice cloning
    • Knowledge of text normalization and inverse text normalization using neural networks or WFST
    • Experience working with grapheme to phoneme systems for multiple languages
    • Interest in linguistics, phonetics, and language technologies
    • Strong C plus plus programming skills
    • Familiarity with GPU technologies such as CUDA, cuDNN, or TensorRT
    • Experience deploying machine learning models to cloud, data center, or embedded systems

    Ready to take the next step?

    If you’re excited about this role and ready to grow with a team that values ambition, ideas, and impact — we’d love to hear from you.

    Apply now and start building your journey with Vailexa.

    Numbers & Facts

    LocationSanta Clara, California

    Skills

    • Analysis Skillsunmatched
    • Arabic Languageunmatched
    • Artificial Intelligence (AI)unmatched
    • Benchmarkingunmatched
    • CUDA (Compute Unified Device Architecture)unmatched
    • Cloningunmatched
    • Cloud Computingunmatched
    • Communication Skillsunmatched
    • Computational Linguisticsunmatched
    • Computer Programmingunmatched
    • Computer Scienceunmatched
    • Cross-Functionalunmatched
    • Deep Learningunmatched
    • Electrical Engineeringunmatched
    • Embedded Systemsunmatched
    • Fast Fourier Transformunmatched
    • French Languageunmatched
    • GPU (Graphics Processing Unit)unmatched
    • German Languageunmatched
    • Gerritunmatched
    • Gitunmatched
    • Identify Issuesunmatched
    • Italian Languageunmatched
    • Japanese Languageunmatched
    • Korean Languageunmatched
    • Linguisticsunmatched
    • Machine Learningunmatched
    • Mandarin Chinese Languageunmatched
    • Mathematicsunmatched
    • Modeling Languagesunmatched
    • Multilingualunmatched
    • Multiplatform/Cross-Platformunmatched
    • Network Operations Centerunmatched
    • Neural Networksunmatched
    • Performance Metricsunmatched
    • Performance Modelingunmatched
    • Portuguese Languageunmatched
    • Process Improvementunmatched
    • Python Programming/Scripting Languageunmatched
    • Quality Metricsunmatched
    • Russian Languageunmatched
    • Signal Processingunmatched
    • Software Designunmatched
    • Software Engineeringunmatched
    • Source Code/Configuration Management (SCM)unmatched
    • Spanish Languageunmatched
    • Speech Synthesisunmatched
    • Speech Technologyunmatched
    • Team Playerunmatched
    • Test Designunmatched
    • Training Data Setsunmatched
    • Use Casesunmatched
    • Voice Applicationsunmatched

    Be found by employers

    5,500+ employers search our resume database daily. Add yours to get found by recruiters looking for candidates like you.

    Level up your application

    Professional resume templates

    Browse dozens of recruiter approved resume templates, layouts and formats. Choose your favorite and make it your own in minutes.

    Free resume templates

    Free resume builder

    Improve your existing resume or start from scratch and create a standout, ATS-friendly resume. Add job-specific content, download and apply.

    Free resume builder