Research Scientist Graduate (Foundation Model-Speech-Interaction

Beijing ByteDance Technology Co Ltd
  • San Jose, CA
    24 days ago

    Job Description

    About the Team Established in 2023, the ByteDance Seed team is dedicated to pioneering new paths toward artificial general intelligence. We aspire to advance the frontier of intelligence to drive progress for both technology and society. With a long-term vision for the AI sector, the Seed team's research spans MLLM, GenMedia, AI for Science, and Robotics. We maintain a global presence with laboratories and career opportunities across China, Singapore, and the United States. To date, we have launched industry-leading general foundation models and cutting-edge multimodal capabilities. Our technology powers over 50 application scenarios - including Doubao, Jimeng, TRAE, Dola and Dreamnia - and serves enterprise customers through Volcano Engine and BytePlus. Third-party data shows that the Doubao App ranks first in user volume in the Chinese market, while Doubao foundation models lead the industry in average daily token consumption.

    The mission of the Seed Speech team is to enrich interactive and creative processes through the application of multimodal speech technologies. The team focuses on the forefront of research and product development in speech and audio, music, natural language understanding, and multimodal deep learning.

    We are looking for talented individuals to join our team. As a graduate, you will get opportunities to pursue bold ideas, tackle complex challenges, and unlock limitless growth. Successful candidates must be able to commit to an onboarding date by the end of the year. Please state your availability and graduation date clearly in your resume.

    Responsibilities

    • Contribute cutting-edge research to ByteDance product evolution (e.g., Douyin, Capcut, and more) to impact billions of users worldwide.
    • Work on advanced science and technology in audio processing and generation (e.g., Dialogue Systems, Audio-Video Models, Speech Synthesis, Voice Conversion, Audio Codec Learning, Audio Language Modeling, etc.)
    • Research, model, design, develop and evaluate novel machine learning models and algorithms.
    • Collaborate with globally based researchers and engineering teams in developing machine learning models and algorithms.Minimum Qualifications
    • Individuals who are completing or have recently completed a PhD degree in Computer Science, Electrical Engineering, Electrical and Computer Engineering, Physics, Mathematics, or a related discipline.
    • Good knowledge of theoretical and empirical research in addressing research problems
    • Solid knowledge and experience with at least one popular deep learning framework (e.g., PyTorch, TensorFlow) and familiarity with deep neural network architectures
    • Experience in both neural and non-neural, classical machine learning models and algorithms

    Preferred Qualifications

    • Research experience in one or more of the following fields: speech synthesis, audio generation, large language model, computer vision, generative models
    • Strong first-author publications at accredited conferences such as(e.g., NeurIPS, ICML, ICLR, ACL, EMNLP, NAACL etc.)
    • Proficient in C / C + +, Python, and shell programming languages, and have a deep understanding of data structure and algorithm design

    As a condition of employment, all successful candidates must be able to establish authorization to work in the United States. For this position, the Company does not provide sponsorship or any immigration-related benefits.

    Numbers & Facts

    LocationSan Jose, CA

    Skills

    • Algorithmsunmatched
    • Artificial Intelligence (AI)unmatched
    • Audio Compressionunmatched
    • Audiovisualunmatched
    • C Programming Languageunmatched
    • C++ Programming Languageunmatched
    • Computer Engineeringunmatched
    • Computer Scienceunmatched
    • Computer Visionunmatched
    • Conferencesunmatched
    • Data Structuresunmatched
    • Deep Learningunmatched
    • Electrical Engineeringunmatched
    • Laboratoryunmatched
    • Machine Learningunmatched
    • Mathematicsunmatched
    • Modeling Languagesunmatched
    • Musicunmatched
    • Network Architecture/Engineeringunmatched
    • Onboardingunmatched
    • Physicsunmatched
    • Product Developmentunmatched
    • Programming Languagesunmatched
    • Publicationsunmatched
    • Python Programming/Scripting Languageunmatched
    • Roboticsunmatched
    • Scientific Researchunmatched
    • Speech Synthesisunmatched
    • Speech Technologyunmatched
    • Unix Shell Programmingunmatched
    • Voice Applicationsunmatched
    • Voice Productsunmatched

    Be found by employers

    5,500+ employers search our resume database daily. Add yours to get found by recruiters looking for candidates like you.

    Level up your application

    Professional resume templates

    Browse dozens of recruiter approved resume templates, layouts and formats. Choose your favorite and make it your own in minutes.

    Free resume templates

    Free resume builder

    Improve your existing resume or start from scratch and create a standout, ATS-friendly resume. Add job-specific content, download and apply.

    Free resume builder