Skip to content
← Back to job listings

Research Scientist in Speech Foundation Model - Seed - Graduates - 2027 Start (BS/MS)

ByteDance · San Jose, California, United States of America

External listingfull-timeRecently

About The Role

About the team

The mission of the Seed Speech team is to enrich interactive and creative processes through the application of multimodal speech technologies. The team focuses on the forefront of research and product development in speech and audio, music, natural language understanding, and multimodal deep learning.

Responsibilities

  • Develop and scale speech foundation models for understanding and generation tasks.
  • Design training pipelines including data construction, instruction tuning, and model alignment.
  • Improve core capabilities such as speech recognition, synthesis, reasoning, and robustness.
  • Optimize model architectures, training efficiency, and system performance.
  • Explore natural and interactive interfaces for speech-based systems.

Minimum Qualifications

  • Currently pursuing a Bachelor's or Master's degree in computer science, mathematics, engineering, or a related field, with an expected graduation date in 2027 and the ability to commit to an onboarding date by the end of 2027.
  • Excellent coding ability, data structures, and fundamental algorithm skills, proficient in C/C++ or Python, etc.
  • Demonstrated interest or project experience in relevant areas.

Preferred Qualifications

  • Experience in speech processing, audio modeling, or related areas through internships is preferred.
  • Strong problem-solving and collaboration skills.

This is an external listing. JobSpring does not represent or verify the employer. Report this listing