ByteDance
ByteDance

Research Scientist Graduate (Seed-Speech Foundation Model) - 2027 Start

职能机器学习
级别中级
地点San Jose, Canada, United States
方式现场办公
类型Regular
发布今天
立即申请

职位介绍

About the team
The mission of the Seed Speech team is to enrich interactive and creative processes through the application of multimodal speech technologies. The team focuses on the forefront of research and product development in speech and audio, music, natural language understanding, and multimodal deep learning.

Responsibilities:

  • Develop and scale speech foundation models for understanding and generation tasks.
  • Design training pipelines including data construction, instruction tuning, and model alignment.
  • Improve core capabilities such as speech recognition, synthesis, reasoning, and robustness.
  • Optimize model architectures, training efficiency, and system performance.
  • Explore natural and interactive interfaces for speech-based systems.

Requirements:

Minimum Qualifications:

  • Individuals who are completing or have recently completed a Bachelor's degree in Computer Science, Electrical Engineering, Electrical and Computer Engineering, Physics, Mathematics, or a related discipline.
  • Excellent coding ability, data structures, and fundamental algorithm skills, proficient in C/C++ or Python, etc.
  • Demonstrated interest or project experience in relevant areas.

Preferred Qualifications:

  • Experience in speech processing, audio modeling, or related areas through internships is preferred.
  • Strong problem-solving and collaboration skills.

必备技能

Machine learning

Model evaluation

Data workflows

关于ByteDance

San Jose

总部位置