Research Scientist Graduate (Seed-Speech Foundation Model) - 2027 Start

ByteDance

$218K — $387K *
Information Technology
Less than 5 years of experience
Job Overview by Ladders

Qualifications

  • Bachelor's degree in Computer Science, Engineering, Physics, Mathematics, or related discipline.
  • Strong coding ability in C/C++ or Python with knowledge in data structures and algorithms.
  • Demonstrated interest in relevant areas through academic or project experience.
  • Internship experience in speech processing or audio modeling is preferred.
  • Excellent problem-solving and collaboration skills.

Responsibilities

  • Develop and scale speech models for understanding and generation tasks.
  • Design training pipelines involving data construction and model alignment.
  • Enhance capabilities in speech recognition, synthesis, reasoning, and robustness.
  • Optimize model architecture and training efficiency.
  • Investigate interactive interfaces for speech-based systems.

Benefits

  • Medical, dental, and vision insurance available on day one.
  • 401(k) savings plan with company match.
  • Paid parental leave, short and long-term disability coverage, and life insurance.
  • Wellbeing benefits included.
  • 10 paid holidays and 10 sick days annually, plus 17 days of Paid Personal Time.
Full Job Description
Responsibilities - Develop and scale speech foundation models for understanding and generation tasks. - Design training pipelines including data construction, instruction tuning, and model alignment. - Improve core capabilities such as speech recognition, synthesis, reasoning, and robustness. - Optimize model architectures, training efficiency, and system performance. - Explore natural and interactive interfaces for speech-based systems. Minimum Qualifications: - Individuals who are completing or have recently completed a Bachelor's degree in Computer Science, Electrical Engineering, Electrical and Computer Engineering, Physics, Mathematics, or a related discipline. - Excellent coding ability, data structures, and fundamental algorithm skills, proficient in C/C++ or Python, etc. - Demonstrated interest or project experience in relevant areas. Preferred Qualifications: - Experience in speech processing, audio modeling, or related areas through internships is preferred. - Strong problem-solving and collaboration skills. 【For Pay Transparency】Compensation Description (Annually) The base salary range for this position in the selected city is $218400 - $387600 annually. Compensation may vary outside of this range depending on a number of factors, including a candidate's qualifications, skills, competencies and experience, and location. Base pay is one part of the Total Package that is provided to compensate and recognize employees for their work, and this role may be eligible for additional discretionary bonuses/incentives, and restricted stock units. Benefits may vary depending on the nature of employment and the country work location. Employees have day one access to medical, dental, and vision insurance, a 401(k) savings plan with company match, paid parental leave, short-term and long-term disability coverage, life insurance, wellbeing benefits, among others. Employees also receive 10 paid holidays per year, 10 paid sick days per year and 17 days of Paid Personal Time (prorated upon hire with increasing accruals by tenure). The Company reserves the right to modify or change these benefits programs at any time, with or without notice.

Similar Jobs

More Jobs at ByteDance

More Information Technology Jobs

Find similar Research Scientist Graduate (Seed-Speech Foundation Model) - 2027 Start jobs: