Senior Interactive Vision Model Researcher

Tencent

$158K — $297K *
Information Technology
Less than 5 years of experience
Job Overview by Ladders

Qualifications

  • Ph.D. in AI-related fields with first-author papers at top-tier conferences.
  • Proficient in diffusion models and autoregressive generation engineering.
  • Experience training video/image generation models from scratch; advanced skills in PyTorch and large-scale distributed training.
  • Deep understanding of design trade-offs in Video VAE/Tokenizers, with 3+ years of relevant research experience.
  • Publications in video generation or diffusion models.
  • Prefer hands-on R&D experience with the latest industry technologies.
  • Preferred experience leading end-to-end video foundation model projects.

Responsibilities

  • Design and develop next-generation video generation architecture.
  • Investigate technologies for long video generation, focusing on attention mechanisms and memory solutions.
  • Advance high-compression video tokenization techniques for various resolutions and frame rates.
  • Lead large-scale video pre-training initiatives, including curriculum learning strategies.
  • Explore Scaling Laws and develop methods to enhance model capabilities.
  • Monitor industry advancements, conduct comparative experiments, and inform the technical roadmap.

Benefits

  • Eligibility for a sign-on payment, relocation package, and restricted stock units.
  • Includes medical, dental, vision, life, and disability benefits.
  • Participation in the company's 401(k) plan.
  • Generous vacation policy ranging from 15 to 25 days based on tenure.
  • Up to 13 holidays and 10 days of paid sick leave per year.
Full Job Description
What the Role Entails

1.Design and iterate on the next-generation video generation foundation architecture.
2.Explore core technologies for long video generation, including long-context attention, KV Cache compression, and memory mechanisms.
3.Research high-compression-ratio video tokenizers and advance unified modeling capabilities for multi-resolution and multi-frame-rate videos.
4.Lead video pre-training at the large token scale, defining data mixture and curriculum learning strategies.
5.Explore Scaling Laws, define scientific scale-up paths, and continuously improve key model capabilities.
6.Track industry state-of-the-art , lead comparative experiments, drive technical roadmap decisions, and publish academic papers.

Who We Look For

1.Ph.D. in AI-related fields with first-author papers at top-tier conferences.
2.Proficient in the principles and engineering implementation of diffusion models and autoregressive generation.
3.Experience training video/image generation models from scratch; highly proficient in PyTorch and large-scale distributed training.
4.Deep understanding of the design trade-offs in Video VAE/Tokenizers, with 3+ years of relevant research experience.
5.Publications related to video generation or diffusion models.
6.Experience in core industry product R&D or hands-on experience with the latest technologies is preferred.
7.Experience leading end-to-end video foundation model projects is preferred.

Location State(s)

US-California-Palo Alto

The expected base pay range for this position in the location(s) listed above is $158,300.00 to $297,000.00 per year. Actual pay may vary depending on job-related knowledge, skills, and experience.Employees hired for this position may be eligible for a sign on payment, relocation package, and restricted stock units, which will be evaluated on a case-by-case basis.Subject to the terms and conditions of the plans in effect, hired applicants are also eligible for medical, dental, vision, life and disability benefits, and participation in the Company's 401(k) plan. The Employee is also eligible for up to 15 to 25 days of vacation per year (depending on the employee's tenure), up to 13 days of holidays throughout the calendar year, and up to 10 days of paid sick leave per year.Your benefits may be adjusted to reflect your location, employment status, duration of employment with the company, and position level. Benefits may also be pro-rated for those who start working during the calendar year.

Similar Jobs

More Jobs at Tencent

More Information Technology Jobs

Find similar Senior Interactive Vision Model Researcher jobs: