Principal Interactive Vision Model Researcher

Tencent

$158K — $326K *
Information Technology
Less than 5 years of experience
Job Overview by Ladders

Qualifications

  • Ph.D. in AI-related fields with first-author papers at top-tier conferences.
  • Proficient in diffusion models and autoregressive generation.
  • Experience training video/image generation models from scratch; proficient in PyTorch and large-scale distributed training.
  • Deep understanding of design trade-offs in Video VAE/Tokenizers with 3+ years of relevant research experience.
  • Publications related to video generation or diffusion models.
  • Preferred experience in core industry product R&D and cutting-edge technologies.
  • Preferred experience managing end-to-end video foundation model projects.

Responsibilities

  • Design and iterate on next-generation video generation architecture.
  • Research technologies for long video generation, including attention mechanisms and memory solutions.
  • Advance high-compression-ratio video tokenizers for multi-resolution and multi-frame-rate capabilities.
  • Lead large-scale pre-training for video models with a focus on data mixture and learning strategies.
  • Explore scaling laws and establish scientific paths for model improvements.
  • Track state-of-the-art advancements and lead comparative experiments to inform technical strategies.
  • Publish findings in academic forums to contribute to the field.

Benefits

  • Sign-on payment and relocation package availability, evaluated on a case-by-case basis.
  • Restricted stock units may be offered as part of compensation.
  • Comprehensive medical, dental, vision, life, and disability benefits.
  • 401(k) plan participation with company contributions.
  • Up to 25 vacation days based on tenure.
  • Up to 13 holidays observed each calendar year.
  • Up to 10 days of paid sick leave annually.
Full Job Description
What the Role Entails

1.Design and iterate on the next-generation video generation foundation architecture.
2.Explore core technologies for long video generation, including long-context attention, KV Cache compression, and memory mechanisms.
3.Research high-compression-ratio video tokenizers and advance unified modeling capabilities for multi-resolution and multi-frame-rate videos.
4.Lead video pre-training at the large token scale, defining data mixture and curriculum learning strategies.
5.Explore Scaling Laws, define scientific scale-up paths, and continuously improve key model capabilities.
6.Track industry state-of-the-art , lead comparative experiments, drive technical roadmap decisions, and publish academic papers.

Who We Look For

1.Ph.D. in AI-related fields with first-author papers at top-tier conferences.
2.Proficient in the principles and engineering implementation of diffusion models and autoregressive generation.
3.Experience training video/image generation models from scratch; highly proficient in PyTorch and large-scale distributed training.
4.Deep understanding of the design trade-offs in Video VAE/Tokenizers, with 3+ years of relevant research experience.
5.Publications related to video generation or diffusion models.
6.Experience in core industry product R&D or hands-on experience with the latest technologies is preferred.
7.Experience leading end-to-end video foundation model projects is preferred.

Location State(s)

US-California-Los Angeles

The expected base pay range for this position in the location(s) listed above is $158,300.00 to $326,000.00 per year. Actual pay may vary depending on job-related knowledge, skills, and experience.Employees hired for this position may be eligible for a sign on payment, relocation package, and restricted stock units, which will be evaluated on a case-by-case basis.Subject to the terms and conditions of the plans in effect, hired applicants are also eligible for medical, dental, vision, life and disability benefits, and participation in the Company's 401(k) plan. The Employee is also eligible for up to 15 to 25 days of vacation per year (depending on the employee's tenure), up to 13 days of holidays throughout the calendar year, and up to 10 days of paid sick leave per year.Your benefits may be adjusted to reflect your location, employment status, duration of employment with the company, and position level. Benefits may also be pro-rated for those who start working during the calendar year.

Similar Jobs

More Jobs at Tencent

More Information Technology Jobs

Find similar Principal Interactive Vision Model Researcher jobs: