Research Scientist (diffusion)

Genmo

• $150K — $180K *
Consumer Technology
Less than 5 years of experience
Job Overview by Ladders

Qualifications

  • Ph.D. in Computer Science, Artificial Intelligence, Machine Learning, or a related field
  • Strong publication record in top-tier conferences focused on generative models
  • Extensive experience with large-scale generative models for image or video
  • Deep understanding of text-to-image and text-to-video generation techniques
  • Proficient in Python and deep learning frameworks (e.g., PyTorch, TensorFlow)
  • Excellent communication skills for diverse audiences
  • Proven collaborative skills in team settings

Responsibilities

  • Lead research initiatives to enhance diffusion models for video generation
  • Develop algorithms to convert textual descriptions into video content
  • Conduct experiments to validate ideas and assess model performance
  • Collaborate with cross-functional teams for production integration
  • Review academic literature and participate in leading conferences
  • Publish research findings and contribute to open-source projects
  • Mentor junior researchers and promote an innovative team culture
  • Align research with user needs in collaboration with product teams

Benefits

  • Collaborative work environment with cross-functional teams
  • Opportunities for professional development and mentorship
  • Access to cutting-edge tools and resources in the field
  • Contributions to impactful research and open-source initiatives
  • Engagement with the research community through conferences
Full Job Description
Role overview:

We are seeking an exceptional Research Scientist to join our team, focusing on developing cutting-edge diffusion models for text-to-video generation. In this role, you will be at the forefront of innovation, creating novel architectures and algorithms that transform written descriptions into stunning, coherent video content.

Key responsibilities:
  • Lead research initiatives in advanced diffusion models for text-to-video generation, focusing on improving visual quality, temporal consistency, and semantic fidelity
  • Develop and implement state-of-the-art algorithms for translating textual descriptions into dynamic video content
  • Design and conduct rigorous experiments to validate new ideas and evaluate model performance
  • Collaborate with cross-functional teams to integrate research breakthroughs into our production pipeline
  • Stay at the cutting edge of the field by regularly reviewing academic literature and attending top-tier conferences
  • Contribute to the research community through high-quality publications and open-source contributions
  • Mentor junior researchers and foster a culture of innovation within the research team
  • Work closely with product teams to align research directions with user needs and market opportunities


Qualifications:
  • Ph.D. in Computer Science, Artificial Intelligence, Machine Learning, or a closely related field
  • Must have:
    • Strong publication record in top-tier conferences (e.g., CVPR, ICCV, NeurIPS, ICML) with a focus on generative models, particularly diffusion models
    • Extensive experience implementing and optimizing large-scale generative models for image or video tasks
    • Deep understanding of state-of-the-art techniques in text-to-image and text-to-video generation
    • Proficiency in Python and deep learning frameworks such as PyTorch or TensorFlow
    • Excellent communication skills with the ability to explain complex technical concepts to diverse audiences
    • Proven ability to work collaboratively in a team environment
  • Ideal candidate will have:
    • Postdoctoral or industrial research experience in generative AI for video
    • Hands-on experience with text-to-video generation projects
    • Expertise in other generative model architectures (e.g., GANs, VAEs) and their applications to video
    • Experience working with large-scale datasets and distributed computing environments
    • Track record of successful collaboration with product teams on technology transfers
    • Familiarity with video codecs, compression techniques, and perceptual quality metrics
    • Contributions to open-source projects in the field of generative AI


Additional information

The role is based in the Bay Area (San Francisco). Candidates are expected to be located near the Bay Area or open to relocation.

Similar Jobs

More Consumer Technology Jobs

Find similar Research Scientist (diffusion) jobs: