Research Scientist - Multi-modal AI & Efficient Generative Models

Meta

$160K — $190K *
Consumer Technology
5 - 7 years of experience
Job Overview by Ladders

Qualifications

  • PhD in Computer Science or related field with machine learning, Deep Learning research focus on vision-language models.
  • Bachelor's degree in Computer Science, Computer Engineering, or equivalent experience.
  • 7+ years leading impactful research projects in the industry.
  • Proficient in Deep Learning development using PyTorch or TensorFlow.
  • Skilled in Python or C/C++ for developing deep learning models or infrastructure.
  • Published first-authored papers in top-tier conferences like ICLR, ICML, and NeurIPS.

Responsibilities

  • Drive the advancement of multi-modal understanding techniques to optimize intelligent systems.
  • Communicate complex features clearly while promoting product quality and engineering efficiency.
  • Conduct research to improve generative models and perception models for Meta's products.
  • Apply research insights to enhance Meta's smart glasses and VR offerings.
  • Define and execute research roadmaps over multi-month periods to push the field forward.
  • Collaborate globally across teams in research and product domains.
  • Present research findings at prestigious peer-reviewed conferences.

Benefits

  • Work in a cutting-edge environment focused on machine learning innovations.
  • Opportunity for significant influence on the development of VR and smartglass technologies.
  • Access to collaboration with cross-functional teams worldwide.
  • Engagement with top-tier research conferences and opportunities to publish findings.
Full Job Description
Drive the organization's goal towards relevant machine learning techniques in the area of multi-modal understanding and generation to build & optimize our intelligent systems that improve Meta's products and experiences
• Effectively communicate complex features and systems in detail while advocating for higher product quality and engineering efficiency
• Conduct applied research to advance the state of the art in efficient generative models (Diffusion models, Multi-modal LLMs, Vision Language Action models) and efficient perception models (Scene and Video understanding)
• Apply research to advance Meta's smartglasses and VR product lines
• Advance the state of the art in your problem area by defining and executing research roadmaps over 6-month or longer timeframes
• Collaborate with different cross-functional teams across the globe in research and product
• Present the outcomes of the research findings as papers in top-tier peer-reviewed conferences in the area

Minimum Qualifications
• PhD in Computer Science or a related field with published projects in the fields of machine learning, Deep learning with a focus on vision-language models
• Bachelor's degree in Computer Science, Computer Engineering, relevant technical field, or equivalent practical experience
• 7+ years of experience leading research projects with industry-wide impact
• Proven development skills in Deep Learning, working with PyTorch or TensorFlow
• Experience developing deep learning models or infrastructure in Python or C/C++
• Experience in one or more of the following areas: deep learning, Computer Vision, language models, Machine Learning or artificial intelligence
• First-authored publications at peer-reviewed conferences, e.g. ICLR, ICML, CVPR, ECCV, ICCV, NeurIPS

Preferred Qualifications
• Experience with CPU/GPU and mobile optimization
• Experience solving complex problems and comparing alternative solutions, trade-offs, and diverse points of view to determine a path forward

Similar Jobs

More Jobs at Meta

More Consumer Technology Jobs

Find similar Research Scientist - Multi-modal AI & Efficient Generative Models jobs: