Researcher - Computer Vision and Multimodal Foundation Models

Huawei Technologies Canada Co., Ltd.

• $110K — $130K *
Information Technology
Less than 5 years of experience
Job Overview by Ladders

Qualifications

  • Advanced degree (Ph.D. preferred) in Computer Science, Robotics, or related field focused on computer vision and multimodal foundation models.
  • Deep understanding of spatial reasoning and foundation models, including VLM, LLM, and MLLM.
  • Proven expertise in applying real-world computer vision algorithms.
  • Strong programming skills in Python and C++, with experience in deep learning frameworks like PyTorch and TensorFlow.
  • Excellent problem-solving and analytical capabilities to address complex challenges.
  • Strong publication record in top AI conferences/journals.

Responsibilities

  • Contribute independently to research and development in computer vision and multimodal foundation models.
  • Collaborate with cross-functional teams to integrate computer vision solutions.
  • Mentor junior engineers and interns, promoting a collaborative atmosphere.
  • Develop intellectual property through patent filings and publications.
  • Stay updated with advancements in computer vision and machine learning, identifying innovative opportunities.
  • Manage projects independently, ensuring timely and quality delivery.

Benefits

  • Flexible working hours.
  • Opportunities for professional development and growth.
  • Access to the latest technology and tools.
  • Collaborative and innovative team environment.
Full Job Description
Responsibilities:
  • Independently contribute to research and development efforts within the area of computer vision and multimodal foundation models.
  • Collaborate closely with cross-functional teams to integrate computer vision solutions into the broader spatial reasoning system.
  • Mentor and guide junior engineers and interns, fostering a collaborative and innovative team environment.
  • Contribute to the development of our intellectual property portfolio through patent filings and publications.
  • Stay abreast of the latest research and advancements in computer vision and machine learning. Proactively identify opportunities for innovation.
  • Independently manage projects and deliver high-quality results within established timelines.


  • Advanced degree (Ph.D. preferred) in Computer Science, Robotics, or a related field with a focus on computer vision and multimodal foundation models.
  • Deep understanding of spatial reasoning and foundation models, e.g., VLM, LLM, MLLM, etc.
  • Proven expertise in applying computer vision algorithms for real-world application.
  • Strong programming skills in Python and C++, with experience using deep learning frameworks (PyTorch, TensorFlow).
  • Excellent problem-solving and analytical skills, with the ability to tackle complex technical challenges.
  • A strong publication record in top-tier AI conferences/journals (e.g., CVPR, ICCV, ECCV, NeurIPS, ICML, ICLR, AAAI, etc.).

Similar Jobs

More Jobs at Huawei Technologies Canada Co., Ltd.

More Information Technology Jobs

Find similar Researcher - Computer Vision and Multimodal Foundation Models jobs: