Dolby Laboratories

Sr. Foundational Audio AI Researcher

Dolby Laboratories$137K — $168K *
Consumer Technology
Less than 5 years of experience
Job Overview by Ladders

Qualifications

  • Ph.D. in Computer Science or a related field.
  • Strong expertise in deep learning, with both theoretical and practical knowledge.
  • Preferred experience in audio, video, or text processing technologies.
  • Proven publication record in top machine learning conferences (e.g., NeurIPS, ICLR).
  • Proficient in Python and popular deep learning frameworks (TensorFlow, PyTorch).
  • Ability to conceptualize new technologies for innovative product development.
  • Strong communication and collaboration skills.

Responsibilities

  • Collaborate with domain experts to refine Dolby's AI and ML strategy.
  • Utilize deep learning to develop new foundation models and improve existing solutions.
  • Advance the state-of-the-art and generate intellectual property.
  • Transfer technology effectively to product teams and draft patent applications.
  • Advise leadership on cutting-edge advancements in deep learning to shape research and business strategies.

Benefits

  • Opportunity to work on cutting-edge AI research and applications.
  • Collaborative environment with world-class audio and vision experts.
  • Involvement in impactful projects that leverage the latest AI technologies.
  • Supports professional growth through exposure to multi-modal learning advancements.
  • Participation in a global team within a leading innovation-driven company.
Full Job Description
Dolby is looking for a talented Senior AI Researcher to join Dolby's research efforts to develop the next generation of AI based audio and video technologies. You will work with Dolby's world-class audio and vision experts to invent new multimedia analysis, processing and rendering technologies. As a part of a global team, the Senior AI Researcher will work on ideas exploring new horizons in multi-modal media processing, analysis, replay and organization. You will be responsible for performing fundamental new research, transferring technology to product groups, and draft patent applications. Summary You will push the boundaries of the state-of-the-art in audio and media technologies. The ideal candidate would have a strong background in deep learning, both in terms of conceptual understanding, as well as practical experience. A core aspect of this role involves being able to keep up to date with the literature, implement, and innovate with the bleeding edge in generative models, self-supervised learning, and multi-modal learning. With the explosion of multi-modal foundation models and the growing capabilities of vision-language and audio-language systems, you will partner closely with Dolby's Applied AI team, which actively pursues the integration of these cutting-edge technologies into next-generation audio and media experiences. You will be able to hit the ground running, innovate, and contribute to impactful projects that leverage the latest advancements in AI. Consequently, experience with audio models, language models, question answering, vision-language models, captioning, etc. would be highly beneficial. What You Will Accomplish • Partner closely with other domain experts to refine and execute Dolby's technical strategy in artificial intelligence and machine learning. • Use deep learning to create new solutions (including foundation models) and enhance existing applications. • Push the state-of-the-art and develop intellectual property. • Transfer technology to product groups and draft patent applications. • Advise internal leaders on recent deep learning advancements in the industry and academia to further influence research direction and business decisions. Key Requirements • Ph.D. in Computer Science or similar field. • A strong background in deep learning, both in terms of conceptual understanding, as well as practical experience. • Knowledge in audio, video, or text processing is desirable. • Strong publication record, with publications in major machine learning conferences (e.g. NeurIPS, ICLR, ICML). Publications in top domain-specific conferences is desirable (e.g., ACL, CVPR, ICASSP). • Good knowledge about current machine learning literature. • Highly skilled in Python and one or more popular deep learning frameworks (TensorFlow or PyTorch). • Ability to envision new technologies and turn them into innovative products. • Good communication and collaboration skills. Consequently, knowledge or experience in any/all of the following are helpful: • Diffusion, autoregressive, or other generative models. • Self-supervised, contrastive learning, auto-encoders. • Audio, image, or text applications - Source separation, text-to-speech, music synthesis, image segmentation, image captioning, question answering, language models, etc. Learn more about our innovative research: https://www.dolby.com/about/innovation/empowering/ The Atlanta Area base salary range for this full-time position is $137,500-$168,200, which can vary if outside this location,plus bonus, benefits, and some roles may also include equity. Our salary ranges are determined by role, level, and location. Within the range, individual pay is determined by work location and additional factors, including job-related skills, competencies, experience, market demands, internal parity, and relevant education or training. Your recruiter can share more about the specific salary range and perks and benefits for your location during the hiring process.

About Dolby Laboratories

Dolby Laboratories, Inc. creates audio and imaging technologies that transform entertainment and communications at the cinema, at home, at work, and on mobile devices. The company develops and licenses its audio technologies, such as AAC & HE-AAC, a digital audio codec solution used for TVs, set-top boxes (STBs), personal computers (PCs), gaming consoles, mobile devices, and digital radio; AVC, a digital video codec with high bandwidth efficiency used in media devices; Dolby AC-4, an audio coding technology that delivers new audio experiences to a range of playback devices; and Dolby Atmos technology for home theaters, cinemas, device speakers, mobile devices, and headphones. Its audio technologies also comprise Dolby Digital, a digital audio coding technology that provides multichannel sound in the home; Dolby Digital Plus, a digital audio coding technology that delivers audio quality for streaming, downloaded, and broadcast content; Dolby TrueHD, a digital audio coding technology for content providers; Dolby Vision, an imaging technology for cinema, digital TV, and other consumer devices; and HEVC, a digital video codec with high bandwidth efficiency to support delivery of Ultra HD and other video content. In addition, the company designs and manufactures audio and imaging products, such as digital cinema servers, Dolby Cinema audio products, and other products for the film production, cinema, television broadcast, and entertainment industries. Further, it offers services to support theatrical and television production for cinema exhibition, broadcast, and home entertainment. The company serves film studios, content creators, post-production facilities, cinema operators, broadcasters, and video game designers. Dolby Laboratories, Inc. was founded in 1965 and is headquartered in San Francisco, California.
Learn more about Dolby Laboratories
Size
2,368 employees
Market Cap
$6.5 billion
Industry
Net Income
$317.8 million
Founded
1965
5 Year Trend
+3%
Revenue
$1.2 billion
NASDAQ

Similar Jobs

More Jobs at Dolby Laboratories

More Consumer Technology Jobs

Find similar Sr. Foundational Audio AI Researcher jobs: