What you will do:You will push the boundaries of the state-of-the-art in audio, video, and multi-modal technologies. The ideal candidate would have a strong background in deep learning, both in terms of conceptual understanding, as well as practical experience. A core aspect of this role involves being able to keep up to date with the literature, implement, and innovate with the bleeding edge in generative models, self-supervised learning, and multi-modal learning.
With the explosion of audio, video, and multi-modal foundational models, you will partner closely with Dolby's worldwide AI research staff, which actively pursues the integration of such models into audio and media experiences. You will be able to hit the ground running, innovate, and contribute to such projects.
Consequently, knowledge or experience in any/all of the following are helpful:
- Diffusion, autoregressive, or other generative models.
- Natural Language processing
- Self-supervised, contrastive learning, auto-encoders.
- Audio, image, or text applications - Source separation, text-to-speech, music synthesis, image segmentation, image captioning, question answering, language models, etc.
Main Responsibilities- Partner closely with other domain experts to refine and execute Dolby's technical strategy in artificial intelligence and machine learning.
- Apply deep learning and foundation models to create new solutions and enhance existing applications.
- Push the state-of-the-art and develop new technology
- Transfer technology to product groups and draft patent applications.
- Advise internal leaders on recent deep learning advancements in the industry and academia to further influence research direction and business decisions.
Requirements- Ph.D. in Computer Science or similar field.
- A strong background in deep learning, both in terms of conceptual understanding, as well as practical experience.
- Knowledge in audio, video, or text processing is desirable.
- Strong publication record, with publications in major machine learning conferences (e.g. NeurIPS, ICLR, ICML). Publications in top domain-specific conferences is desirable (e.g., ACL, CVPR, ICASSP).
- Good knowledge about current machine learning literature.
- Highly skilled in Python and one or more popular deep learning frameworks (TensorFlow or PyTorch).
- Ability to envision new technologies and turn them into innovative products. Creativity.
- Good communication and collaboration skills.
The Atlanta Area base salary range for this full-time position is $138.000-170,000, which can vary if outside this location,plus bonus, benefits, equity and profit sharing. Our salary ranges are determined by role, level, and location. Within the range, individual pay is determined by work location and additional factors, including job-related skills, competencies, experience, market demands, internal parity, and relevant education or training. Your recruiter can share more about the specific salary range and perks and benefits for your location during the hiring process.
Dolby will consider qualified applicants with criminal histories in a manner consistent with the requirements of San Francisco Police Code, Article 49, and Administrative Code, Article 12