Senior Researcher - Multimodal Lab
Dolby Laboratories
Job Description
Summary
Dolby’s research division is currently looking for a talented, self-motivated AI researcher to push the boundaries of the state-of-the-art in audio and media technologies. An ideal candidate would have a strong background in deep learning, both in terms of conceptual understanding, as well as practical experience. A core aspect of this role involves being able to keep up to date with the literature, implement, and innovate with the bleeding edge in generative models, self-supervised learning, and multi-modal learning. Consequently, knowledge or experience in any/all the following are helpful:
• Diffusion, autoregressive, or other generative models.
• Self-supervised, contrastive learning, auto-encoders.
• Audio, image, or text applications – Source separation, text-to-speech, music synthesis, image segmentation, image captioning, question answering, language models, etc.
With the explosion of large language models and natural language processing, the candidate will work closely with Dolby’s Applied AI team, which actively pursues the integration of such models into audio and media experiences. Prospective candidates would be expected to hit the ground running, innovate, and contribute to such projects. Consequently, experience with language models, question answering, vision-language models, captioning, etc. would be highly beneficial.
Main responsibilities:
Requirements:
All official communication regarding employment opportunities at Dolby will come from an official dolby.com email address. We will never request payment as part of the hiring process. If you receive a suspicious message, please verify its authenticity before responding.