Develop and evaluate machine-learning algorithms, use behavioral and customer data to propose value, build prototype applications, collaborate with business teams, and publish research.
About this role
This internship is based in an R&D organization researching large-scale generative AI for content creation and production in music, film, and games. The role sits within a research program focused on natural language processing and multimodal machine learning, with opportunities to address problems connected to entertainment content and creative workflows.
The work covers fundamental research in multimodal learning, multimodal large language models, music and video understanding, agents, reasoning, controllable generation, deep generative models, image and audio captioning, text-to-image and text-to-audio systems, vision-language pre-training, commonsense knowledge graphs, and large-scale data development. The stated development environment includes Python, C/C++, Linux, Windows, PyTorch, and TensorFlow. Research output may include submissions to conferences such as ACL, EMNLP, NeurIPS, ICLR, and CVPR.
Applicants must have a master’s degree in NLP, AI, machine learning, or a closely related field, or equivalent practical experience. The posting requires three years of experience with Python, C/C++, and Linux/Unix; two years in machine learning and NLP using frameworks such as PyTorch and TensorFlow; and research ability shown through papers, open-source software, or other scientific work.
This is an AI-generated summary of the employer's original posting — details can be incomplete, out of date or simply wrong. Always confirm everything on the official posting before applying.