Ad
Skip to content

OpenAI hires three Google DeepMind researchers for multimodal AI work

Image description
OpenAI / Midjourney promted by THE DECODER

December 5 update:

According to OpenAI, the new research unit in Zurich will initially focus on multimodal research for AI systems that can "understand and combine different types of information like text, images, and sound to complete tasks more effectively." The company says this work is essential for its broader goal of developing artificial general intelligence.

Original article from December 4:

OpenAI has recruited three researchers from Google DeepMind who specialize in multimodal AI: Lucas Beyer, Alexander Kolesnikov, and Xiaohua Zhai. The trio has worked together in recent years, making progress in computer vision model scaling and developing the Vision Transformer (ViT) architecture. OpenAI plans to put their expertise to use developing technologies that can process multiple types of data and handle complex interactions. The company is setting up a new office in Zurich as part of this effort, adding to its existing European locations in Dublin, London, Paris, and Brussels. The company still says its goal is to "develop artificial general intelligence that benefits everyone."

AI News Without the Hype – Curated by Humans

Subscribe to THE DECODER for ad-free reading, a weekly AI newsletter, our exclusive "AI Radar" frontier report six times a year, full archive access, and access to our comment section.

Read on for the full picture.
Subscribe for hype-free coverage.

  • Full access to every article on THE DECODER
  • No ads
  • Join the comments and community discussions
  • A weekly AI news recap via mail
  • 6x/year: "AI Radar" — deep dives on the AI topics that matter most
  • Daily AI news, always up to date
  • Our full ten-year archive
  • Covered by a team with 10+ years in AI
Subscribe to The Decoder