Artificial Intelligence
Ai2 Releases Molmo 2, an Open Video Model That Outperforms Qwen 3, GPT-5, and Gemini 2.5 Pro While Knowing Where the Action Happens
Ai2 has unveiled Molmo 2, the latest iteration of its open-source vision-language model (VLM). Arriving over a year after the original, this state-of-the-art update brings the most notable upgrades yet: support for multiple images and video, and grounding. The next-generation Molmo can now count and track objects or actions within videos. And just like its […]