Artificial Intelligence #ver0#open rl recipe
Vero: An Open RL Recipe for General Visual Reasoning — A Fully Open Vision-Language Model Family
A new research paper introduces Vero, a family of fully open vision-language models (VLMs) that use reinforcement learning (RL) to achieve strong general visual reasoning. The team constructed a 600K-sample dataset from 59 datasets and designed task-routed rewards. Vero variants outperformed their base models by 2.9-5.4 points on average across a 30-benchmark suite, and the best variant surpassed a stronger closed model by 3.8 points. All code, data, and models are released publicly.
Jun 21, 2026 1 source