iGEN
Visit IGEN World Explore IGEN Expo
EXPLORE UPGRADE PLANS
BREAKING
Relay Q: London Startup's AI Microphone Puts Hands-Free Voice Dictation on the Desktop Google Pixel 10a Crowned Best Budget Pixel in WIRED's Updated 2026 Buying Guide Global Steel Wire seeks fresh Santander terminal concession Veritas Shipmanagement books fresh ultramax pair at COSCO yard, Splash247 reports Seanergy linked to fresh newcastlemax at Hengli as dry bulk orderbook grows Weaker rupee may push foreign assets over FAST-DS Rs 1 crore limit, raising tax bill 45 Indian power plants face critically low coal stocks as monsoon hits supply SFL Makes Fresh $363m Car Carrier Play With Four LNG Dual-Fuel Newbuilds Iran Blacklist Threatens Hormuz Shuttle Tanker Lifeline for Gulf Crude Keyfield International Enters Dredging Market with $24.7m Vessel Acquisition Relay Q: London Startup's AI Microphone Puts Hands-Free Voice Dictation on the Desktop Google Pixel 10a Crowned Best Budget Pixel in WIRED's Updated 2026 Buying Guide Global Steel Wire seeks fresh Santander terminal concession Veritas Shipmanagement books fresh ultramax pair at COSCO yard, Splash247 reports Seanergy linked to fresh newcastlemax at Hengli as dry bulk orderbook grows Weaker rupee may push foreign assets over FAST-DS Rs 1 crore limit, raising tax bill 45 Indian power plants face critically low coal stocks as monsoon hits supply SFL Makes Fresh $363m Car Carrier Play With Four LNG Dual-Fuel Newbuilds Iran Blacklist Threatens Hormuz Shuttle Tanker Lifeline for Gulf Crude Keyfield International Enters Dredging Market with $24.7m Vessel Acquisition
Home ›› Topics ›› action

Topic

action

3 stories
ROSE Benchmark Reveals Perception-to-Action Gap in Multimodal AI Models Technology
Artificial Intelligence #ai#multimodal

ROSE Benchmark Reveals Perception-to-Action Gap in Multimodal AI Models

The ROSE benchmark measures how reliably multimodal large language models (MLLMs) convert visual evidence into context-appropriate actions. Testing nine recent models, researchers found performance drops of up to 44.5 percentage points from counting to region-conditioned action, while humans achieve 98.8% accuracy.

Jun 22, 2026 3 sources
New Robotic Architecture AVP Improves Pick-and-Place Success Rate by 37% over Existing Models Technology
Artificial Intelligence #action#visual

New Robotic Architecture AVP Improves Pick-and-Place Success Rate by 37% over Existing Models

A new research paper introduces AVP (Action with Visual Primitives), an end-to-end architecture for robotic manipulation that decouples visual-language reasoning from action generation. In real-robot pick-and-place experiments, AVP achieved a 37.04% higher success rate than the pi_0.5 baseline, with gains in data efficiency, spatial-compositional generalization, and object-level transfer.

Jun 17, 2026 1 source
Survey on Medical Embodied AI Highlights Integration of Perception, Decision-Making, and Action Technology
Artificial Intelligence #medical ai#embodied ai

Survey on Medical Embodied AI Highlights Integration of Perception, Decision-Making, and Action

A systematic survey of medical embodied AI examines its core components — perception, decision-making, and action — and their coordinated integration for real-world clinical workflows. The paper reviews representative applications, datasets, and challenges, highlighting the need for unified system-level organization beyond individual functional aspects.

Jun 16, 2026 1 source