Artificial Intelligence #ai#reinforcement learning
MetaResearcher AI Framework Trains Deep Research Agents via Self-Reflective Reinforcement Learning in Adversarial Environments
MetaResearcher is a novel AI framework for training deep research agents using self-reflective reinforcement learning in adversarial virtual environments. It introduces four synergistic dimensions: Evolving Virtual World, Discovery-Oriented Tasks, Self-Reflective Meta-Reward (GRPO), and Heterogeneous Multi-Agent Swarm. Built on LiteResearcher, it requires zero marginal API cost and targets improvements on GAIA and Xbench-DS benchmarks.
Jun 20, 2026 1 source