Topic
standardization
LLM Agent With Ontology Constraints Automates Standardization of Legacy Biomedical Metadata
Researchers have developed an LLM-based metadata standardization system that queries standard reporting guidelines and biomedical terminology services in real time. Tested on 839 legacy records from the Human BioMolecular Atlas Program, the approach consistently improves prediction accuracy over LLM-only methods for both ontology-constrained and non-ontology-constrained fields.
AgentBeats Proposes Open Standard for Reproducible AI Agent Evaluation Across Benchmarks
A new research paper introduces AgentBeats, a framework for open, standardized, and reproducible AI agent assessment. The approach uses judge agents and protocols A2A and MCP to unify evaluation, demonstrated through a five-month competition with 298 judge agents and 467 subject agents.
PrologMCP: A Standardized Prolog Tool Interface That Boosts LLM Agents’ Deductive Accuracy
A team of researchers introduced PrologMCP, an open-source server that exposes Prolog as a stateful tool through the Model Context Protocol, allowing LLM agents to delegate deductive reasoning tasks. In evaluations on the PARARULE-Plus benchmark, an agent powered by PrologMCP achieved accuracy of 1.00 on a general sample, matching or exceeding reasoning LLMs, and 1.00/0.99 on a challenging subset where reasoning models dropped to 0.95/0.94.