iGEN
Visit IGEN World Explore IGEN Expo
EXPLORE UPGRADE PLANS
BREAKING
CMA CGM and Stonepeak Launch United Ports LLC in $2.4 Billion Terminal Joint Venture UPS shift away from Amazon shows bigger payoff Lanesurf: 62% of Loads Get Vetted Carrier Offers Before Brokers Arrive India-China Border Trade Via Lipulekh Resumes Aug 1; China Permits 20 Traders Geopolitics Drives CMA CGM Q2 Profit Surge of 42% as Volumes and Rates Climb Benchmark Diesel Price Rises Third Week as Futures Plunge; Spread Hits Record Indian Government Limits Sugar Dealers to 400 Tonnes Stock Until November to Curb Hoarding Tenants signing longer leases for larger warehouses as 3PLs lock in capacity US stock market flat as S&P 500 and Dow barely move, Nasdaq slides over 1% on chip rout TruAlt Bioenergy Q1 Net Zooms to ₹59.27 Crore on Higher Revenues, Capacity Expansion CMA CGM and Stonepeak Launch United Ports LLC in $2.4 Billion Terminal Joint Venture UPS shift away from Amazon shows bigger payoff Lanesurf: 62% of Loads Get Vetted Carrier Offers Before Brokers Arrive India-China Border Trade Via Lipulekh Resumes Aug 1; China Permits 20 Traders Geopolitics Drives CMA CGM Q2 Profit Surge of 42% as Volumes and Rates Climb Benchmark Diesel Price Rises Third Week as Futures Plunge; Spread Hits Record Indian Government Limits Sugar Dealers to 400 Tonnes Stock Until November to Curb Hoarding Tenants signing longer leases for larger warehouses as 3PLs lock in capacity US stock market flat as S&P 500 and Dow barely move, Nasdaq slides over 1% on chip rout TruAlt Bioenergy Q1 Net Zooms to ₹59.27 Crore on Higher Revenues, Capacity Expansion
Home ›› Technology ›› Ai ›› Robotics ›› AI Isn't Smarter Than a Baby—Yet: New Test Reveals Limits of Vision Language Models

AI Isn't Smarter Than a Baby—Yet: New Test Reveals Limits of Vision Language Models

Researchers at Meta, Stanford, and other institutions developed EgoBabyVLM, a test comparing vision language models to infant learning. Current VLMs fail when fed realistic baby-camera footage, highlighting the efficiency gap between human and AI learning. The findings suggest that building baby-like AI could reduce costs and improve robot learning.

iG
iGEN Editorial
July 15, 2026
AI Isn't Smarter Than a Baby—Yet: New Test Reveals Limits of Vision Language Models

Despite the immense computing power behind today's generative AI models, a new test from researchers at Meta, Stanford University, the University of Tokyo, and France's École Normale Supérieure reveals that these systems still cannot match the learning efficiency of a one-year-old baby, according to WIRED.

The EgoBabyVLM Challenge

The EgoBabyVLM Challenge judges how well vision language models (VLMs), which learn from both text and imagery, make sense of the world as a baby sees it. The test requires a model to describe the world after ingesting about a thousand hours of video collected from cameras strapped to the heads of infants and toddlers. According to WIRED, "it turns out that the cutting-edge models fail miserably when fed this realistic and messy footage, which suggests there may be something different about the design of the baby brain that enables it to learn so rapidly from so little information."

Babies learn from a kaleidoscopic view: parents talking about objects no longer visible, indicating things using gaze or gestures, or discussing past or future events. Michael Frank, a cognitive scientist at Stanford who specializes in language learning and was involved with EgoBabyVLM's development, said: "it's clear that there's more [than just language] that's needed."

BabyLM and Language Learning

EgoBabyVLM follows a related challenge called BabyLM, introduced in 2023, which tasked AI models with learning the syntax of language using about the same amount of data a 10-year-old takes in—tens of millions of words, compared to trillions for AI models. Remarkably, transformer-based AI models can do this quite well, a finding that challenges Noam Chomsky's ideas about syntax being hardwired into the human brain. However, Ryan Cotterell, a linguist at ETH Zurich who first developed BabyLM, noted that understanding the physical world is different. "There isn't going to be a large corpus of human interactions—there's no internet of human interactions," he said.

Joshua Tenenbaum, a cognitive scientist at the Massachusetts Institute of Technology, observed that BabyLM showed models do not acquire "common sense" about the physical world, social dynamics, or theory of mind. "Transformers are very good at finding patterns in data," Tenenbaum said. "But it does seem that just pure pattern learning systems are not able to take the kind of data that a baby or a child receives and learn all the things that they do."

Implications for Enterprise AI

For enterprise decision-makers investing in AI, the findings carry practical weight. According to WIRED, building a more baby-like version of AI "could make frontier models less costly and less energy intensive." It might also be valuable if AI-powered robots are to learn about their environments in a more natural way. This is particularly relevant for industries like logistics and manufacturing, where robots often operate in unpredictable settings. Current large models require massive datasets and energy; a baby-like learning algorithm could dramatically lower the bar for deployment.


Sources: WIRED – Top Stories

Keep Reading

Recommended Stories

Thermodynamic Measure of Intelligence Proposed: Rare-Valid Futures Amplification as Universal Scale Technology

Thermodynamic Measure of Intelligence Proposed: Rare-Valid Futures Amplification as Universal Scale

A new theoretical framework proposes intelligence as the lawful amplification of rare but valid futures. The research presents a thermodynamic measure and shows that recursive self-simulation is necessary and nearly sufficient for high thermodynamic intelligence. This makes intelligence measurable on a universal scale.

July 8, 2026
Cognitive Trajectory Modeling: A New Framework for Quantifying Human-AI Co-Creation Technology

Cognitive Trajectory Modeling: A New Framework for Quantifying Human-AI Co-Creation

Cognitive Trajectory Modeling (CTM) is a novel cognitive theory of interaction dynamics that conceptualizes cognition and creative processes as temporally organized trajectories. It provides a framework for quantifying how human-AI co-creation evolves over time, distinguishing cognitive trajectories from mere interaction traces.

June 16, 2026
Bombay High Court to Hear Gadkari's Suit Against Meta, X, Google Over Deepfakes on August 5 Technology

Bombay High Court to Hear Gadkari's Suit Against Meta, X, Google Over Deepfakes on August 5

The Bombay High Court has scheduled August 5 for hearing Union Minister Nitin Gadkari's civil suit against Meta, X Corp, Google LLC and unknown persons over defamatory deepfakes and AI-generated posts. Gadkari alleges the fake content falsely portrays him as personally responsible for the ethanol-blending programme and claims financial benefit to him and his family, causing irreparable harm to his reputation and personality rights.

July 28, 2026
Can the New York Times Save Journalism From Our AI Overlords? Technology

Can the New York Times Save Journalism From Our AI Overlords?

New York Times publisher A.G. Sulzberger discusses the $20 million copyright lawsuit against OpenAI and Microsoft, the Trump administration's attacks on press freedom, and the existential challenges facing journalism as AI reshapes information consumption.

July 28, 2026