iGEN
Visit IGEN World Explore IGEN Expo
EXPLORE UPGRADE PLANS
BREAKING
Werner Enterprises Posts Highest Revenue Per Truck Growth in One-Way Segment in a Decade CMA CGM and Stonepeak Launch United Ports LLC in $2.4 Billion Terminal Joint Venture UPS shift away from Amazon shows bigger payoff Lanesurf: 62% of Loads Get Vetted Carrier Offers Before Brokers Arrive India-China Border Trade Via Lipulekh Resumes Aug 1; China Permits 20 Traders Geopolitics Drives CMA CGM Q2 Profit Surge of 42% as Volumes and Rates Climb Benchmark Diesel Price Rises Third Week as Futures Plunge; Spread Hits Record Indian Government Limits Sugar Dealers to 400 Tonnes Stock Until November to Curb Hoarding Tenants signing longer leases for larger warehouses as 3PLs lock in capacity US stock market flat as S&P 500 and Dow barely move, Nasdaq slides over 1% on chip rout Werner Enterprises Posts Highest Revenue Per Truck Growth in One-Way Segment in a Decade CMA CGM and Stonepeak Launch United Ports LLC in $2.4 Billion Terminal Joint Venture UPS shift away from Amazon shows bigger payoff Lanesurf: 62% of Loads Get Vetted Carrier Offers Before Brokers Arrive India-China Border Trade Via Lipulekh Resumes Aug 1; China Permits 20 Traders Geopolitics Drives CMA CGM Q2 Profit Surge of 42% as Volumes and Rates Climb Benchmark Diesel Price Rises Third Week as Futures Plunge; Spread Hits Record Indian Government Limits Sugar Dealers to 400 Tonnes Stock Until November to Curb Hoarding Tenants signing longer leases for larger warehouses as 3PLs lock in capacity US stock market flat as S&P 500 and Dow barely move, Nasdaq slides over 1% on chip rout
Home ›› Technology ›› Ai ›› Robotics ›› OnDeFog: Online Decision Transformer That Handles Frame Dropping Outperforms Prior Methods

OnDeFog: Online Decision Transformer That Handles Frame Dropping Outperforms Prior Methods

Researchers propose OnDeFog, an online variant of the Decision Transformer under Random Frame Dropping (DeFog) that integrates DeFog's mechanisms with the Online Decision Transformer (ODT). Experiments show OnDeFog outperforms ODT in high dropping-rate environments and surpasses DeFog when datasets contain large amounts of low-reward data.

iG
iGEN Editorial
June 20, 2026
OnDeFog: Online Decision Transformer That Handles Frame Dropping Outperforms Prior Methods

Frame dropping – the loss of state and reward information due to communication delays or sensor failures – poses a significant challenge for reinforcement learning (RL) agents deployed in the physical world. Existing solutions such as the Decision Transformer under Random Frame Dropping (DeFog) can mitigate performance degradation, but because DeFog is an offline learning method, it struggles to generalise to novel states not adequately represented in its training dataset.

To address this limitation, researchers Yotsufuji, Daiki, Nishihara, Kenta, Shimizu, Shoma, Uchida, Kento, and Shirakawa, Shinichi have developed OnDeFog, an online approach that integrates the mechanisms of DeFog with the Online Decision Transformer (ODT). OnDeFog learns policies through direct environmental interaction, enabling it to handle frame dropping while retaining the ability to adapt to new situations.

How OnDeFog Works

OnDeFog builds on two prior lines of work:

  • Decision Transformer (DT): a model that frames RL as a sequence-modelling problem using a transformer architecture.
  • DeFog (Decision Transformer under Random Frame Dropping): an offline extension of DT that adds mechanisms to cope with missing states and rewards caused by frame dropping.

By combining DeFog's specialised mechanisms with ODT's online learning capability, OnDeFog can update its policy continuously even when frames are dropped during deployment.

Performance Results

The researchers conducted a comprehensive experimental evaluation. Key findings include:

  • OnDeFog achieves superior performance compared to ODT in environments characterised by high dropping frame rates.
  • OnDeFog outperforms DeFog on datasets containing a large amount of low-reward data.

The results indicate that the online nature of OnDeFog helps it generalise better than offline DeFog when the training data is noisy or sparse, while its frame-dropping mechanisms allow it to maintain robustness where ODT fails.

Method Environment with High Drop Rate Dataset with Low-Reward Data
ODT Lower performance Not reported
DeFog Moderate performance Lower performance
OnDeFog Superior performance Superior performance

Table: Relative performance of OnDeFog versus ODT and DeFog as reported in the study.

Implications for Real-World AI

While the research is presented in a machine-learning context, its relevance extends to any domain where RL agents operate under imperfect sensing – including robotics in logistics, autonomous vehicles, and industrial automation. For enterprise technology buyers evaluating AI deployment in supply chain or warehouse automation, frame robustness is a critical factor. Systems that can gracefully handle sensor dropouts without retraining from scratch offer higher uptime and lower maintenance costs.

The work also highlights a broader trend: moving from offline to online learning to close the gap between training and deployment conditions. OnDeFog's architecture suggests that hybrid approaches – combining the structural advantages of transformers with online adaptation – may become a standard design pattern for production RL systems.

No information was provided in the source about specific industries, deployment costs, or integration timelines. The paper is available on arXiv under a Creative Commons license.


Sources:

Keep Reading

Recommended Stories

1X Neo Robot's Freaky Fast Fingers Bring Human-Level Dexterity to Home and Office Technology

1X Neo Robot's Freaky Fast Fingers Bring Human-Level Dexterity to Home and Office

1X, a Norwegian-American robotics company, revealed its Neo robot's five-finger hands capable of 25 degrees of freedom, gripping odd shapes, and detecting slippage. The robot is partly teleoperated via Expert Mode, with early access pricing of $20,000 or $500 per month. The hands are IP68 waterproof, enabling the robot to wash itself.

July 9, 2026
How Automation Erodes Human Control: Lessons from the Decline of the Manual Transmission Technology

How Automation Erodes Human Control: Lessons from the Decline of the Manual Transmission

In a WIRED book excerpt, Ian Bogost explores how automation has quietly reduced direct human-machine interaction, using the near-extinction of manual transmissions as a metaphor. Data from CarMax shows stick-shift sales dropped from over 15% in 2000 to 2.4% in 2020, as automakers like Mercedes and Volkswagen phase out manuals. Philosopher Matthew Crawford argues that maintaining 'natural bonds between action and perception' is essential for autonomy and meaning.

July 7, 2026
Flexion Robotics' AI Turns Humanoid Robots Into Competent Office Interns for Logistics Automation Technology

Flexion Robotics' AI Turns Humanoid Robots Into Competent Office Interns for Logistics Automation

Flexion Robotics, a Swiss startup founded by ex-Nvidia researchers, has developed an AI system that trains humanoid robots to perform multistep tasks like retrieving parcels and navigating offices autonomously. The approach uses reinforcement learning in simulation and video-based skill matching, aiming to make humanoids viable for logistics and supply chain automation.

June 29, 2026
Reinforcement-Aware Knowledge Distillation Boosts LLM Reasoning Efficiency Technology

Reinforcement-Aware Knowledge Distillation Boosts LLM Reasoning Efficiency

Researchers propose RL-aware distillation (RLAD) to address distribution mismatch and objective interference in knowledge distillation for LLM reasoning. The method uses Trust Region Ratio Distillation (TRRD) to selectively imitate teacher policies during reinforcement learning. RLAD outperforms offline distillation, standard GRPO, and KL-based on-policy distillation across logic and math benchmarks.

June 21, 2026