iGEN
Visit IGEN World Explore IGEN Expo
EXPLORE UPGRADE PLANS
BREAKING
FNP targets over Rs 1,400 crore revenue this fiscal as quick commerce and global expansion accelerate CMA CGM and Stonepeak launch $2.4bn United Ports joint venture for global terminal expansion Inflexor Ventures Secures Rs 400 Crore for Fund III, Targets 22-25 Technology Startups For Ease of Liquidity and Business, Government to Amend MSME Act India Industrial Output Grows 7.3% in June, Fastest Pace in 22 Months India Public Sector Banks' Bad Debt at Lowest Levels in Decades, Government Says India's Factory Output Surges 7.3% in June, Fastest Pace in Nearly Two Years Hindustan Unilever Announces More Price Increases Amid Persistent Inflation Crude Prices Climb Over 4% as Renewed Middle East Tensions and Inventory Draw Fuel Supply Fears Werner CEO Leathers Says Driver Attrition Only in 'Third Inning' as Regulatory Pressures Tighten Capacity FNP targets over Rs 1,400 crore revenue this fiscal as quick commerce and global expansion accelerate CMA CGM and Stonepeak launch $2.4bn United Ports joint venture for global terminal expansion Inflexor Ventures Secures Rs 400 Crore for Fund III, Targets 22-25 Technology Startups For Ease of Liquidity and Business, Government to Amend MSME Act India Industrial Output Grows 7.3% in June, Fastest Pace in 22 Months India Public Sector Banks' Bad Debt at Lowest Levels in Decades, Government Says India's Factory Output Surges 7.3% in June, Fastest Pace in Nearly Two Years Hindustan Unilever Announces More Price Increases Amid Persistent Inflation Crude Prices Climb Over 4% as Renewed Middle East Tensions and Inventory Draw Fuel Supply Fears Werner CEO Leathers Says Driver Attrition Only in 'Third Inning' as Regulatory Pressures Tighten Capacity
Home ›› Technology ›› Ai ›› New Reinforcement Learning Framework WeCAN Improves DAG Scheduling Efficiency in Heterogeneous Environments

New Reinforcement Learning Framework WeCAN Improves DAG Scheduling Efficiency in Heterogeneous Environments

Researchers propose WeCAN, an end-to-end reinforcement learning framework for scheduling directed acyclic graphs (DAGs) on heterogeneous resources. The method uses a two-stage single-pass design with a weighted cross-attention encoder and a skip-extended generation map to close optimality gaps. Experiments on TPC-H query DAGs and ML-compiler graphs show improved makespan with inference time comparable to classical heuristics.

iG
iGEN Editorial
June 17, 2026
New Reinforcement Learning Framework WeCAN Improves DAG Scheduling Efficiency in Heterogeneous Environments

Efficient scheduling of dependent tasks in data-intensive computing remains a critical bottleneck for enterprises running large-scale workloads. Directed acyclic graphs (DAGs), which represent task dependencies in query plans, data processing pipelines, and computation graphs, must be mapped to limited heterogeneous resource pools under tight runtime budgets. Existing schedulers often struggle to adapt across different environments or generate schedules quickly enough for real-time decisions.

A team of researchers has introduced WeCAN, an end-to-end reinforcement learning framework for heterogeneous DAG scheduling that addresses two key challenges: compatibility between tasks and resource pools, and optimality gaps introduced by the schedule-generation process. The work is detailed in a paper titled "A Learning Method with Gap-Aware Generation for Heterogeneous DAG Scheduling," authored by Zhou, Ruisong; Zou, Haijun; Li, Sun; Chumin; Wen; and Zaiwen, and published on arXiv.

Two-Stage Single-Pass Design

WeCAN adopts a two-stage single-pass architecture. In a single forward pass, the framework produces task-pool scores and global parameters. A subsequent generation map then constructs schedules without repeated network calls, making it computationally efficient. The weighted cross-attention encoder models task-pool interactions using compatibility coefficients and is size-agnostic to environment fluctuations, meaning it can handle varying numbers of tasks and resources.

Addressing Generation-Induced Optimality Gaps

The researchers identify that widely used list-scheduling maps can suffer from generation-induced optimality gaps due to restricted reachability. They introduce an order-space analysis that characterizes the reachable set of generation maps via feasible schedule orders, explains the mechanism behind these gaps, and yields sufficient conditions for elimination. Guided by these conditions, the team designs a skip-extended realization with an analytically parameterized decreasing skip rule. This enlarges the reachable order set while preserving single-pass efficiency.

Experimental Results

WeCAN was evaluated on three types of workloads: TPC-H query DAGs, resource-intensive workload datasets, and ML-compiler computation graphs. The framework demonstrated improved makespan over strong baselines.

Metric WeCAN Classical Heuristics Multi-Round Neural Schedulers
Makespan improvement Improved over baselines Baseline Baseline
Inference time Comparable to classical heuristics Fast Slower than WeCAN

According to the paper, WeCAN achieves inference time comparable to classical heuristics and faster than multi-round neural schedulers, making it suitable for runtime-constrained environments.

Implications for Enterprise Computing

For technology leaders managing cloud or on-premise heterogeneous clusters, the ability to generate high-quality schedules quickly can directly reduce job completion times and improve resource utilization. WeCAN's size-agnostic design and single-pass generation mean it can scale to large, dynamic workloads without retraining for each new environment. While the experiments focus on database queries and ML graphs, the underlying method could extend to any DAG-based workflow, including those in supply chain planning, logistics orchestration, and scientific computing.

The paper provides a mathematical framework for understanding and eliminating schedule-generation inefficiencies, which may inspire further optimizations in production schedulers. Enterprises investing in AI-driven operations should monitor such advances, as they promise to close the gap between theoretical scheduling quality and practical deployment constraints.


Sources:

Keep Reading

Recommended Stories

AWS Billing Glitch Shows Customers Erroneous Fees of Up to $7.1 Trillion Technology

AWS Billing Glitch Shows Customers Erroneous Fees of Up to $7.1 Trillion

A global billing glitch in Amazon Web Services displayed incorrect estimated charges to customers, with some seeing fees in the billions and trillions. The root cause was an issue with the estimated billing computation subsystem, which AWS is now rolling back. The issue is expected to be resolved by the weekend.

July 17, 2026
New York Governor Signs First Statewide Data Center Moratorium, Halting Hyperscale Development for One Year Technology

New York Governor Signs First Statewide Data Center Moratorium, Halting Hyperscale Development for One Year

New York Governor Kathy Hochul signed an executive order enacting a one-year moratorium on hyperscale data centers over 50 megawatts, the first statewide pause in the US. The order directs the Department of Public Service to evaluate environmental and energy impacts and proposes ending tax incentives. The move follows growing opposition and a legislative bill with stricter limits.

July 14, 2026
Microsoft Emissions Jump 25% as AI Data Center Expansion Fuels Greenhouse Gas Rise Technology

Microsoft Emissions Jump 25% as AI Data Center Expansion Fuels Greenhouse Gas Rise

Microsoft’s greenhouse gas emissions rose 25% in the last fiscal year, driven primarily by data center expansion for AI workloads. The company matched 100% of electricity with carbon-free sources but has signed new fossil-fuel-powered deals. Google and Amazon reported similar double-digit increases, highlighting the sustainability challenge of the AI race.

July 10, 2026
DynAMO: Dynamic Asset Management Orchestration via Topological Multi-Agent Scheduling Technology

DynAMO: Dynamic Asset Management Orchestration via Topological Multi-Agent Scheduling

A new research paper introduces DynAMO, a deployment-ready engine for LLM-powered agent orchestration in Industry 4.0. Using Plan-then-Execute architecture with sequential and parallel workflows, it achieves 1.6x median latency reduction, rising to 1.8x on parallelizable tasks. Structured context pruning cuts inference time 30%. Tests on AssetOpsBench show robust performance under fault injection.

July 8, 2026