iGEN
Visit IGEN World Explore IGEN Expo
EXPLORE UPGRADE PLANS
BREAKING
The Chemical Blind Spot in Ballast Water Treatment: Unmonitored Bromoform Risks to Ports and Shipping AstraZeneca India Buys First Group Cancer Top-Up Cover for 25,000 Employees Banks Collected Rs 7,100 Crore as Minimum Balance Penalty in FY26, Private Banks Account for 70% Gold loans surge over 50% in FY26 on larger ticket sizes, shift from unsecured credit India Smartphone Exports Surge 23% to $9.8 Billion in Q1 Driven by iPhone Shipments India Seeks Sunflower Oil Alternatives as Black Sea Disruptions Delay Shipments and Raise Prices GenAI to Reshape Indian Workforce, Global Capability Centres to Cushion Job Impact: Goldman Sachs Mahindra Arm and SML Merge to Become India's 4th Largest Commercial Vehicle Player Exporters Flag Concerns Over West Asia War Impact on Trade and Logistics Asian stocks trade mixed after Fed holds rates; Kospi rallies 4%, Shenzhen slips 300 points The Chemical Blind Spot in Ballast Water Treatment: Unmonitored Bromoform Risks to Ports and Shipping AstraZeneca India Buys First Group Cancer Top-Up Cover for 25,000 Employees Banks Collected Rs 7,100 Crore as Minimum Balance Penalty in FY26, Private Banks Account for 70% Gold loans surge over 50% in FY26 on larger ticket sizes, shift from unsecured credit India Smartphone Exports Surge 23% to $9.8 Billion in Q1 Driven by iPhone Shipments India Seeks Sunflower Oil Alternatives as Black Sea Disruptions Delay Shipments and Raise Prices GenAI to Reshape Indian Workforce, Global Capability Centres to Cushion Job Impact: Goldman Sachs Mahindra Arm and SML Merge to Become India's 4th Largest Commercial Vehicle Player Exporters Flag Concerns Over West Asia War Impact on Trade and Logistics Asian stocks trade mixed after Fed holds rates; Kospi rallies 4%, Shenzhen slips 300 points
Home ›› Topics ›› reliability

Topic

reliability

9 stories
Last-Mile Delivery: Why Reliability Now Trumps Speed for Consumer Satisfaction Logistics
Last Mile Delivery #last-mile delivery#reliability

Last-Mile Delivery: Why Reliability Now Trumps Speed for Consumer Satisfaction

According to FreightWaves, last-mile delivery reliability has overtaken speed as the second-most important consumer factor after price. Jake Stein of Burq explains how hybrid, API-connected carrier networks can improve reliability and reduce hidden costs.

Jul 30, 2026 1 source
Why Essential Technology Fails When Temperatures Rise: Lessons from June Heatwave Technology
Hardware #heat#tech failure

Why Essential Technology Fails When Temperatures Rise: Lessons from June Heatwave

The June heatwave exposed how essential technology—from electricity transformers to healthcare IT—fails under extreme temperatures. A transformer failure in Brittany left over 100,000 without power, while six NHS trusts in England declared critical incidents. Experts warn that rising temperatures are reducing efficiency across energy networks and requiring climate resilience strategies.

Jul 1, 2026 1 source
PM Intervals: Fleet Cost Savings vs. Hidden Violation Costs in Preventive Maintenance Manufacturing
Industrial Machinery #preventive maintenance#lifecycle extension

PM Intervals: Fleet Cost Savings vs. Hidden Violation Costs in Preventive Maintenance

Fleets often extend preventive maintenance intervals to cut costs, but the savings can be offset by roadside violations and repairs. FMCSA data shows brake and hub seal failures cluster at extended intervals, costing carriers more in the long run.

Jun 23, 2026 1 source
Efficient and Sound Probabilistic Verification Secures AI Agents Against Policy Violations Technology
Artificial Intelligence #ai#artificial intelligence

Efficient and Sound Probabilistic Verification Secures AI Agents Against Policy Violations

Researchers introduce a sound and efficient framework for probabilistic verification of AI agents, addressing the need for enforcing security policies under ambiguity. The approach computes upper bounds on violation probability without independence assumptions, outperforming prior art on standard benchmarks.

Jun 20, 2026 1 source
AI Safety Monitors May Fail After Model Updates, New Benchmarking Study Finds Technology
Artificial Intelligence #ai safety#model monitoring

AI Safety Monitors May Fail After Model Updates, New Benchmarking Study Finds

A new research paper presents the first systematic test of whether activation monitors remain reliable after common model updates such as quantization and fine-tuning. The study finds that while quantization largely preserves performance, fine-tuning frequently makes monitors stale, with privacy monitors most affected. Degradation is predictable, enabling triaged revalidation.

Jun 16, 2026 1 source
XFlow: A New Programming System for Reliable Multi-Agent Workflows Addresses Prompt–Harness Boundary Technology
Software #xflow#executable protocol

XFlow: A New Programming System for Reliable Multi-Agent Workflows Addresses Prompt–Harness Boundary

Researchers present XFlow, an executable protocol programming system designed to improve reliability in LLM-based multi-agent workflows. By introducing the XPF protocol language and lifecycle-governed symbols, XFlow makes constraints and process requirements explicit and enforceable, addressing the underspecified prompt–harness boundary that limits current systems.

Jun 16, 2026 1 source
Do LLMs Reliably Identify Correct Information Units in Aphasic Discourse? A New Study Evaluates Four Models Technology
Artificial Intelligence #llms#aphasia

Do LLMs Reliably Identify Correct Information Units in Aphasic Discourse? A New Study Evaluates Four Models

A study examined whether instruction-tuned large language models (LLMs) can reliably perform token-level classification of Correct Information Units (CIUs) from aphasic discourse transcripts. Four models—Llama-3.1-8B, Qwen2.5-7B, Mistral-7B, and Phi-3-mini—were tested under zero-shot and few-shot prompting conditions. Results showed that few-shot prompting yielded competitive mean F1 scores between 0.776 and 0.817 for three models, but zero-shot was insufficient and Phi-3-mini was unstable. The authors recommend a human-in-the-loop approach for automated CIU scoring.

Jun 16, 2026 1 source
Metric Match: New Subset Selection Method Improves LLM Judge Reliability Evaluation, Cuts Annotation Costs by 32.5% Technology
Artificial Intelligence #llm#judge

Metric Match: New Subset Selection Method Improves LLM Judge Reliability Evaluation, Cuts Annotation Costs by 32.5%

Researchers developed Metric Match, a subset selection method that reduces costly human annotations needed to evaluate LLM judge reliability. The approach achieves a 0.838 win-rate over random selection, cuts estimation error by 18.7%, and reduces annotation needs by 32.5%. A medical case study showed $1,041.67 in savings.

Jun 16, 2026 1 source
Mythos AI Exploits Hidden Fault Lines: 81% of Teams Still Ship Vulnerable Code Technology
Software #software#code quality

Mythos AI Exploits Hidden Fault Lines: 81% of Teams Still Ship Vulnerable Code

TechRadar reports that AI models like Claude Mythos have become dangerously adept at tracing connections across enterprise systems and exploiting hidden fault lines. Meanwhile, a Checkmarx study found that 81% of global AppSec leaders knowingly ship vulnerable code. The article argues that traditional AppSec is obsolete and calls for continuous, embedded security in development workflows.

Jun 14, 2026 1 source