iGEN
Visit IGEN World Explore IGEN Expo
EXPLORE UPGRADE PLANS
BREAKING
OpenAI’s Rogue AI Agent Hacked More Than Just Hugging Face Inside the rogue ChatGPT hack of Hugging Face: AI agents operate at superhuman speed but make clumsy mistakes Landstar Expects to Emerge a Winner After Supreme Court’s Montgomery Ruling Widens Broker Liability New Senate bill targets 'chameleon carriers' that reopen to escape penalties Werner Enterprises Posts Highest Revenue Per Truck Growth in One-Way Segment in a Decade CMA CGM and Stonepeak Launch United Ports LLC in $2.4 Billion Terminal Joint Venture UPS shift away from Amazon shows bigger payoff Lanesurf: 62% of Loads Get Vetted Carrier Offers Before Brokers Arrive India-China Border Trade Via Lipulekh Resumes Aug 1; China Permits 20 Traders Geopolitics Drives CMA CGM Q2 Profit Surge of 42% as Volumes and Rates Climb OpenAI’s Rogue AI Agent Hacked More Than Just Hugging Face Inside the rogue ChatGPT hack of Hugging Face: AI agents operate at superhuman speed but make clumsy mistakes Landstar Expects to Emerge a Winner After Supreme Court’s Montgomery Ruling Widens Broker Liability New Senate bill targets 'chameleon carriers' that reopen to escape penalties Werner Enterprises Posts Highest Revenue Per Truck Growth in One-Way Segment in a Decade CMA CGM and Stonepeak Launch United Ports LLC in $2.4 Billion Terminal Joint Venture UPS shift away from Amazon shows bigger payoff Lanesurf: 62% of Loads Get Vetted Carrier Offers Before Brokers Arrive India-China Border Trade Via Lipulekh Resumes Aug 1; China Permits 20 Traders Geopolitics Drives CMA CGM Q2 Profit Surge of 42% as Volumes and Rates Climb
Home ›› Technology ›› Ai ›› Computer Vision ›› SLUM-i: AI Semi-Supervised Learning Maps Informal Settlements with Benchmark Dataset

SLUM-i: AI Semi-Supervised Learning Maps Informal Settlements with Benchmark Dataset

A new AI framework called SLUM-i uses semi-supervised learning to map informal settlements in cities like Lahore, Karachi, and Mumbai. It introduces a benchmark dataset and achieves up to +5.9 pp mIoU improvement over existing methods.

iG
iGEN Editorial
June 17, 2026
SLUM-i: AI Semi-Supervised Learning Maps Informal Settlements with Benchmark Dataset

Rapid urban expansion in low- and middle-income countries has driven the growth of informal settlements in major cities such as Lahore, Karachi, and Mumbai. However, large-scale mapping of these areas is hampered by annotation scarcity and data quality issues, including high spectral ambiguity between formal and informal structures and significant annotation noise, according to a research paper published on arXiv. To address this, a team of researchers led by Muhammad Taha Mukhtar and colleagues has introduced SLUM-i, a semi-supervised segmentation framework designed to improve informal settlement mapping.

The Benchmark Dataset

The researchers constructed a benchmark dataset for Lahore from scratch, and derived companion datasets for Karachi and Mumbai from verified administrative boundaries, totaling approximately 900 km² of urban area. This collection is supplemented by four cities from prior literature across Sub-Saharan Africa and Latin America, with comprehensive data quality assessments provided for each city, the study reports.

City Country Source
Lahore Pakistan Constructed from scratch
Karachi Pakistan Derived from verified administrative boundaries
Mumbai India Derived from verified administrative boundaries
Four additional cities Sub-Saharan Africa and Latin America From prior literature

The SLUM-i Framework

SLUM-i is a semi-supervised segmentation framework that mitigates the class imbalance and distribution mismatch inherent in standard semi-supervised learning pipelines, the authors explain. It integrates two key components:

  • Class-Aware Adaptive Thresholding: A mechanism that dynamically adjusts confidence thresholds to prevent minority class suppression.
  • DINOv2-based unlabeled pool filter: Removes out-of-distribution tiles prior to training to reduce covariate shift.

Both components are architecture-agnostic and add no inference overhead, according to the paper.

Performance Gains

Extensive experiments across seven cities spanning three continents, repeated over five random seeds, demonstrated gains of up to +5.9 pp mIoU over state-of-the-art semi-supervised baselines. The study notes that both components contribute to this improvement without increasing computational burden during inference.

Gains of up to +5.9 pp mIoU over state-of-the-art semi-supervised baselines, with both components being architecture-agnostic and adding no inference overhead.

Implications for Urban Mapping

The research directly addresses critical challenges in mapping informal settlements, which are often poorly documented yet vital for urban planning and resource allocation. By providing a benchmark dataset and an effective semi-supervised framework, SLUM-i enables more accurate and scalable mapping in regions where labeled data is scarce. The methodology's architecture-agnostic design allows it to be integrated into existing computer vision pipelines, making it practical for deployment by urban planners and development organizations.

For enterprise technology leaders, the approach demonstrates how advanced AI techniques like semi-supervised learning and vision transformers (DINOv2) can overcome data quality and annotation bottlenecks. Such capabilities are relevant for applications beyond informal settlements, including infrastructure monitoring and logistics planning in complex urban environments.


Sources:

Keep Reading

Recommended Stories

RSRCC Benchmark Uses Retrieval-Augmented Best-of-N Ranking for Remote Sensing Change Comprehension Technology

RSRCC Benchmark Uses Retrieval-Augmented Best-of-N Ranking for Remote Sensing Change Comprehension

RSRCC is a new benchmark for remote sensing change question-answering, containing 126k questions focused on localized, semantic changes. It uses a hierarchical semi-supervised curation pipeline with retrieval-augmented Best-of-N ranking to filter noisy candidates. The dataset is available online.

June 16, 2026
New Research Reveals How Visual Tokens Evolve Inside Vision-Language Models Technology

New Research Reveals How Visual Tokens Evolve Inside Vision-Language Models

A new computer vision paper from arXiv investigates how visual tokens are integrated into large language models (LLMs) under two paradigms: in-context prompting and layer-wise injection. The authors find that visual tokens enter the LLM as 'disguised visual context' lacking linguistic structure, then evolve differently depending on the integration architecture. They show that attention allocation alone is insufficient, and performance depends on the quality of visual representations at each layer.

July 8, 2026
ROSE Benchmark Reveals Perception-to-Action Gap in Multimodal AI Models Technology

ROSE Benchmark Reveals Perception-to-Action Gap in Multimodal AI Models

The ROSE benchmark measures how reliably multimodal large language models (MLLMs) convert visual evidence into context-appropriate actions. Testing nine recent models, researchers found performance drops of up to 44.5 percentage points from counting to region-conditioned action, while humans achieve 98.8% accuracy.

June 22, 2026
New AI Research Shows Vision-Language Models Think Better with Visual Grounding Technology

New AI Research Shows Vision-Language Models Think Better with Visual Grounding

Researchers introduce visually grounded thinking, a reasoning process that interleaves natural-language thoughts with explicit point or box groundings to image regions. The method, using a scalable synthesis pipeline and grounding-aware reinforcement learning, consistently improves performance of Gemma3-4B-IT on counting and spatial reasoning benchmarks, with the 4B model matching or surpassing the 27B variant.

June 21, 2026