iGEN
Visit IGEN World Explore IGEN Expo
EXPLORE UPGRADE PLANS
BREAKING
Home ›› Technology ›› Ai ›› Computer Vision ›› Mutual Distillation of Dual Foundation Models Achieves State-of-the-Art PET/CT Segmentation with Only 5 Labeled Cases

Mutual Distillation of Dual Foundation Models Achieves State-of-the-Art PET/CT Segmentation with Only 5 Labeled Cases

Researchers propose MuDuo, a mutual distillation framework that leverages two foundation models (SAM-Med3D for CT, SegAnyPET for PET) to distill knowledge into a lightweight student network for semi-supervised PET/CT segmentation. Achieving state-of-the-art performance on the AutoPET dataset with only 5 labeled cases, the approach eliminates manual prompts and maximizes unlabeled data utility.

iG
iGEN Editorial
June 16, 2026
Mutual Distillation of Dual Foundation Models Achieves State-of-the-Art PET/CT Segmentation with Only 5 Labeled Cases

Organ segmentation from PET/CT is critical for quantitative analysis and radiotherapy planning in oncology, but the high cost of expert annotation limits the development of deep learning models. A team of researchers has proposed MuDuo, a mutual distillation framework that exploits both structural and functional foundation models to achieve state-of-the-art performance on the AutoPET dataset using only 5 labeled cases.

The Annotation Bottleneck in Medical Imaging

According to the research paper published on arXiv (arXiv:2606.15611), semi-supervised learning (SSL) provides a practical and effective solution for developing deep models with limited labeled data. Recent developments in visual foundation models have demonstrated remarkable adaptability with improved efficiency. The team's work bridges the gap between the task-specific precision of student models and the segmentation priors of generalist foundation models.

MuDuo: Mutual Distillation Framework

The proposed framework, MuDuo, synergistically leverages two modality-specific foundation models:

  • SAM-Med3D for structural CT imaging
  • SegAnyPET for metabolic PET imaging

Both act as generalists that distill their knowledge into a lightweight student network. The approach eliminates the need for manual prompts while maximizing the utility of unlabeled data for automatic segmentation.

Technical Details and Performance

The key innovation is mutual distillation: the two foundation models are used as teachers, each specializing in one modality, and the student network learns from both. The authors report state-of-the-art performance on the AutoPET dataset with only 5 labeled cases. The source code is publicly available at the project's GitHub repository.

Implications for Enterprise AI Adoption

While this work focuses on medical imaging, the concept of leveraging pre-trained foundation models through distillation to reduce labeled data requirements has broad applications. For enterprise technology leaders, the ability to deploy high-performance AI models with minimal annotated data translates directly into lower costs and faster time-to-value. The framework demonstrates that combining multiple large models as teachers can produce lightweight, efficient student models suitable for deployment in resource-constrained environments.

The research was conducted by Mao, Fuyou, Wu, Beining, Jiang, Yanfeng, Xu, Bohan, Lin, Lixin, Naye, Zhang, Hao, and Tang. The full paper is available under a CC BY 4.0 license on arXiv.


Sources:

Keep Reading

Recommended Stories

First Billion-Parameter Generative Foundation Model for Chest Radiography Achieves Expert-Level Synthesis Fidelity Technology

First Billion-Parameter Generative Foundation Model for Chest Radiography Achieves Expert-Level Synthesis Fidelity

Ribeiro et al. present the largest specialist generative foundation model for chest radiographs, with over 1.3 billion parameters. Trained on 1.2 million radiographs, the model supports controllable generation across demographics, views, and pathologies, advancing synthesis fidelity to clinical indistinguishability.

June 20, 2026
BrainG3N Tokenizer Enables Controllable 3D Brain MRI Generation with Clinical-Grade Embeddings Technology

BrainG3N Tokenizer Enables Controllable 3D Brain MRI Generation with Clinical-Grade Embeddings

BrainG3N, a novel tokenizer for 3D brain MRI latent diffusion, decouples encoder and decoder to preserve clinical information while enabling high-quality reconstruction. Pretrained on 35,309 volumes, it outperforms SOTA models on 21 of 23 clinical tasks and supports controllable generation for disease simulation and privacy-preserving data sharing.

June 20, 2026
UniBrain: A Unified Multimodal Model for Brain MRI Imputation and Understanding Technology

UniBrain: A Unified Multimodal Model for Brain MRI Imputation and Understanding

Researchers propose UniBrain, a unified multimodal large language model for brain MRI analysis that handles missing data through joint imputation and understanding. The model uses interleaved data flow, self-alignment, and dynamic hidden state mechanisms to achieve high performance on multi-disease MRI datasets.

June 16, 2026
Deep Learning Automates Doppler Angle Estimation in Ultrasound, Reducing Measurement Errors Technology

Deep Learning Automates Doppler Angle Estimation in Ultrasound, Reducing Measurement Errors

A deep learning approach developed using 2100 carotid ultrasound images can automatically estimate Doppler angle, reducing error. The best model achieved mean absolute error less than clinical threshold, potentially improving blood velocity measurements.

June 16, 2026