iGEN
Visit IGEN World Explore IGEN Expo
EXPLORE UPGRADE PLANS
BREAKING
Moody's Assigns First-Time Baa2 Rating to RBL Bank, One Notch Above India's Sovereign Sebi Bars Zee's Subhash Chandra, Punit Goenka From Market for One Year Zepto Defers IPO by Two to Three Quarters After Tepid Investor Response Tim Cook: India Among Apple's Best Global Markets as June Quarter Records Revenue Domestic funds reach record 21% stake in Indian companies as FPI ownership drops to 17% Cybercriminals widen net as assessees rush to meet I-T return filing deadline Bloomberg Delays India's Sovereign Bond Index Inclusion as Market Reforms Need Further Testing Gold loans jump 93.8% y-o-y, fuel bank credit growth in Q1FY27 Snapchat joins YouTube, LinkedIn and Substack in fight against 'AI slop' Amazon speeds last-mile delivery, expands robotics fleet past 1 million Moody's Assigns First-Time Baa2 Rating to RBL Bank, One Notch Above India's Sovereign Sebi Bars Zee's Subhash Chandra, Punit Goenka From Market for One Year Zepto Defers IPO by Two to Three Quarters After Tepid Investor Response Tim Cook: India Among Apple's Best Global Markets as June Quarter Records Revenue Domestic funds reach record 21% stake in Indian companies as FPI ownership drops to 17% Cybercriminals widen net as assessees rush to meet I-T return filing deadline Bloomberg Delays India's Sovereign Bond Index Inclusion as Market Reforms Need Further Testing Gold loans jump 93.8% y-o-y, fuel bank credit growth in Q1FY27 Snapchat joins YouTube, LinkedIn and Substack in fight against 'AI slop' Amazon speeds last-mile delivery, expands robotics fleet past 1 million
Home ›› Technology ›› Ai ›› Computer Vision ›› AI Video Generation Method for Cardiac MRI Addresses Data Scarcity with Latent Motion Modeling

AI Video Generation Method for Cardiac MRI Addresses Data Scarcity with Latent Motion Modeling

Researchers propose a generative method for synthesizing temporally coherent and anatomically consistent cardiac sequences from clinical text prompts. The model decouples spatial structure from temporal motion using a fine-tuned diffusion model and latent flow conditioning, achieving strong fidelity metrics. This approach addresses the scarcity of public cardiac MRI datasets.

iG
iGEN Editorial
June 16, 2026
AI Video Generation Method for Cardiac MRI Addresses Data Scarcity with Latent Motion Modeling

A team of researchers has introduced a generative method for synthesizing temporally coherent and anatomically consistent cardiac sequences, according to a paper published on arXiv. The work, titled "Temporally Consistent and Controllable Video Generation of 2D Cine CMR via Latent Space Motion Modeling," addresses the scarcity of public datasets that limits the development of advanced data-driven models for cine cardiac magnetic resonance (CMR)—the gold standard for assessing cardiac function.

The Data Scarcity Problem

Cine CMR is essential for evaluating cardiac function, but the limited availability of public datasets hinders the training of sophisticated AI models. The researchers propose a text-to-video framework that generates high-fidelity, on-demand medical data, offering a scalable solution to this data shortage.

How the Model Works: Decoupling Structure and Motion

The framework decouples cardiac spatial structure from temporal motion. First, a fine-tuned diffusion model synthesizes an initial frame from a clinical text prompt, controlling anatomical features. Then, a latent flow model conditioned on a cardiac phase embedding generates the complete cardiac motion, ensuring spatial consistency and temporal control. This two-stage approach allows the model to generate anatomically and pathologically diverse sequences with high temporal coherence and strong fidelity to input prompts.

Quantitative Results

The model's performance was evaluated using two key metrics. The Frechet Inception Distance (FID), which measures image realism, achieved a score of 31.68. The CLIP score, which measures alignment between text prompts and generated images, reached 31.04. These experimental results highlight its potential to produce high-fidelity medical data.

Metric Value Interpretation
FID (Frechet Inception Distance) 31.68 Lower is better; indicates realism of generated frames
CLIP score 31.04 Higher is better; measures text-image alignment

Implications for Medical AI

By enabling controlled generation of cardiac sequences from text prompts, this method could reduce reliance on scarce real-world datasets. The ability to produce diverse pathological variations on demand may accelerate research and model development in cardiac imaging. While the paper focuses on medical applications, the underlying technique of decoupling structure and motion in latent space could inform video generation tasks in other domains that require temporal consistency.


Sources:

Keep Reading

Recommended Stories

Breast MRI AI Challenge Reveals Trade-Offs Between Accuracy and Fairness Across Patient Subgroups Technology

Breast MRI AI Challenge Reveals Trade-Offs Between Accuracy and Fairness Across Patient Subgroups

The MAMA-MIA Challenge provided a standardized benchmark for breast MRI tumor segmentation and pathologic complete response prediction. Using a training cohort of 1,506 patients from US institutions and an external test set of 574 patients from three European centers, 26 international teams showed substantial performance variability and trade-offs between overall accuracy and subgroup fairness across age, menopausal status, and breast density.

June 21, 2026
BrainG3N Tokenizer Enables Controllable 3D Brain MRI Generation with Clinical-Grade Embeddings Technology

BrainG3N Tokenizer Enables Controllable 3D Brain MRI Generation with Clinical-Grade Embeddings

BrainG3N, a novel tokenizer for 3D brain MRI latent diffusion, decouples encoder and decoder to preserve clinical information while enabling high-quality reconstruction. Pretrained on 35,309 volumes, it outperforms SOTA models on 21 of 23 clinical tasks and supports controllable generation for disease simulation and privacy-preserving data sharing.

June 20, 2026
First Billion-Parameter Generative Foundation Model for Chest Radiography Achieves Expert-Level Synthesis Fidelity Technology

First Billion-Parameter Generative Foundation Model for Chest Radiography Achieves Expert-Level Synthesis Fidelity

Ribeiro et al. present the largest specialist generative foundation model for chest radiographs, with over 1.3 billion parameters. Trained on 1.2 million radiographs, the model supports controllable generation across demographics, views, and pathologies, advancing synthesis fidelity to clinical indistinguishability.

June 20, 2026
DySink: Dynamic Frame Sinks Enable Adaptive Long Video Generation Without Context Collapse Technology

DySink: Dynamic Frame Sinks Enable Adaptive Long Video Generation Without Context Collapse

Researchers propose DySink, a retrieval-based framework that replaces static early-frame sinks with dynamic, visually relevant historical frames for autoregressive long video generation. This approach prevents sink collapse and improves temporal quality in minute-long videos.

June 16, 2026