Geodesic-informed Generative Diffusion Model For Topology-preserved Image Video Generation
Nian Wu1
, Nivetha Jayakumar1
, Jiarui Xing2
, Miaomiao Zhang1,3
1: Department of Electrical and Computer Engineering, University of Virginia, Charlottesville, VA, USA, 2: School of Medicine, Yale University, New Haven, CT, USA, 3: Department of Computer Science, University of Virginia, Charlottesville, VA, USA
Publication date: 2026/08/30
https://doi.org/10.59275/j.melba.2026-eec1
Abstract
Generative diffusion models have emerged as a class of powerful techniques for various imaging applications, including but not limited to synthesis, reconstruction, and segmen- tation. Despite their success, current generative models pose two key limitations. First, they primarily rely on image intensity and texture information, with limited attention to underlying object geometry. As a result, they do not guarantee geometric or topological consistency during the generation process, which is a crucial requirement for high-stakes domains such as computational anatomy, biology, and robotics, where preserving object structure is critical. Second, existing models fail to explicitly learn or represent shape changes in the generative process. Such deformation dynamics remain occluded within network parameters; hence leaving the transformation process uninterpretable and phys- ically uninformed. To address these challenges, we introduce IGG (Image Generation informed by Geodesic dynamics), a novel framework that integrates topology-preserving geodesic principles into the diffusion-based generative process. In contrast to conventional methods that operate in image intensity space, IGG learns and synthesizes diverse sam- ples within geodesic deformation spaces, where geometric object changes are learned as smooth and invertible smooth mappings from a given template/source image. Specifically, IGG employs a two-stage architecture: (i) a geodesic-informed image registration (GIR) module that directly encodes geodesic paths of image deformations into a compact latent space, and (ii) a latent geodesic diffusion (LGD) model that captures the distribution of these deformation representations, conditioned on a template image and/or text prompts. We validate IGG on the datasets of plant growth and brain MRI scans. Experimental results demonstrate that IGG outperforms state-of-the-art image generation and editing models, producing realistic, high-quality images with preserved topology and fewer arti- facts. Furthermore, when used for data augmentation in downstream segmentation tasks, IGG substantially improves segmentation accuracy, particularly in low-data regimes. Our code is publicly available at https://github.com/nellie689/IGG
Keywords
Geodesics · Diffeomorphisms · Generative diffusion model · Neural Operator
Bibtex
@article{melba:2026:021:wu,
title = "Geodesic-informed Generative Diffusion Model For Topology-preserved Image Video Generation",
author = "Wu, Nian and Jayakumar, Nivetha and Xing, Jiarui and Zhang, Miaomiao",
journal = "Machine Learning for Biomedical Imaging",
volume = "2026",
issue = "IPMI 2025 special issue",
year = "2026",
pages = "440--456",
issn = "2766-905X",
doi = "https://doi.org/10.59275/j.melba.2026-eec1",
url = "https://melba-journal.org/2026:021"
}
RIS
TY - JOUR
AU - Wu, Nian
AU - Jayakumar, Nivetha
AU - Xing, Jiarui
AU - Zhang, Miaomiao
PY - 2026
TI - Geodesic-informed Generative Diffusion Model For Topology-preserved Image Video Generation
T2 - Machine Learning for Biomedical Imaging
VL - 2026
IS - IPMI 2025 special issue
SP - 440
EP - 456
SN - 2766-905X
DO - https://doi.org/10.59275/j.melba.2026-eec1
UR - https://melba-journal.org/2026:021
ER -