anatomical representation learning

Design and train models that encode anatomical structure into compact, interpretable latent representations from imaging or volumetric anatomical data, typically using variational autoencoders and their semi‑supervised variants. These representations are constructed to reconstruct or generate aligned segmentation masks, be conditioned on clinical priors, and to disentangle static anatomy from temporal or dynamic factors for downstream analysis and tasks.

anatomicalrepresentationlearning

Recent Skill Trend

Momentum and market value over time
Trending
Score
No comparison yet
-0.54
Oct 01, 2026Oct 01, 2026
Career
Value
No comparison yet
$200K/year
Oct 01, 2026Oct 01, 2026

Recommended Survey Paper

Quick overview of the field
View more

Must-Read Papers

Most classic and influential ideas
View more

Latent space projections and atlases: A cautionary tale in deep neuroimaging using autoencoders

Sep 03, 2025
JM
J. M. Gorriz
🏛️ University of Granada | University of Cambridge

Early detection of Alzheimer’s disease (AD) remains challenging due to the difficulty in identifying interpretable, imaging-based biomarkers from structural MRI. Method: We propose an interpretable unsupervised deep learning framework centered on a lightweight 3D convolutional autoencoder that learns compact latent representations of MRI data. Multi-stage dimensionality reduction—integrating PCA and UMAP—is combined with the AAL brain atlas for neuroanatomically grounded visualization. Critically, we introduce Latent Region Correlation Profiling (LRCP), a novel analytical framework that jointly applies SHAP-based regression and assumption-free statistical testing to quantify region-specific contributions to cognitive status variability. Results: Our lightweight model robustly captures AD-progressive anatomical patterns without supervision. LRCP substantially enhances both clinical interpretability and neuroanatomical fidelity of latent features, enabling precise identification of AD-relevant brain regions. This framework establishes a new paradigm for unsupervised AD biomarker discovery—balancing discriminative power with biological interpretability.

Analyzing 3D brain MRI data using unsupervised autoencoders for neuroanatomical patternsIdentifying clinically relevant brain regions correlated with Alzheimer's disease progressionValidating interpretability of latent features through statistical and SHAP-based methods

This work addresses the scarcity of high-quality brain segmentation masks in non-contrast CT neuroimaging, particularly due to the high annotation cost and substantial variability of ischemic infarcts. To tackle this challenge, the authors propose an anatomy-preserving generative framework that, for the first time, integrates a diffusion model into the latent space of a variational autoencoder (VAE) trained on segmentation masks. The method enables unconditional generation of multi-class brain tissue masks containing ischemic infarcts and supports coarse control over lesion presence via binary prompts. By leveraging a frozen VAE decoder to reconstruct masks, the approach effectively preserves global anatomical structure, discrete semantic labels, and realistic pathological variations while avoiding common structural artifacts associated with pixel-level generative models.

data scarcityischemic infarctmedical image analysis

A Review of Latent Representation Models in Neuroimaging

Dec 24, 2024
CV
C. V'azquez-Garc'ia
🏛️ University of Granada

Extracting biologically meaningful and clinically interpretable representations from high-dimensional neuroimaging data (e.g., MRI/PET) remains challenging due to inherent complexity and limited interpretability of latent features. Method: This study systematically reviews and empirically evaluates generative latent-variable models—including autoencoders, GANs, and latent diffusion models (LDMs)—across two complementary pathways: clinical neuroimaging and computational neuroscience. It pioneers the integration of predictive coding theory with deep generative modeling to establish a multimodal alignment and interpretable latent-space analysis framework, accompanied by a cross-model performance evaluation protocol. Contribution/Results: The work delineates the applicability boundaries of each model class for Alzheimer’s disease and Parkinson’s disease subtyping, longitudinal tracking, and brain-age estimation. It significantly enhances the biological interpretability and clinical transferability of learned latent representations, providing a methodological foundation for interpretable brain-computational modeling.

Brain FunctionNeuroimagingPathological Changes

Steerable Anatomical Shape Synthesis with Implicit Neural Representations

Apr 04, 2025
BD
B. D. Wilde
🏛️ University of Twente

In virtual imaging trials, generating anatomically accurate, clinically relevant patient-specific phantoms with controllable population-level anatomical variations remains challenging. Method: We propose the first implicit neural representation framework for editable anatomical modeling, integrating geometry-prior-guided implicit surface reconstruction, disentangled latent space learning, and topology-adaptive deformation—enabling fine-grained, target-specific morphological editing of topologically variable organs (e.g., thyroid). Contribution/Results: Our approach is the first to achieve explicit shape–topology disentanglement in anatomical implicit neural representations, supporting clinically interpretable, parameterized editing. Quantitative and qualitative evaluations demonstrate state-of-the-art performance in reconstruction accuracy and anatomical plausibility. Generated phantoms exhibit high fidelity, clinical interpretability, and strong controllability—facilitating reproducible, patient-population-aware virtual imaging studies.

Generative modeling for anatomical structures in virtual imaging trialsImplicit neural representations for topology-varying anatomical structuresTargeted control to simulate specific patient populations

This study investigates the structure and information content of latent representations in 3D brain MRI generative models, with a focus on their efficacy for clinical discrimination of Down syndrome. Employing various variational autoencoder (VAE) architectures, we compress 3D brain MRIs into compact latent codes that enable high-fidelity reconstruction while supporting downstream multitask analysis. Through principal component analysis visualization and systematic evaluation, we demonstrate that the learned latent space clearly clusters individuals with Down syndrome and neurotypical controls, exhibiting strong discriminative power and interpretability. These findings validate the potential of such latent representations for clinical neuroimaging applications and offer a novel approach to disease representation learning based on generative models.

3D brain MRIDown syndromedownstream analysis

Latest Papers

What's happening recently
View more

This work addresses the limitations of traditional statistical shape modeling, which relies on dense annotations and fixed latent representations, thereby struggling to flexibly capture complex anatomical variations. The authors propose MorphoFlow, a framework that learns compact probabilistic shape representations from only sparse surface annotations. MorphoFlow integrates neural implicit representations, a self-decoder architecture, and autoregressive normalizing flows, augmented with an adaptive latent correlation weighting mechanism. This mechanism leverages a sparsity-inducing prior to automatically modulate the contribution of each latent dimension to anatomical variability, eliminating the need for manual hyperparameter tuning. The method enables high-resolution 3D shape generation and uncertainty quantification. Evaluated on lumbar spine and femur datasets, MorphoFlow achieves high-fidelity reconstructions and accurately recovers population-consistent, structured patterns of anatomical variation.

3D shape reconstructionanatomical variabilitylatent representation

This study addresses the challenge of effectively integrating structural and functional neuroimaging data by proposing a multimodal graph variational autoencoder (gMMVAE). The method introduces a modality-aware graph encoding mechanism that maps gray matter volume and static functional connectivity into a unified low-dimensional latent space. It further provides a systematic comparison of diverse generative architectures—including VAEs, Transformers, GANs, and diffusion models—in modeling graph-structured brain data. Experimental results demonstrate that gMMVAE consistently outperforms existing approaches in terms of generation fidelity, reconstruction quality, computational efficiency, and discriminative power of the latent representation. This work thus establishes an efficient and interpretable generative modeling paradigm for multimodal brain network analysis.

functional connectivitygenerative AIgraph encoding

This study addresses the limitation of existing accelerated multi-contrast MRI reconstruction methods, which typically process each contrast independently and fail to fully exploit shared anatomical information across contrasts. To overcome this, we propose MAX, a framework that learns subject-specific anatomical representations via manifold expansion for efficient reconstruction. By employing intensity augmentation to expand the manifold, MAX effectively decouples shared anatomical structures from contrast-dependent components, integrating decoupled implicit neural representations with unrolled optimization techniques. Experimental results demonstrate that MAX achieves state-of-the-art PSNR and SSIM performance in both brain and knee MRI reconstruction, yielding improvements exceeding 1 dB. Furthermore, the proposed method exhibits significant robustness against motion artifacts and noise, highlighting its potential for reliable clinical application.

accelerated MRIanatomical representationimage reconstruction

Existing self-supervised methods for ultrasound imaging often neglect anatomical context, hindering the learning of clinically aligned representations. This work proposes ANAUS, a novel framework that, for the first time, leverages anatomical structures as anchors in self-supervised learning. ANAUS introduces a learnable latent prompt engine combined with one-shot domain adaptation to enable annotation-free anatomical segmentation. It further incorporates a dual-strategy self-supervised mechanism—cross-view anatomical region semantic alignment and masked reconstruction of contextual core regions—to enhance representation invariance and fine-grained detail perception. Evaluated across six public datasets, ANAUS significantly outperforms state-of-the-art methods while maintaining computational efficiency suitable for clinical deployment.

anatomical contextmedical imagingrepresentation learning

Hot Scholars

JY

Joseph Y. Lo

Professor of Radiology, Biomed. Engineering, Elec. Engineering, Med Physics
medical imagingmachine learning
FI

Fakrul Islam Tushar

PhD Student at ECE, Duke University
DigitalTwinVirtual Imaging TrialsComputer VisionMulti-model AI
QZ

Qingyu Zhao

Assistant Professor, WCM, Cornell
Machine LearningNeuroimagingComputational Neuroscience
PL

Peirong Liu

Assistant Professor of ECE, Johns Hopkins University
AI for HealthcareComputer VisionMedical Imaging