MONAI: An open-source framework for deep learning in healthcare
2022/11/04 by M. Jorge Cardoso, Cardoso, M. Jorge, Wenqi Li +118 · 140 citations
Computer Science · Medicine · #Artificial Intelligence (cs.AI) #Artificial Intelligence in Healthcare and Education #Computer Vision and Pattern Recognition (cs.CV) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Machine Learning in Healthcare #Radiomics and Machine Learning in Medical Imaging #cs.AI #cs.CV #cs.LG
paper · pdf · doi:10.48550/arxiv.2211.02701
www.monai.io
arxiv created 2022/11/04 · openalex publication_date 2022/11/04 · arxiv updated 2022/11/08 · openalex created_date 2025/10/10 · openalex updated_date 2026/07/30
Abstract
Artificial Intelligence (AI) is having a tremendous impact across most areas of science. Applications of AI in healthcare have the potential to improve our ability to detect, diagnose, prognose, and intervene on human disease. For AI models to be used clinically, they need to be made safe, reproducible and robust, and the underlying software framework must be aware of the particularities (e.g. geometry, physiology, physics) of medical data being processed. This work introduces MONAI, a freely available, community-supported, and consortium-led PyTorch-based framework for deep learning in healthcare. MONAI extends PyTorch to support medical data, with a particular focus on imaging, and provide purpose-specific AI model architectures, transformations and utilities that streamline the development and deployment of medical AI models. MONAI follows best practices for software-development, providing an easy-to-use, robust, well-documented, and well-tested software framework. MONAI preserves the simple, additive, and compositional approach of its underlying PyTorch libraries. MONAI is being used by and receiving contributions from research, clinical and industrial teams from around the world, who are pursuing applications spanning nearly every aspect of healthcare.
Cited by
- FluenceFormer: Transformer-Driven Multi-Beam Fluence Map Regression for Radiotherapy Planning
- Are Vision Foundation Models Foundational for Electron Microscopy Image Segmentation?
- Flash-CNNCap: Capacitance Extraction via Image Mapping
- When Average Calibration Fails: Site-Conditional Federated Conformal Risk Control
- AI for Mycetoma Diagnosis in Histopathological Images: The MICCAI 2024 Challenge
- BertsWin: Resolving Topological Sparsity in 3D Masked Autoencoders via Component-Balanced Structural Optimization
- SDUM: A Scalable Deep Unrolled Model for Universal MRI Reconstruction
- GANeXt: A Fully ConvNeXt-Enhanced Generative Adversarial Network for MRI- and CBCT-to-CT Synthesis
- Placenta Accreta Spectrum Detection Using an MRI-based Hybrid CNN-Transformer Model
- NodMAISI: Nodule-Oriented Medical AI for Synthetic Imaging
- A unified FLAIR hyperintensity segmentation model for various CNS tumor types and acquisition time points
- In search of truth: Evaluating concordance of AI-based anatomy segmentation models
- DBT-DINO
- Comparison of different segmentation algorithms on brain volume and fractal dimension in infant brain MRIs
- Kaapana: A Comprehensive Open-Source Platform for Integrating AI in Medical Imaging Research Environments
- Demo: Generative AI helps Radiotherapy Planning with User Preference
- Vision Foundry: A System for Training Foundational Vision AI Models
- Comparing Baseline and Day-1 Diffusion MRI Using Multimodal Deep Embeddings for Stroke Outcome Prediction
- TAP-CT: 3D Task-Agnostic Pretraining of Computed Tomography Foundation Models
- Hard Spatial Gating for Precision-Driven Brain Metastasis Segmentation: Addressing the Over-Segmentation Paradox in Deep Attention Networks
- BackSplit: The Importance of Sub-dividing the Background in Biomedical Lesion Segmentation
- NeuroVascU-Net: A Unified Multi-Scale and Cross-Domain Adaptive Feature Fusion U-Net for Precise 3D Segmentation of Brain Vessels in Contrast-Enhanced T1 MRI
- MAFM3: Modular Adaptation of Foundation Models for Multi-Modal Medical AI
- Cancer-Net PCa-MultiSeg: Multimodal Enhancement of Prostate Cancer Lesion Segmentation Using Synthetic Correlated Diffusion Imaging
- Anatomy-Aware Lymphoma Lesion Detection in Whole-Body PET/CT
- Submanifold Sparse Convolutional Networks for Automated 3D Segmentation of Kidneys and Kidney Tumours in Computed Tomography
- When Swin Transformer Meets KANs: An Improved Transformer Architecture for Medical Image Segmentation
- HERMES: A Hybrid Ensemble for Head-and-Neck Tumor Segmentation, TN Staging, and Recurrence-Free Survival on PET/CT
- Self-Supervised Text-Vision Alignment for Automated Brain MRI Abnormality Detection: A Multicenter Study (ALIGN Study)
- Diffusion-Driven Generation of Minimally Preprocessed Brain MRI
- Towards the Automatic Segmentation, Modeling and Meshing of the Aortic Vessel Tree from Multicenter Acquisitions: An Overview of the SEG.A. 2023 Segmentation of the Aorta Challenge
- Progressive Growing of Patch Size: Curriculum Learning for Accelerated and Improved Medical Image Segmentation
- SemiETPicker: Fast and Label-Efficient Particle Picking for CryoET Tomography Using Semi-Supervised Learning
- Better Tokens for Better 3D: Advancing Vision-Language Modeling in 3D Medical Imaging
- Learning from Disagreement: A Group Decision Simulation Framework for Robust Medical Image Segmentation
- BioVault: A privacy-first data visitation platform for equitable global collaboration in biomedicine
- A Biophysically-Conditioned Generative Framework for 3D Brain Tumor MRI Synthesis
- MIP-Based Tumor Segmentation: A Radiologist-Inspired Approach
- Random Window Augmentations for Deep Learning Robustness in CT and Liver Tumor Segmentation
- Limited-Angle Tomography Reconstruction via Projector Guided 3D Diffusion
- Fit Pixels, Get Labels: Meta-learned Implicit Networks for Image Segmentation
- Deep learning motion correction of quantitative stress perfusion cardiovascular magnetic resonance
- Latent Representation Learning from 3D Brain MRI for Interpretable Prediction in Multiple Sclerosis
- Pancreas Part Segmentation under Federated Learning Paradigm
- Confidence-Calibrating Regularization for Robust Brain MRI Segmentation Under Domain Shift
- COMPASS: Robust Feature Conformal Prediction for Medical Segmentation Metrics
- Sex-based Bias Inherent in the Dice Similarity Coefficient: A Model Independent Analysis for Multiple Anatomical Structures
- AuricularWorld: Hierarchical Action-Guided World Modeling for Fine-Grained Auricular Structure Segmentation from CT Scans
- RadHarmony: Radiological Data Handling in the Era of Agentic AI
- Evaluating Synthetic Data Generation for Domain Generalization in Fetal Brain MRI Segmentation
- Impact of partial volume correction on radiomics reproducibility in theranostic SPECT/CT imaging
- Automated Labeling of Intracranial Arteries with Uncertainty Quantification Using Deep Learning
- Unified Multimodal Coherent Field: Synchronous Semantic-Spatial-Vision Fusion for Brain Tumor Segmentation
- MRN: Harnessing 2D Vision Foundation Models for Diagnosing Parkinson's Disease with Limited 3D MR Data
- Opportunities and Challenges in Applying AI to Evolutionary Morphology
- Temporal Representation Learning of Phenotype Trajectories for pCR Prediction in Breast Cancer
- Radiology Report Conditional 3D CT Generation with Multi Encoder Latent diffusion Model
- ORCA: A Comprehensive AI-Driven Platform for Digital Pathology Analysis and Biomarker Discovery
- Physics-Informed Deep Unrolled Network for Portable MR Image Reconstruction
- Building a General SimCLR Self-Supervised Foundation Model Across Neurological Diseases to Advance 3D Brain MRI Diagnoses
- SSL-AD: Spatiotemporal Self-Supervised Learning for Generalizability and Adaptability Across Alzheimer's Prediction Tasks and Datasets
- Invisible Attributes, Visible Biases: Exploring Demographic Shortcuts in MRI-based Alzheimer's Disease Classification
- Enhancing 3D Medical Image Understanding with Pretraining Aided by 2D Multimodal Large Language Models
- Physics-Guided Diffusion Transformer with Spherical Harmonic Posterior Sampling for High-Fidelity Angular Super-Resolution in Diffusion MRI
- Vector-based loss functions for turbulent flow field inpainting
- SynBT: High-quality Tumor Synthesis for Breast Tumor Segmentation by 3D Diffusion Model
- Anisotropic Fourier Features for Positional Encoding in Medical Imaging
- A Modality-agnostic Multi-task Foundation Model for Human Brain Imaging
- Federated Fine-tuning of SAM-Med3D for MRI-based Dementia Classification
- Mask-Guided Multi-Channel SwinUNETR Framework for Robust MRI Classification
- Learning What is Worth Learning: Active and Sequential Domain Adaptation for Multi-modal Gross Tumor Volume Segmentation
- Random forest-based out-of-distribution detection for robust lung cancer segmentation
- Understanding Benefits and Pitfalls of Current Methods for the Segmentation of Undersampled MRI Data
- An MRI Atlas of the Human Fetal Brain: Reference and Segmentation Tools for Fetal Brain MRI Analysis
- UNICON: UNIfied CONtinual Learning for Medical Foundational Models
- Quantitative analysis of miniature synaptic calcium transients using positive unlabeled deep learning
- Subcortical Masks Generation in CT Images via Ensemble-Based Cross-Domain Label Transfer
- FunduSegmenter: Leveraging the RETFound Foundation Model for Joint Optic Disc and Optic Cup Segmentation in Retinal Fundus Images
- KonfAI: A Modular and Fully Configurable Framework for Deep Learning in Medical Imaging
- Automatic and standardized surgical reporting for central nervous system tumors
- Large-scale Multi-sequence Pretraining for Generalizable MRI Analysis in Versatile Clinical Applications
- MAISI-v2: Accelerated 3D High-Resolution Medical Image Synthesis with Rectified Flow and Region-specific Contrastive Loss
- Unmasking Interstitial Lung Diseases: Leveraging Masked Autoencoders for Diagnosis
- ERDES: A Benchmark Video Dataset for Retinal Detachment and Macular Status Classification in Ocular Ultrasound
- GRASPing Anatomy to Improve Pathology Segmentation
- Decoding the Alzheimer's Continuum: Interpretable Multi-Gate Routing for Diagnosis and Transition Prediction
- Register Anything: Estimating "Corresponding Prompts" for Segment Anything Model
- Skip priors and add graph-based anatomical information, for point-based Couinaud segmentation
- Consensus-Driven Active Model Selection
- MRpro - open PyTorch-based MR reconstruction and processing package
- Learning from Heterogeneous Structural MRI via Collaborative Domain Adaptation for Late-Life Depression Assessment
- Cyst-X: A Federated AI System Outperforms Clinical Guidelines to Detect Pancreatic Cancer Precursors and Reduce Unnecessary Surgery
- Rethink Domain Generalization in Heterogeneous Sequence MRI Segmentation
- A Study of Anatomical Priors for Deep Learning-Based Segmentation of Pheochromocytoma in Abdominal CT
- Parameter-Efficient Fine-Tuning of 3D DDPM for MRI Image Generation Using Tensor Networks
- Benchmarking GANs, Diffusion Models, and Flow Matching for T1w-to-T2w MRI Translation
- Interpretable Prediction of Lymph Node Metastasis in Rectal Cancer MRI Using Variational Autoencoders
- A Privacy-Preserving Federated Learning Framework for Generalizable CBCT to Synthetic CT Translation in Head and Neck
- A Unified Benchmark of Deep Learning Models for Multi-task 3D Brain Tumor Segmentation from Magnetic Resonance Imaging
- Knowledge-Guided Brain Tumor Segmentation via Synchronized Visual-Semantic-Topological Prior Fusion
- First Investigation of Deep Learning for Intraoperative Gauze Segmentation in Minimally Invasive Abdominal Surgery
- PanTS: The Pancreatic Tumor Segmentation Dataset
- SonoGym: High Performance Simulation for Challenging Surgical Tasks with Robotic Ultrasound
- ShapeKit
- Three-dimensional end-to-end deep learning for brain MRI analysis
- Mapping intratumoral heterogeneity through PET-derived washout and deep learning after proton therapy
- TextBraTS: Text-Guided Volumetric Brain Tumor Segmentation with Innovative Dataset Development and Fusion Module Exploration
- Classification of Multi-Parametric Body MRI Series Using Deep Learning
- D2Diff : A Dual Domain Diffusion Model for Accurate Multi-Contrast MRI Synthesis
- Improving Prostate Gland Segmentation Using Transformer based Architectures
- Modality-AGnostic Image Cascade (MAGIC) for Multi-Modality Cardiac Substructure Segmentation
- ADNP-15: An Open-Source Histopathological Dataset for Neuritic Plaque Segmentation in Human Brain Whole Slide Images with Frequency Domain Image Enhancement for Stain Normalization
- Text2CT: Towards 3D CT Volume Generation from Free-text Descriptions Using Diffusion Model
- Histo-Miner: Deep Learning based Tissue Features Extraction Pipeline from H&E Whole Slide Images of Cutaneous Squamous Cell Carcinoma
- Average Calibration Losses for Reliable Uncertainty in Medical Image Segmentation
- Medical World Model: Generative Simulation of Tumor Evolution for Treatment Planning
- MAIA: A Collaborative Medical AI Platform for Integrated Healthcare Innovation
- Towards Scalable Language-Image Pre-training for 3D Medical Imaging
- A Foundation Model Framework for Multi-View MRI Classification of Extramural Vascular Invasion and Mesorectal Fascia Invasion in Rectal Cancer
- TAGS: 3D Tumor-Adaptive Guidance for SAM
- CT-PrepAgent: Bounded Policy and Controlled Execution for Adaptive CT Data Preparation
- Automated Real-time Assessment of Intracranial Hemorrhage Detection AI Using an Ensembled Monitoring Model (EMM)
- Open‐source magnetic resonance imaging: Improving access, science, and education through global collaboration
- Generalizable cardiac substructures segmentation from contrast and non-contrast CTs using pretrained transformers
- VIViT: Variable-Input Vision Transformer Framework for 3D MR Image Segmentation
- Multi-Plane Vision Transformer for Hemorrhage Classification Using Axial and Sagittal MRI Data
- Semi-MedRef: Semi-Supervised Medical Referring Image Segmentation with Cross-Modal Alignment
- Multi-Scale Target-Aware Representation Learning for Fundus Image Enhancement
- SPROUT: A User-friendly, Scalable Toolkit for Multi-class Segmentation of Volumetric Images
- Multimodal Masked Autoencoder Pre-training for 3D MRI-Based Brain Tumor Analysis with Missing Modalities
- NeuroAgent: LLM Agents for Multimodal Neuroimaging Analysis and Research
- Innovative Integration of 4D Cardiovascular Reconstruction and Hologram: A New Visualization Tool for Coronary Artery Bypass Grafting Planning
- A tissue-informed deep learning-based method for positron range correction in preclinical 68Ga PET imaging
- Towards a deep learning approach for classifying treatment response in glioblastomas
- 3D Transport-based Morphometry (3D-TBM) for medical image analysis
- Enhancing Low Back Pain Assessment with Diffusion Models for Lumbar Spine MRI Segmentation
- Comprehensive Evaluation of Quantitative Measurements from Automated Deep Segmentations of PSMA PET/CT Images
- BrainMRDiff: A Diffusion Model for Anatomically Consistent Brain MRI Synthesis
- Latent Diffusion Autoencoders: Toward Efficient and Meaningful Unsupervised Representation Learning in Medical Imaging
- CTI-Unet: Cascaded Threshold Integration for Improved U-Net Segmentation of Pathology Images
Related