TCAM-Diff: Triplane-Aware Cross-Attention Medical Diffusion Model
Zhenkai Zhang, Krista A. Ehinger, Tom Drummond
摘要
We introduce TCAM-Diff, a novel 3D medical image generation model that reduces the memory requirements to encode and generate high-resolution 3D data. This model utilizes a decoder-only autoencoder method to learn triplane representation from dense volume and leverages generalization operations to prevent overfitting. Subsequently, it uses a triplaneaware cross-attention diffusion model to learn and integrate these features effectively. Furthermore, the features generated by the diffusion model can be rapidly transformed into 3D volumes using a pre-trained decoder module. Our experiments on three different scales of medical datasets, BrainTumour 128 × 128 × 128, Pancreas 256 × 256 × 256, and Colon 512×512×512, demonstrate outstanding results. We utilized MSE and SSIM to assess reconstruction quality and leveraged the Wasserstein Generative Adversarial Network (W-GAN) critic to assess generative quality. Comparisons with existing approaches show that our method gives better reconstruction and generation results than other encoder-decoder methods with similar sized latent spaces.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper1
问问它们各自怎么用它它引用的顶会 Paper6
- High-Resolution Image Synthesis with Latent Diffusion ModelsRobin Rombach, Andreas Blattmann, Dominik Lorenz, Patrick Esser 等CVPR 2022 · 被引用 13,123 次
- Efficient Geometry-aware 3D Generative Adversarial NetworksEric R. Chan, Connor Z. Lin, Matthew A. Chan, Koki Nagano 等CVPR 2022 · 被引用 984 次
- Autodecoding Latent 3D Diffusion ModelsEvangelos Ntavelis, Aliaksandr Siarohin, Kyle Olszewski, Chaoyang Wang 等NeurIPS 2023 · 被引用 65 次
- Class-Balancing Diffusion ModelsYiming Qin, Huangjie Zheng, Jiangchao Yao, Mingyuan Zhou 等CVPR 2023
- 3D Neural Field Generation Using Triplane DiffusionJ. Ryan Shue, Eric Ryan Chan, Ryan Po, Zachary Ankner 等CVPR 2023
相关 Paper
- GuideGen: A Text-Guided Framework for Paired Full-torso Anatomy and CT Volume GenerationLinrui Dai, Rongzhao Zhang, Yongrui Yu, Xiaofan ZhangAAAI 2026
- Improving 3D Imaging with Pre-Trained Perpendicular 2D Diffusion ModelsSuhyeon Lee, Hyungjin Chung, Minyoung Park, Jonghyuk Park 等ICCV 2023 · 被引用 72 次
- DMV3D: Denoising Multi-view Diffusion Using 3D Large Reconstruction ModelYinghao Xu, Hao Tan, Fujun Luan, Sai Bi 等ICLR 2024 · 被引用 234 次
- DiffusionBlend: Learning 3D Image Prior through Position-aware Diffusion Score Blending for 3D Computed Tomography ReconstructionBowen Song, Jason Hu, Zhaoxu Luo, Jeffrey A. Fessler 等NeurIPS 2024 · 被引用 20 次
- Foundation VAE for CT Reconstruction, Augmentation, and GenerationQi Chen, Shuhan Ding, Yu Gu, Nan Liu 等ICML 2026 · 被引用 1 次
