Can Shape-Infused Joint Embeddings Improve Image-Conditioned 3D Diffusion?
Fuente:
arXiv
Saved in:
| Main Authors: | Sbrolli, Cristian, Cudrano, Paolo, Matteucci, Matteo |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
No Captions, No Problem: Captionless 3D-CLIP Alignment with Hard Negatives via CLIP Knowledge and LLMs
by: Sbrolli, Cristian, et al.
Published: (2024)
by: Sbrolli, Cristian, et al.
Published: (2024)
Auto-Comp: An Automated Pipeline for Scalable Compositional Probing of Contrastive Vision-Language Models
by: Sbrolli, Cristian, et al.
Published: (2026)
by: Sbrolli, Cristian, et al.
Published: (2026)
Neuro-Symbolic Scene Graph Conditioning for Synthetic Image Dataset Generation
by: Savazzi, Giacomo, et al.
Published: (2025)
by: Savazzi, Giacomo, et al.
Published: (2025)
Your Image Generator Is Your New Private Dataset
by: Resmini, Nicolo, et al.
Published: (2025)
by: Resmini, Nicolo, et al.
Published: (2025)
SCENEFORGE: Enhancing 3D-text alignment with Structured Scene Compositions
by: Sbrolli, Cristian, et al.
Published: (2025)
by: Sbrolli, Cristian, et al.
Published: (2025)
PolyGen: Fully Synthetic Vision-Language Training via Multi-Generator Ensembles
by: Brusini, Leonardo, et al.
Published: (2026)
by: Brusini, Leonardo, et al.
Published: (2026)
The Empirical Impact of Forgetting and Transfer in Continual Visual Odometry
by: Cudrano, Paolo, et al.
Published: (2024)
by: Cudrano, Paolo, et al.
Published: (2024)
Few Shot Semantic Segmentation: a review of methodologies, benchmarks, and open challenges
by: Catalano, Nico, et al.
Published: (2023)
by: Catalano, Nico, et al.
Published: (2023)
Stable Diffusion Dataset Generation for Downstream Classification Tasks
by: Lomurno, Eugenio, et al.
Published: (2024)
by: Lomurno, Eugenio, et al.
Published: (2024)
Synthetic Image Learning: Preserving Performance and Preventing Membership Inference Attacks
by: Lomurno, Eugenio, et al.
Published: (2024)
by: Lomurno, Eugenio, et al.
Published: (2024)
Federated Knowledge Recycling: Privacy-Preserving Synthetic Data Sharing
by: Lomurno, Eugenio, et al.
Published: (2024)
by: Lomurno, Eugenio, et al.
Published: (2024)
Audio-Infused Automatic Image Colorization by Exploiting Audio Scene Semantics
by: Zhao, Pengcheng, et al.
Published: (2024)
by: Zhao, Pengcheng, et al.
Published: (2024)
Representing 3D Shapes With 64 Latent Vectors for 3D Diffusion Models
by: Cho, In, et al.
Published: (2025)
by: Cho, In, et al.
Published: (2025)
Image-Conditional Diffusion Transformer for Underwater Image Enhancement
by: Nie, Xingyang, et al.
Published: (2024)
by: Nie, Xingyang, et al.
Published: (2024)
POMONAG: Pareto-Optimal Many-Objective Neural Architecture Generator
by: Lomurno, Eugenio, et al.
Published: (2024)
by: Lomurno, Eugenio, et al.
Published: (2024)
MultiDiffSense: Diffusion-Based Multi-Modal Visuo-Tactile Image Generation Conditioned on Object Shape and Contact Pose
by: Bhouri, Sirine, et al.
Published: (2026)
by: Bhouri, Sirine, et al.
Published: (2026)
Uncovering the Text Embedding in Text-to-Image Diffusion Models
by: Yu, Hu, et al.
Published: (2024)
by: Yu, Hu, et al.
Published: (2024)
A Lightweight Neural Architecture Search Model for Medical Image Classification
by: Xie, Lunchen, et al.
Published: (2024)
by: Xie, Lunchen, et al.
Published: (2024)
MetaVoxel: Joint Diffusion Modeling of Imaging and Clinical Metadata
by: Liu, Yihao, et al.
Published: (2025)
by: Liu, Yihao, et al.
Published: (2025)
On Improved Conditioning Mechanisms and Pre-training Strategies for Diffusion Models
by: Ifriqi, Tariq Berrada, et al.
Published: (2024)
by: Ifriqi, Tariq Berrada, et al.
Published: (2024)
InpDiffusion: Image Inpainting Localization via Conditional Diffusion Models
by: Wang, Kai, et al.
Published: (2025)
by: Wang, Kai, et al.
Published: (2025)
A Structured Benchmark for Text-Guided Anomaly Detection: When Language Stops Conditioning the Decision
by: Samele, Stefano, et al.
Published: (2026)
by: Samele, Stefano, et al.
Published: (2026)
ShapeShifter: 3D Variations Using Multiscale and Sparse Point-Voxel Diffusion
by: Maruani, Nissim, et al.
Published: (2025)
by: Maruani, Nissim, et al.
Published: (2025)
Co-generation of Layout and Shape from Text via Autoregressive 3D Diffusion
by: Tang, Zhenggang, et al.
Published: (2026)
by: Tang, Zhenggang, et al.
Published: (2026)
Identifying and Solving Conditional Image Leakage in Image-to-Video Diffusion Model
by: Zhao, Min, et al.
Published: (2024)
by: Zhao, Min, et al.
Published: (2024)
Conditional Diffusion Model for Longitudinal Medical Image Generation
by: Dao, Duy-Phuong, et al.
Published: (2024)
by: Dao, Duy-Phuong, et al.
Published: (2024)
Conditional Image Synthesis with Diffusion Models: A Survey
by: Zhan, Zheyuan, et al.
Published: (2024)
by: Zhan, Zheyuan, et al.
Published: (2024)
Graph Conditioned Diffusion for Controllable Histopathology Image Generation
by: Cechnicka, Sarah, et al.
Published: (2025)
by: Cechnicka, Sarah, et al.
Published: (2025)
Magic-Boost: Boost 3D Generation with Multi-View Conditioned Diffusion
by: Yang, Fan, et al.
Published: (2024)
by: Yang, Fan, et al.
Published: (2024)
CoDiff: Conditional Diffusion Model for Collaborative 3D Object Detection
by: Huang, Zhe, et al.
Published: (2025)
by: Huang, Zhe, et al.
Published: (2025)
3D-Consistent Image Inpainting with Diffusion Models
by: Antsfeld, Leonid, et al.
Published: (2024)
by: Antsfeld, Leonid, et al.
Published: (2024)
LTM3D: Bridging Token Spaces for Conditional 3D Generation with Auto-Regressive Diffusion Framework
by: Kang, Xin, et al.
Published: (2025)
by: Kang, Xin, et al.
Published: (2025)
Fusion Embedding for Pose-Guided Person Image Synthesis with Diffusion Model
by: Lee, Donghwna, et al.
Published: (2024)
by: Lee, Donghwna, et al.
Published: (2024)
Tooth-Diffusion: Guided 3D CBCT Synthesis with Fine-Grained Tooth Conditioning
by: Said, Said Djafar, et al.
Published: (2025)
by: Said, Said Djafar, et al.
Published: (2025)
Bidirectional Mammogram View Translation with Column-Aware and Implicit 3D Conditional Diffusion
by: Li, Xin, et al.
Published: (2025)
by: Li, Xin, et al.
Published: (2025)
MINT: Memory-Infused Prompt Tuning at Test-time for CLIP
by: Yi, Jiaming, et al.
Published: (2025)
by: Yi, Jiaming, et al.
Published: (2025)
Image-Conditioned 3D Gaussian Splat Quantization
by: Liu, Xinshuang, et al.
Published: (2025)
by: Liu, Xinshuang, et al.
Published: (2025)
Enhancing Text-to-Image Diffusion Transformer via Split-Text Conditioning
by: Zhang, Yu, et al.
Published: (2025)
by: Zhang, Yu, et al.
Published: (2025)
Video Representation Learning with Joint-Embedding Predictive Architectures
by: Drozdov, Katrina, et al.
Published: (2024)
by: Drozdov, Katrina, et al.
Published: (2024)
Can Diffusion Models Learn Hidden Inter-Feature Rules Behind Images?
by: Han, Yujin, et al.
Published: (2025)
by: Han, Yujin, et al.
Published: (2025)
Similar Items
-
No Captions, No Problem: Captionless 3D-CLIP Alignment with Hard Negatives via CLIP Knowledge and LLMs
by: Sbrolli, Cristian, et al.
Published: (2024) -
Auto-Comp: An Automated Pipeline for Scalable Compositional Probing of Contrastive Vision-Language Models
by: Sbrolli, Cristian, et al.
Published: (2026) -
Neuro-Symbolic Scene Graph Conditioning for Synthetic Image Dataset Generation
by: Savazzi, Giacomo, et al.
Published: (2025) -
Your Image Generator Is Your New Private Dataset
by: Resmini, Nicolo, et al.
Published: (2025) -
SCENEFORGE: Enhancing 3D-text alignment with Structured Scene Compositions
by: Sbrolli, Cristian, et al.
Published: (2025)