Skip Mamba Diffusion for Monocular 3D Semantic Scene Completion
Fuente:
arXiv
Salvato in:
| Autori principali: | Liang, Li, Akhtar, Naveed, Vice, Jordan, Kong, Xiangrui, Mian, Ajmal Saeed |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Manipulating and Mitigating Generative Model Biases without Retraining
di: Vice, Jordan, et al.
Pubblicazione: (2024)
di: Vice, Jordan, et al.
Pubblicazione: (2024)
Exploring Bias in over 100 Text-to-Image Generative Models
di: Vice, Jordan, et al.
Pubblicazione: (2025)
di: Vice, Jordan, et al.
Pubblicazione: (2025)
Safety Without Semantic Disruptions: Editing-free Safe Image Generation via Context-preserving Dual Latent Reconstruction
di: Vice, Jordan, et al.
Pubblicazione: (2024)
di: Vice, Jordan, et al.
Pubblicazione: (2024)
CymbaDiff: Structured Spatial Diffusion for Sketch-based 3D Semantic Urban Scene Generation
di: Liang, Li, et al.
Pubblicazione: (2025)
di: Liang, Li, et al.
Pubblicazione: (2025)
On the Fairness, Diversity and Reliability of Text-to-Image Generative Models
di: Vice, Jordan, et al.
Pubblicazione: (2024)
di: Vice, Jordan, et al.
Pubblicazione: (2024)
Multistream Network for LiDAR and Camera-based 3D Object Detection in Outdoor Scenes
di: Ibrahim, Muhammad, et al.
Pubblicazione: (2025)
di: Ibrahim, Muhammad, et al.
Pubblicazione: (2025)
On the Reliability of Vision-Language Models Under Adversarial Frequency-Domain Perturbations
di: Vice, Jordan, et al.
Pubblicazione: (2025)
di: Vice, Jordan, et al.
Pubblicazione: (2025)
Mitigating Memorization in Text-to-Image Diffusion via Region-Aware Prompt Augmentation and Multimodal Copy Detection
di: Chen, Yunzhuo, et al.
Pubblicazione: (2026)
di: Chen, Yunzhuo, et al.
Pubblicazione: (2026)
Attribution-Guided Model Rectification of Unreliable Neural Network Behaviors
di: Yang, Peiyu, et al.
Pubblicazione: (2026)
di: Yang, Peiyu, et al.
Pubblicazione: (2026)
Simultaneous Multiple Object Detection and Pose Estimation using 3D Model Infusion with Monocular Vision
di: Li, Congliang, et al.
Pubblicazione: (2022)
di: Li, Congliang, et al.
Pubblicazione: (2022)
Auto-Regressive Diffusion for Generating 3D Human-Object Interactions
di: Geng, Zichen, et al.
Pubblicazione: (2025)
di: Geng, Zichen, et al.
Pubblicazione: (2025)
Text-guided 3D Human Motion Generation with Keyframe-based Parallel Skip Transformer
di: Geng, Zichen, et al.
Pubblicazione: (2024)
di: Geng, Zichen, et al.
Pubblicazione: (2024)
Enhancing Monocular 3D Scene Completion with Diffusion Model
di: Song, Changlin, et al.
Pubblicazione: (2025)
di: Song, Changlin, et al.
Pubblicazione: (2025)
D3Seg: Dependency-Aware Diffusion for Brain Tumor Segmentation with Missing Modalities
di: Ali, Danish, et al.
Pubblicazione: (2026)
di: Ali, Danish, et al.
Pubblicazione: (2026)
DRBD-Mamba for Robust and Efficient Brain Tumor Segmentation with Analytical Insights
di: Ali, Danish, et al.
Pubblicazione: (2025)
di: Ali, Danish, et al.
Pubblicazione: (2025)
Efficient Diffusion Models for Vision: A Survey
di: Ulhaq, Anwaar, et al.
Pubblicazione: (2022)
di: Ulhaq, Anwaar, et al.
Pubblicazione: (2022)
Artificial intelligence techniques in inherited retinal diseases: A review
di: Trinh, Han, et al.
Pubblicazione: (2024)
di: Trinh, Han, et al.
Pubblicazione: (2024)
One Step Closer: Creating the Future to Boost Monocular Semantic Scene Completion
di: Lu, Haoang, et al.
Pubblicazione: (2025)
di: Lu, Haoang, et al.
Pubblicazione: (2025)
Deeper Diffusion Models Amplify Bias
di: Hakemi, Shahin, et al.
Pubblicazione: (2025)
di: Hakemi, Shahin, et al.
Pubblicazione: (2025)
UniHM: Universal Human Motion Generation with Object Interactions in Indoor Scenes
di: Geng, Zichen, et al.
Pubblicazione: (2025)
di: Geng, Zichen, et al.
Pubblicazione: (2025)
Plug-and-Play Interpretable Responsible Text-to-Image Generation via Dual-Space Multi-facet Concept Control
di: Azam, Basim, et al.
Pubblicazione: (2025)
di: Azam, Basim, et al.
Pubblicazione: (2025)
Suitability of KANs for Computer Vision: A preliminary investigation
di: Azam, Basim, et al.
Pubblicazione: (2024)
di: Azam, Basim, et al.
Pubblicazione: (2024)
FastBO: Fast HPO and NAS with Adaptive Fidelity Identification
di: Jiang, Jiantong, et al.
Pubblicazione: (2024)
di: Jiang, Jiantong, et al.
Pubblicazione: (2024)
Mantis: Mamba-native Tuning is Efficient for 3D Point Cloud Foundation Models
di: Guo, Zihao, et al.
Pubblicazione: (2026)
di: Guo, Zihao, et al.
Pubblicazione: (2026)
Disentangled Hierarchical VAE for 3D Human-Human Interaction Generation
di: Geng, Zichen, et al.
Pubblicazione: (2026)
di: Geng, Zichen, et al.
Pubblicazione: (2026)
Single-weight Model Editing for Post-hoc Spurious Correlation Neutralization
di: Hakemi, Shahin, et al.
Pubblicazione: (2025)
di: Hakemi, Shahin, et al.
Pubblicazione: (2025)
TextMamba: Scene Text Detector with Mamba
di: Zhao, Qiyan, et al.
Pubblicazione: (2025)
di: Zhao, Qiyan, et al.
Pubblicazione: (2025)
Dynamic watermarks in images generated by diffusion models
di: Chen, Yunzhuo, et al.
Pubblicazione: (2025)
di: Chen, Yunzhuo, et al.
Pubblicazione: (2025)
Deepfake Detection with Spatio-Temporal Consistency and Attention
di: Chen, Yunzhuo, et al.
Pubblicazione: (2025)
di: Chen, Yunzhuo, et al.
Pubblicazione: (2025)
4D3R: Motion-Aware Neural Reconstruction and Rendering of Dynamic Scenes from Monocular Videos
di: Guo, Mengqi, et al.
Pubblicazione: (2025)
di: Guo, Mengqi, et al.
Pubblicazione: (2025)
Class-Partitioned VQ-VAE and Latent Flow Matching for Point Cloud Scene Generation
di: Edirimuni, Dasith de Silva, et al.
Pubblicazione: (2026)
di: Edirimuni, Dasith de Silva, et al.
Pubblicazione: (2026)
HexPlane Representation for 3D Semantic Scene Understanding
di: Chen, Zeren, et al.
Pubblicazione: (2025)
di: Chen, Zeren, et al.
Pubblicazione: (2025)
MetaSSC: Enhancing 3D Semantic Scene Completion for Autonomous Driving through Meta-Learning and Long-sequence Modeling
di: Qu, Yansong, et al.
Pubblicazione: (2024)
di: Qu, Yansong, et al.
Pubblicazione: (2024)
RoomCraft: Controllable and Complete 3D Indoor Scene Generation
di: Zhou, Mengqi, et al.
Pubblicazione: (2025)
di: Zhou, Mengqi, et al.
Pubblicazione: (2025)
GeoSAM-3D: Geodesic Prompt Propagation for Open-Vocabulary 3D Scene Segmentation from Monocular Video
di: Sharma, Arun
Pubblicazione: (2026)
di: Sharma, Arun
Pubblicazione: (2026)
LT3SD: Latent Trees for 3D Scene Diffusion
di: Meng, Quan, et al.
Pubblicazione: (2024)
di: Meng, Quan, et al.
Pubblicazione: (2024)
PaSCo: Urban 3D Panoptic Scene Completion with Uncertainty Awareness
di: Cao, Anh-Quan, et al.
Pubblicazione: (2023)
di: Cao, Anh-Quan, et al.
Pubblicazione: (2023)
DefAn: Definitive Answer Dataset for LLMs Hallucination Evaluation
di: Rahman, A B M Ashikur, et al.
Pubblicazione: (2024)
di: Rahman, A B M Ashikur, et al.
Pubblicazione: (2024)
Scene Splatter: Momentum 3D Scene Generation from Single Image with Video Diffusion Model
di: Zhang, Shengjun, et al.
Pubblicazione: (2025)
di: Zhang, Shengjun, et al.
Pubblicazione: (2025)
Visual Attention Methods in Deep Learning: An In-Depth Survey
di: Hassanin, Mohammed, et al.
Pubblicazione: (2022)
di: Hassanin, Mohammed, et al.
Pubblicazione: (2022)
Documenti analoghi
-
Manipulating and Mitigating Generative Model Biases without Retraining
di: Vice, Jordan, et al.
Pubblicazione: (2024) -
Exploring Bias in over 100 Text-to-Image Generative Models
di: Vice, Jordan, et al.
Pubblicazione: (2025) -
Safety Without Semantic Disruptions: Editing-free Safe Image Generation via Context-preserving Dual Latent Reconstruction
di: Vice, Jordan, et al.
Pubblicazione: (2024) -
CymbaDiff: Structured Spatial Diffusion for Sketch-based 3D Semantic Urban Scene Generation
di: Liang, Li, et al.
Pubblicazione: (2025) -
On the Fairness, Diversity and Reliability of Text-to-Image Generative Models
di: Vice, Jordan, et al.
Pubblicazione: (2024)