MirrorVerse: Pushing Diffusion Models to Realistically Reflect the World
Fuente:
arXiv
Saved in:
| Main Authors: | Dhiman, Ankit, Shah, Manan, Babu, R Venkatesh |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Reflecting Reality: Enabling Diffusion Models to Produce Faithful Mirror Reflections
by: Dhiman, Ankit, et al.
Published: (2024)
by: Dhiman, Ankit, et al.
Published: (2024)
ChromaDistill: Colorizing Monochrome Radiance Fields with Knowledge Distillation
by: Dhiman, Ankit, et al.
Published: (2023)
by: Dhiman, Ankit, et al.
Published: (2023)
UniC-Lift: Unified 3D Instance Segmentation via Contrastive Learning
by: Dhiman, Ankit, et al.
Published: (2025)
by: Dhiman, Ankit, et al.
Published: (2025)
VerseCrafter: Dynamic Realistic Video World Model with 4D Geometric Control
by: Zheng, Sixiao, et al.
Published: (2026)
by: Zheng, Sixiao, et al.
Published: (2026)
Turbo-GS: Accelerating 3D Gaussian Fitting for High-Quality Radiance Fields
by: Dhiman, Ankit, et al.
Published: (2024)
by: Dhiman, Ankit, et al.
Published: (2024)
Harnessing Diffusion-Generated Synthetic Images for Fair Image Classification
by: Basu, Abhipsa, et al.
Published: (2025)
by: Basu, Abhipsa, et al.
Published: (2025)
PreciseControl: Enhancing Text-To-Image Diffusion Models with Fine-Grained Attribute Control
by: Parihar, Rishubh, et al.
Published: (2024)
by: Parihar, Rishubh, et al.
Published: (2024)
MirrorGaussian: Reflecting 3D Gaussians for Reconstructing Mirror Reflections
by: Liu, Jiayue, et al.
Published: (2024)
by: Liu, Jiayue, et al.
Published: (2024)
Do Vision Language Models Need to Process Image Tokens?
by: Ghosh, Sambit, et al.
Published: (2026)
by: Ghosh, Sambit, et al.
Published: (2026)
Balancing Act: Distribution-Guided Debiasing in Diffusion Models
by: Parihar, Rishubh, et al.
Published: (2024)
by: Parihar, Rishubh, et al.
Published: (2024)
DeepVerse: 4D Autoregressive Video Generation as a World Model
by: Chen, Junyi, et al.
Published: (2025)
by: Chen, Junyi, et al.
Published: (2025)
NeoVerse: Enhancing 4D World Model with in-the-wild Monocular Videos
by: Yang, Yuxue, et al.
Published: (2026)
by: Yang, Yuxue, et al.
Published: (2026)
DynamicVerse: A Physically-Aware Multimodal Framework for 4D World Modeling
by: Wen, Kairun, et al.
Published: (2025)
by: Wen, Kairun, et al.
Published: (2025)
Leveraging Vision-Language Models for Improving Domain Generalization in Image Classification
by: Addepalli, Sravanti, et al.
Published: (2023)
by: Addepalli, Sravanti, et al.
Published: (2023)
PROPEX-RAG: Enhanced GraphRAG using Prompt-Driven Prompt Execution
by: Sarnaik, Tejas, et al.
Published: (2025)
by: Sarnaik, Tejas, et al.
Published: (2025)
BiDM: Pushing the Limit of Quantization for Diffusion Models
by: Zheng, Xingyu, et al.
Published: (2024)
by: Zheng, Xingyu, et al.
Published: (2024)
CogniVerse: Revolutionizing Multi-Modal Retrieval-Augmented Generation with Cognitive Reflection and Geometric Reasoning
by: Fang, Xiang, et al.
Published: (2026)
by: Fang, Xiang, et al.
Published: (2026)
GameVerse: Can Vision-Language Models Learn from Video-based Reflection?
by: Zhang, Kuan, et al.
Published: (2026)
by: Zhang, Kuan, et al.
Published: (2026)
GeoDiv: Framework For Measuring Geographical Diversity In Text-To-Image Models
by: Basu, Abhipsa, et al.
Published: (2026)
by: Basu, Abhipsa, et al.
Published: (2026)
UniVerse: Unleashing the Scene Prior of Video Diffusion Models for Robust Radiance Field Reconstruction
by: Cao, Jin, et al.
Published: (2025)
by: Cao, Jin, et al.
Published: (2025)
VideoVerse: Does Your T2V Generator Have World Model Capability to Synthesize Videos?
by: Wang, Zeqing, et al.
Published: (2025)
by: Wang, Zeqing, et al.
Published: (2025)
Mirror-3DGS: Incorporating Mirror Reflections into 3D Gaussian Splatting
by: Meng, Jiarui, et al.
Published: (2024)
by: Meng, Jiarui, et al.
Published: (2024)
Compass Control: Multi Object Orientation Control for Text-to-Image Generation
by: Parihar, Rishubh, et al.
Published: (2025)
by: Parihar, Rishubh, et al.
Published: (2025)
Text2Place: Affordance-aware Text Guided Human Placement
by: Parihar, Rishubh, et al.
Published: (2024)
by: Parihar, Rishubh, et al.
Published: (2024)
MapVerse: A Benchmark for Geospatial Question Answering on Diverse Real-World Maps
by: Bhat, Sharat, et al.
Published: (2026)
by: Bhat, Sharat, et al.
Published: (2026)
ShareVerse: Multi-Agent Consistent Video Generation for Shared World Modeling
by: Zhu, Jiayi, et al.
Published: (2026)
by: Zhu, Jiayi, et al.
Published: (2026)
Realistic Human Motion Generation with Cross-Diffusion Models
by: Ren, Zeping, et al.
Published: (2023)
by: Ren, Zeping, et al.
Published: (2023)
Efficient Label Refinement for Face Parsing Under Extreme Poses Using 3D Gaussian Splatting
by: Gahlawat, Ankit, et al.
Published: (2025)
by: Gahlawat, Ankit, et al.
Published: (2025)
BlendFusion -- Scalable Synthetic Data Generation for Diffusion Model Training
by: Venkatesh, Thejas, et al.
Published: (2026)
by: Venkatesh, Thejas, et al.
Published: (2026)
Reflect3r: Single-View 3D Stereo Reconstruction Aided by Mirror Reflections
by: Wu, Jing, et al.
Published: (2025)
by: Wu, Jing, et al.
Published: (2025)
ProFeAT: Projected Feature Adversarial Training for Self-Supervised Learning of Robust Representations
by: Addepalli, Sravanti, et al.
Published: (2024)
by: Addepalli, Sravanti, et al.
Published: (2024)
Objects in Generated Videos Are Slower Than They Appear: Models Suffer Sub-Earth Gravity and Don't Know Galileo's Principle...for now
by: Thozhiyoor, Varun Varma, et al.
Published: (2025)
by: Thozhiyoor, Varun Varma, et al.
Published: (2025)
Instructive3D: Editing Large Reconstruction Models with Text Instructions
by: Kathare, Kunal, et al.
Published: (2025)
by: Kathare, Kunal, et al.
Published: (2025)
EgoVerse: An Egocentric Human Dataset for Robot Learning from Around the World
by: Punamiya, Ryan, et al.
Published: (2026)
by: Punamiya, Ryan, et al.
Published: (2026)
Reinforced Diffusion: Learning to Push the Limits of Anisotropic Diffusion for Image Denoising
by: Qin, Xinran, et al.
Published: (2025)
by: Qin, Xinran, et al.
Published: (2025)
Multimodal Emotion Recognition via Causal-Diffusion Bridge (Affect-Diff)
by: Sanjyal, Ankit
Published: (2026)
by: Sanjyal, Ankit
Published: (2026)
Reproducibility Study of CDUL: CLIP-Driven Unsupervised Learning for Multi-Label Image Classification
by: Shah, Manan, et al.
Published: (2024)
by: Shah, Manan, et al.
Published: (2024)
ComboVerse: Compositional 3D Assets Creation Using Spatially-Aware Diffusion Guidance
by: Chen, Yongwei, et al.
Published: (2024)
by: Chen, Yongwei, et al.
Published: (2024)
MonoPlace3D: Learning 3D-Aware Object Placement for 3D Monocular Detection
by: Parihar, Rishubh, et al.
Published: (2025)
by: Parihar, Rishubh, et al.
Published: (2025)
Mirror in the Model: Ad Banner Image Generation via Reflective Multi-LLM and Multi-modal Agents
by: Wang, Zhao, et al.
Published: (2025)
by: Wang, Zhao, et al.
Published: (2025)
Similar Items
-
Reflecting Reality: Enabling Diffusion Models to Produce Faithful Mirror Reflections
by: Dhiman, Ankit, et al.
Published: (2024) -
ChromaDistill: Colorizing Monochrome Radiance Fields with Knowledge Distillation
by: Dhiman, Ankit, et al.
Published: (2023) -
UniC-Lift: Unified 3D Instance Segmentation via Contrastive Learning
by: Dhiman, Ankit, et al.
Published: (2025) -
VerseCrafter: Dynamic Realistic Video World Model with 4D Geometric Control
by: Zheng, Sixiao, et al.
Published: (2026) -
Turbo-GS: Accelerating 3D Gaussian Fitting for High-Quality Radiance Fields
by: Dhiman, Ankit, et al.
Published: (2024)