Embedding Physical Reasoning into Diffusion-Based Shadow Generation
Fuente:
arXiv
Guardado en:
| Autores principales: | Hu, Shilin, Xu, Jingyi, Dave, Akshat, Samaras, Dimitris, Le, Hieu |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Cast and Attached Shadow Detection via Iterative Light and Geometry Reasoning
por: Hu, Shilin, et al.
Publicado: (2025)
por: Hu, Shilin, et al.
Publicado: (2025)
Shadow Removal Refinement via Material-Consistent Shadow Edges
por: Hu, Shilin, et al.
Publicado: (2024)
por: Hu, Shilin, et al.
Publicado: (2024)
Assessing Sample Quality via the Latent Space of Generative Models
por: Xu, Jingyi, et al.
Publicado: (2024)
por: Xu, Jingyi, et al.
Publicado: (2024)
Importance-Based Token Merging for Efficient Image and Video Generation
por: Wu, Haoyu, et al.
Publicado: (2024)
por: Wu, Haoyu, et al.
Publicado: (2024)
One Attention, One Scale: Phase-Aligned Rotary Positional Embeddings for Mixed-Resolution Diffusion Transformer
por: Wu, Haoyu, et al.
Publicado: (2025)
por: Wu, Haoyu, et al.
Publicado: (2025)
Talking Head Generation via AU-Guided Landmark Prediction
por: Chang, Shao-Yu, et al.
Publicado: (2025)
por: Chang, Shao-Yu, et al.
Publicado: (2025)
Learning 3D Reconstruction with Priors in Test Time
por: Zhou, Lei, et al.
Publicado: (2026)
por: Zhou, Lei, et al.
Publicado: (2026)
CORA: Consistency-Guided Semi-Supervised Framework for Reasoning Segmentation
por: Howlader, Prantik, et al.
Publicado: (2025)
por: Howlader, Prantik, et al.
Publicado: (2025)
Weighting Pseudo-Labels via High-Activation Feature Index Similarity and Object Detection for Semi-Supervised Segmentation
por: Howlader, Prantik, et al.
Publicado: (2024)
por: Howlader, Prantik, et al.
Publicado: (2024)
Improving Contrastive Learning for Referring Expression Counting
por: Triaridis, Kostas, et al.
Publicado: (2025)
por: Triaridis, Kostas, et al.
Publicado: (2025)
Beyond Pixels: Semi-Supervised Semantic Segmentation with a Multi-scale Patch-based Multi-Label Classifier
por: Howlader, Prantik, et al.
Publicado: (2024)
por: Howlader, Prantik, et al.
Publicado: (2024)
Few-shot Personalized Scanpath Prediction
por: Xue, Ruoyu, et al.
Publicado: (2025)
por: Xue, Ruoyu, et al.
Publicado: (2025)
Poppy: Polarization-based Plug-and-Play Guidance for Enhancing Monocular Normal Estimation
por: Kim, Irene, et al.
Publicado: (2026)
por: Kim, Irene, et al.
Publicado: (2026)
Phrase-Instance Alignment for Generalized Referring Segmentation
por: Nguyen, E-Ro, et al.
Publicado: (2024)
por: Nguyen, E-Ro, et al.
Publicado: (2024)
Personalized Image Descriptions from Attention Sequences
por: Xue, Ruoyu, et al.
Publicado: (2025)
por: Xue, Ruoyu, et al.
Publicado: (2025)
TopoDiffusionNet: A Topology-aware Diffusion Model
por: Gupta, Saumya, et al.
Publicado: (2024)
por: Gupta, Saumya, et al.
Publicado: (2024)
Learning Relighting and Intrinsic Decomposition in Neural Radiance Fields
por: Yang, Yixiong, et al.
Publicado: (2024)
por: Yang, Yixiong, et al.
Publicado: (2024)
MLI-NeRF: Multi-Light Intrinsic-Aware Neural Radiance Fields
por: Yang, Yixiong, et al.
Publicado: (2024)
por: Yang, Yixiong, et al.
Publicado: (2024)
GriDiT: Factorized Grid-Based Diffusion for Efficient Long Image Sequence Generation
por: Tomar, Snehal Singh, et al.
Publicado: (2025)
por: Tomar, Snehal Singh, et al.
Publicado: (2025)
Self-supervised co-salient object detection via feature correspondence at multiple scales
por: Chakraborty, Souradeep, et al.
Publicado: (2024)
por: Chakraborty, Souradeep, et al.
Publicado: (2024)
PathSegDiff: Pathology Segmentation using Diffusion model representations
por: Danisetty, Sachin Kumar, et al.
Publicado: (2025)
por: Danisetty, Sachin Kumar, et al.
Publicado: (2025)
$\infty$-Brush: Controllable Large Image Synthesis with Diffusion Models in Infinite Dimensions
por: Le, Minh-Quan, et al.
Publicado: (2024)
por: Le, Minh-Quan, et al.
Publicado: (2024)
What about gravity in video generation? Post-Training Newton's Laws with Verifiable Rewards
por: Le, Minh-Quan, et al.
Publicado: (2025)
por: Le, Minh-Quan, et al.
Publicado: (2025)
JEAN: Joint Expression and Audio-guided NeRF-based Talking Face Generation
por: Chakkera, Sai Tanmay Reddy, et al.
Publicado: (2024)
por: Chakkera, Sai Tanmay Reddy, et al.
Publicado: (2024)
Multi-view Gaze Target Estimation
por: Miao, Qiaomu, et al.
Publicado: (2025)
por: Miao, Qiaomu, et al.
Publicado: (2025)
All Seeds Are Not Equal: Enhancing Compositional Text-to-Image Generation with Reliable Random Seeds
por: Li, Shuangqi, et al.
Publicado: (2024)
por: Li, Shuangqi, et al.
Publicado: (2024)
SuperEx: Enhancing Indoor Mapping and Exploration using Non-Line-of-Sight Perception
por: Garg, Kush, et al.
Publicado: (2025)
por: Garg, Kush, et al.
Publicado: (2025)
MI-NeRF: Learning a Single Face NeRF from Multiple Identities
por: Chatziagapi, Aggelina, et al.
Publicado: (2024)
por: Chatziagapi, Aggelina, et al.
Publicado: (2024)
MIGS: Multi-Identity Gaussian Splatting via Tensor Decomposition
por: Chatziagapi, Aggelina, et al.
Publicado: (2024)
por: Chatziagapi, Aggelina, et al.
Publicado: (2024)
Diffusion-Refined VQA Annotations for Semi-Supervised Gaze Following
por: Miao, Qiaomu, et al.
Publicado: (2024)
por: Miao, Qiaomu, et al.
Publicado: (2024)
TopoCellGen: Generating Histopathology Cell Topology with a Diffusion Model
por: Xu, Meilong, et al.
Publicado: (2024)
por: Xu, Meilong, et al.
Publicado: (2024)
Modeling Deep Learning Based Privacy Attacks on Physical Mail
por: Huang, Bingyao, et al.
Publicado: (2020)
por: Huang, Bingyao, et al.
Publicado: (2020)
Fast constrained sampling in pre-trained diffusion models
por: Graikos, Alexandros, et al.
Publicado: (2024)
por: Graikos, Alexandros, et al.
Publicado: (2024)
Pairwise-Constrained Implicit Functions for 3D Human Heart Modelling
por: Le, Hieu, et al.
Publicado: (2023)
por: Le, Hieu, et al.
Publicado: (2023)
Mitigating Diffusion Model Hallucinations with Dynamic Guidance
por: Triaridis, Kostas, et al.
Publicado: (2025)
por: Triaridis, Kostas, et al.
Publicado: (2025)
Learning to Weight Parameters for Training Data Attribution
por: Li, Shuangqi, et al.
Publicado: (2025)
por: Li, Shuangqi, et al.
Publicado: (2025)
Rig3DGS: Creating Controllable Portraits from Casual Monocular Videos
por: Rivero, Alfredo, et al.
Publicado: (2024)
por: Rivero, Alfredo, et al.
Publicado: (2024)
PISCES: Annotation-free Text-to-Video Post-Training via Optimal Transport-Aligned Rewards
por: Le, Minh-Quan, et al.
Publicado: (2026)
por: Le, Minh-Quan, et al.
Publicado: (2026)
ZoomLDM: Latent Diffusion Model for multi-scale image generation
por: Yellapragada, Srikar, et al.
Publicado: (2024)
por: Yellapragada, Srikar, et al.
Publicado: (2024)
CDG-MAE: Learning Correspondences from Diffusion Generated Views
por: Belagali, Varun, et al.
Publicado: (2025)
por: Belagali, Varun, et al.
Publicado: (2025)
Ejemplares similares
-
Cast and Attached Shadow Detection via Iterative Light and Geometry Reasoning
por: Hu, Shilin, et al.
Publicado: (2025) -
Shadow Removal Refinement via Material-Consistent Shadow Edges
por: Hu, Shilin, et al.
Publicado: (2024) -
Assessing Sample Quality via the Latent Space of Generative Models
por: Xu, Jingyi, et al.
Publicado: (2024) -
Importance-Based Token Merging for Efficient Image and Video Generation
por: Wu, Haoyu, et al.
Publicado: (2024) -
One Attention, One Scale: Phase-Aligned Rotary Positional Embeddings for Mixed-Resolution Diffusion Transformer
por: Wu, Haoyu, et al.
Publicado: (2025)