Generative Models: What Do They Know? Do They Know Things? Let's Find Out!
Fuente:
arXiv
Saved in:
| Main Authors: | Du, Xiaodan, Kolkin, Nicholas, Shakhnarovich, Greg, Bhattad, Anand |
|---|---|
| Format: | Preprint |
| Published: |
2023
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Teaching an Agent to Sketch One Part at a Time
by: Du, Xiaodan, et al.
Published: (2026)
by: Du, Xiaodan, et al.
Published: (2026)
Shadows Don't Lie and Lines Can't Bend! Generative Models don't know Projective Geometry...for now
by: Sarkar, Ayush, et al.
Published: (2023)
by: Sarkar, Ayush, et al.
Published: (2023)
Generative Blocks World: Moving Things Around in Pictures
by: Vavilala, Vaibhav, et al.
Published: (2025)
by: Vavilala, Vaibhav, et al.
Published: (2025)
Do generative video models understand physical principles?
by: Motamed, Saman, et al.
Published: (2025)
by: Motamed, Saman, et al.
Published: (2025)
Annotated Hands for Generative Models
by: Yang, Yue, et al.
Published: (2024)
by: Yang, Yue, et al.
Published: (2024)
LumiNet: Latent Intrinsics Meets Diffusion Models for Indoor Scene Relighting
by: Xing, Xiaoyan, et al.
Published: (2024)
by: Xing, Xiaoyan, et al.
Published: (2024)
SignMusketeers: An Efficient Multi-Stream Approach for Sign Language Translation at Scale
by: Gueuwou, Shester, et al.
Published: (2024)
by: Gueuwou, Shester, et al.
Published: (2024)
LoopDraw: a Loop-Based Autoregressive Model for Shape Synthesis and Editing
by: Dinh, Nam Anh, et al.
Published: (2022)
by: Dinh, Nam Anh, et al.
Published: (2022)
Monkey See, Monkey Do: Harnessing Self-attention in Motion Diffusion for Zero-shot Motion Transfer
by: Raab, Sigal, et al.
Published: (2024)
by: Raab, Sigal, et al.
Published: (2024)
What You See is What You GAN: Rendering Every Pixel for High-Fidelity Geometry in 3D GANs
by: Trevithick, Alex, et al.
Published: (2024)
by: Trevithick, Alex, et al.
Published: (2024)
What matters for Representation Alignment: Global Information or Spatial Structure?
by: Singh, Jaskirat, et al.
Published: (2025)
by: Singh, Jaskirat, et al.
Published: (2025)
ComboStoc: Combinatorial Stochasticity for Diffusion Generative Models
by: Xu, Rui, et al.
Published: (2024)
by: Xu, Rui, et al.
Published: (2024)
SliderSpace: Decomposing the Visual Capabilities of Diffusion Models
by: Gandikota, Rohit, et al.
Published: (2025)
by: Gandikota, Rohit, et al.
Published: (2025)
NeuSDFusion: A Spatial-Aware Generative Model for 3D Shape Completion, Reconstruction, and Generation
by: Cui, Ruikai, et al.
Published: (2024)
by: Cui, Ruikai, et al.
Published: (2024)
COLLAGE: Collaborative Human-Agent Interaction Generation using Hierarchical Latent Diffusion and Language Models
by: Daiya, Divyanshu, et al.
Published: (2024)
by: Daiya, Divyanshu, et al.
Published: (2024)
ImagenWorld: Stress-Testing Image Generation Models with Explainable Human Evaluation on Open-ended Real-World Tasks
by: Sani, Samin Mahdizadeh, et al.
Published: (2026)
by: Sani, Samin Mahdizadeh, et al.
Published: (2026)
Improving Video Generation with Human Feedback
by: Liu, Jie, et al.
Published: (2025)
by: Liu, Jie, et al.
Published: (2025)
Image Generation from Contextually-Contradictory Prompts
by: Huberman, Saar, et al.
Published: (2025)
by: Huberman, Saar, et al.
Published: (2025)
Training-Free Consistent Text-to-Image Generation
by: Tewel, Yoad, et al.
Published: (2024)
by: Tewel, Yoad, et al.
Published: (2024)
GenHMR: Generative Human Mesh Recovery
by: Saleem, Muhammad Usama, et al.
Published: (2024)
by: Saleem, Muhammad Usama, et al.
Published: (2024)
Be Decisive: Noise-Induced Layouts for Multi-Subject Generation
by: Dahary, Omer, et al.
Published: (2025)
by: Dahary, Omer, et al.
Published: (2025)
Learning to Infer Generative Template Programs for Visual Concepts
by: Jones, R. Kenny, et al.
Published: (2024)
by: Jones, R. Kenny, et al.
Published: (2024)
RealFill: Reference-Driven Generation for Authentic Image Completion
by: Tang, Luming, et al.
Published: (2023)
by: Tang, Luming, et al.
Published: (2023)
PASTA: Controllable Part-Aware Shape Generation with Autoregressive Transformers
by: Li, Songlin, et al.
Published: (2024)
by: Li, Songlin, et al.
Published: (2024)
Boosting 3D Object Generation through PBR Materials
by: Wang, Yitong, et al.
Published: (2024)
by: Wang, Yitong, et al.
Published: (2024)
Diffusion Self-Distillation for Zero-Shot Customized Image Generation
by: Cai, Shengqu, et al.
Published: (2024)
by: Cai, Shengqu, et al.
Published: (2024)
Revisiting Deepfake Detection: Chronological Continual Learning and the Limits of Generalization
by: Fontana, Federico, et al.
Published: (2025)
by: Fontana, Federico, et al.
Published: (2025)
Be Yourself: Bounded Attention for Multi-Subject Text-to-Image Generation
by: Dahary, Omer, et al.
Published: (2024)
by: Dahary, Omer, et al.
Published: (2024)
Diverse Text-to-Image Generation via Contrastive Noise Optimization
by: Kim, Byungjun, et al.
Published: (2025)
by: Kim, Byungjun, et al.
Published: (2025)
RAGDiffusion: Faithful Cloth Generation via External Knowledge Assimilation
by: Tan, Xianfeng, et al.
Published: (2024)
by: Tan, Xianfeng, et al.
Published: (2024)
PhysGaussian: Physics-Integrated 3D Gaussians for Generative Dynamics
by: Xie, Tianyi, et al.
Published: (2023)
by: Xie, Tianyi, et al.
Published: (2023)
FlexMotion: Lightweight, Physics-Aware, and Controllable Human Motion Generation
by: Tashakori, Arvin, et al.
Published: (2025)
by: Tashakori, Arvin, et al.
Published: (2025)
BulletGen: Improving 4D Reconstruction with Bullet-Time Generation
by: Rozumny, Denis, et al.
Published: (2025)
by: Rozumny, Denis, et al.
Published: (2025)
SALAD: Skeleton-aware Latent Diffusion for Text-driven Motion Generation and Editing
by: Hong, Seokhyeon, et al.
Published: (2025)
by: Hong, Seokhyeon, et al.
Published: (2025)
Stylized Text-to-Motion Generation via Hypernetwork-Driven Low-Rank Adaptation
by: Jeon, Junhyuk, et al.
Published: (2026)
by: Jeon, Junhyuk, et al.
Published: (2026)
BloomScene: Lightweight Structured 3D Gaussian Splatting for Crossmodal Scene Generation
by: Hou, Xiaolu, et al.
Published: (2025)
by: Hou, Xiaolu, et al.
Published: (2025)
RealmDreamer: Text-Driven 3D Scene Generation with Inpainting and Depth Diffusion
by: Shriram, Jaidev, et al.
Published: (2024)
by: Shriram, Jaidev, et al.
Published: (2024)
InfiniteDiffusion: Bridging Learned Fidelity and Procedural Utility for Open-World Terrain Generation
by: Goslin, Alexander
Published: (2025)
by: Goslin, Alexander
Published: (2025)
ArchComplete: Autoregressive 3D Architectural Design Generation with Hierarchical Diffusion-Based Upsampling
by: Rasoulzadeh, S., et al.
Published: (2024)
by: Rasoulzadeh, S., et al.
Published: (2024)
Meta 3D TextureGen: Fast and Consistent Texture Generation for 3D Objects
by: Bensadoun, Raphael, et al.
Published: (2024)
by: Bensadoun, Raphael, et al.
Published: (2024)
Similar Items
-
Teaching an Agent to Sketch One Part at a Time
by: Du, Xiaodan, et al.
Published: (2026) -
Shadows Don't Lie and Lines Can't Bend! Generative Models don't know Projective Geometry...for now
by: Sarkar, Ayush, et al.
Published: (2023) -
Generative Blocks World: Moving Things Around in Pictures
by: Vavilala, Vaibhav, et al.
Published: (2025) -
Do generative video models understand physical principles?
by: Motamed, Saman, et al.
Published: (2025) -
Annotated Hands for Generative Models
by: Yang, Yue, et al.
Published: (2024)