Vision-as-Inverse-Graphics Agent via Interleaved Multimodal Reasoning
Fuente:
arXiv
Saved in:
| Main Authors: | Yin, Shaofeng, Ge, Jiaxin, Wang, Zora Zhiruo, Wang, Chenyang, Li, Xiuyu, Black, Michael J., Darrell, Trevor, Kanazawa, Angjoo, Feng, Haiwen |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Agent-to-Sim: Learning Interactive Behavior Models from Casual Longitudinal Videos
by: Yang, Gengshan, et al.
Published: (2024)
by: Yang, Gengshan, et al.
Published: (2024)
SOAR: Self-Occluded Avatar Recovery from a Single Video In the Wild
by: Pan, Zhuoyang, et al.
Published: (2024)
by: Pan, Zhuoyang, et al.
Published: (2024)
NeRF-XL: Scaling NeRFs with Multiple GPUs
by: Li, Ruilong, et al.
Published: (2024)
by: Li, Ruilong, et al.
Published: (2024)
ETCH: Generalizing Body Fitting to Clothed Humans via Equivariant Tightness
by: Li, Boqian, et al.
Published: (2025)
by: Li, Boqian, et al.
Published: (2025)
GARField: Group Anything with Radiance Fields
by: Kim, Chung Min, et al.
Published: (2024)
by: Kim, Chung Min, et al.
Published: (2024)
Fillerbuster: Unified Generative Scene Completion Model for Casual Captures
by: Weber, Ethan, et al.
Published: (2025)
by: Weber, Ethan, et al.
Published: (2025)
Graphical X Splatting (GraphiXS): A Graphical Model for 4D Gaussian Splatting under Uncertainty
by: Yılmaz, Doğa, et al.
Published: (2026)
by: Yılmaz, Doğa, et al.
Published: (2026)
GenLit: Reformulating Single-Image Relighting as Video Generation
by: Bharadwaj, Shrisha, et al.
Published: (2024)
by: Bharadwaj, Shrisha, et al.
Published: (2024)
Rethinking Score Distillation as a Bridge Between Image Distributions
by: McAllister, David, et al.
Published: (2024)
by: McAllister, David, et al.
Published: (2024)
Self‐Supervised Image Harmonization via Region‐Aware Harmony Classification
by: Chenyang Tian, et al.
Published: (2025)
by: Chenyang Tian, et al.
Published: (2025)
Multiphysics Simulation Methods in Computer Graphics
by: Daniel Holz, et al.
Published: (2025)
by: Daniel Holz, et al.
Published: (2025)
SUPQA: LLM‐based Geo‐Visualization for Subjective Urban Performance Question‐Answering
by: Haiwen Huang, et al.
Published: (2025)
by: Haiwen Huang, et al.
Published: (2025)
St4RTrack: Simultaneous 4D Reconstruction and Tracking in the World
by: Feng, Haiwen, et al.
Published: (2025)
by: Feng, Haiwen, et al.
Published: (2025)
PosterReward: Unlocking Accurate Evaluation for High-Quality Graphic Design Generation
by: Lai, Jianyu, et al.
Published: (2026)
by: Lai, Jianyu, et al.
Published: (2026)
Statistical Blendshape Calculation and Analysis for Graphics Applications
by: Li, Shuxian, et al.
Published: (2026)
by: Li, Shuxian, et al.
Published: (2026)
Role of Graphics in Disaster Communication: Practitioner Perspectives on Use, Challenges, and Inclusivity
by: Madugalla, Anuradha, et al.
Published: (2026)
by: Madugalla, Anuradha, et al.
Published: (2026)
BlenderAlchemy: Editing 3D Graphics with Vision-Language Models
by: Huang, Ian, et al.
Published: (2024)
by: Huang, Ian, et al.
Published: (2024)
Can any model be fabricated? Inverse operation based planning for hybrid additive-subtractive manufacturing
by: Chen, Yongxue, et al.
Published: (2025)
by: Chen, Yongxue, et al.
Published: (2025)
LayoutRectifier: An Optimization-based Post-processing for Graphic Design Layout Generation
by: Shen, I-Chao, et al.
Published: (2025)
by: Shen, I-Chao, et al.
Published: (2025)
Towards Understanding Graphical Perception in Large Multimodal Models
by: Zhang, Kai, et al.
Published: (2025)
by: Zhang, Kai, et al.
Published: (2025)
Deep Inverse Shading: Consistent Albedo and Surface Detail Recovery via Generative Refinement
by: Wu, Jiacheng, et al.
Published: (2025)
by: Wu, Jiacheng, et al.
Published: (2025)
Inverse Discrete Elastic Rod
by: Li, Jiahao, et al.
Published: (2025)
by: Li, Jiahao, et al.
Published: (2025)
The Racial Character of Computer Graphics Research
by: Kim, Theodore, et al.
Published: (2026)
by: Kim, Theodore, et al.
Published: (2026)
MATStruct: High-Quality Medial Mesh Computation via Structure-aware Variational Optimization
by: Wang, Ningna, et al.
Published: (2025)
by: Wang, Ningna, et al.
Published: (2025)
Dynamic-Interactive Graphics for Statistics (26 Years Later)
by: Pedro Valero-Mora
Published: (2014)
by: Pedro Valero-Mora
Published: (2014)
pyGANDALF -- An open-source, Geometric, ANimation, Directed, Algorithmic, Learning Framework for Computer Graphics
by: Petropoulos, John, et al.
Published: (2024)
by: Petropoulos, John, et al.
Published: (2024)
CrowdVLA: Embodied Vision-Language-Action Agents for Context-Aware Crowd Simulation
by: Hwang, Juyeong, et al.
Published: (2026)
by: Hwang, Juyeong, et al.
Published: (2026)
Unified Smooth Vector Graphics: Modeling Gradient Meshes and Curve-based Approaches Jointly as Poisson Problem
by: Tian, Xingze, et al.
Published: (2024)
by: Tian, Xingze, et al.
Published: (2024)
Seamless and Aligned Texture Optimization for 3D Reconstruction
by: Lei Wang, et al.
Published: (2024)
by: Lei Wang, et al.
Published: (2024)
Inverse Garment and Pattern Modeling with a Differentiable Simulator
by: Yu, Boyang, et al.
Published: (2024)
by: Yu, Boyang, et al.
Published: (2024)
Impact and Educational Effectiveness of the Graphic Adaptation of Sapiens in Depicting Human Evolution
by: Kumar, Sanjiv
Published: (2022)
by: Kumar, Sanjiv
Published: (2022)
How Does a Virtual Agent Decide Where to Look? Symbolic Cognitive Reasoning for Embodied Head Rotation
by: Hwang, Juyeong, et al.
Published: (2025)
by: Hwang, Juyeong, et al.
Published: (2025)
ETBHD‐HMF: A Hierarchical Multimodal Fusion Architecture for Enhanced Text‐Based Hair Design
by: Rong He, et al.
Published: (2024)
by: Rong He, et al.
Published: (2024)
WaSP: Warp Scheduling to Mimic Prefetching in Graphics Workloads
by: Joseph, Diya, et al.
Published: (2024)
by: Joseph, Diya, et al.
Published: (2024)
BlenderGym: Benchmarking Foundational Model Systems for Graphics Editing
by: Gu, Yunqi, et al.
Published: (2025)
by: Gu, Yunqi, et al.
Published: (2025)
NCD: Normal‐Guided Chamfer Distance Loss for Watertight Mesh Reconstruction from Unoriented Point Clouds
by: Jiaxin Li, et al.
Published: (2025)
by: Jiaxin Li, et al.
Published: (2025)
ViRAC: A Vision-Reasoning Agent Head Movement Control Framework in Arbitrary Virtual Environments
by: Hwang, Juyeong, et al.
Published: (2025)
by: Hwang, Juyeong, et al.
Published: (2025)
AnnoGram: An Annotative Grammar of Graphics Extension
by: Rahman, Md Dilshadur, et al.
Published: (2025)
by: Rahman, Md Dilshadur, et al.
Published: (2025)
Refined Inverse Rigging: A Balanced Approach to High-fidelity Blendshape Animation
by: Racković, Stevo, et al.
Published: (2024)
by: Racković, Stevo, et al.
Published: (2024)
WildCap: Facial Albedo Capture in the Wild via Hybrid Inverse Rendering
by: Han, Yuxuan, et al.
Published: (2025)
by: Han, Yuxuan, et al.
Published: (2025)
Similar Items
-
Agent-to-Sim: Learning Interactive Behavior Models from Casual Longitudinal Videos
by: Yang, Gengshan, et al.
Published: (2024) -
SOAR: Self-Occluded Avatar Recovery from a Single Video In the Wild
by: Pan, Zhuoyang, et al.
Published: (2024) -
NeRF-XL: Scaling NeRFs with Multiple GPUs
by: Li, Ruilong, et al.
Published: (2024) -
ETCH: Generalizing Body Fitting to Clothed Humans via Equivariant Tightness
by: Li, Boqian, et al.
Published: (2025) -
GARField: Group Anything with Radiance Fields
by: Kim, Chung Min, et al.
Published: (2024)