Seeing the Wind from a Falling Leaf
Fuente:
arXiv
Saved in:
| Main Authors: | Gao, Zhiyuan, Mao, Jiageng, Yu, Hong-Xing, Lou, Haozhe, Jia, Emily Yue-Ting, Barbic, Jernej, Wu, Jiajun, Wang, Yue |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Learning an Implicit Physics Model for Image-based Fluid Simulation
by: Jia, Emily Yue-Ting, et al.
Published: (2025)
by: Jia, Emily Yue-Ting, et al.
Published: (2025)
Volume Rendering of Human Hand Anatomy
by: Huang, Jingtao, et al.
Published: (2024)
by: Huang, Jingtao, et al.
Published: (2024)
InstantSfM: Towards GPU-Native SfM for the Deep Learning Era
by: Zhong, Jiankun, et al.
Published: (2025)
by: Zhong, Jiankun, et al.
Published: (2025)
FluidNexus: 3D Fluid Reconstruction and Prediction from a Single Video
by: Gao, Yue, et al.
Published: (2025)
by: Gao, Yue, et al.
Published: (2025)
D-REX: Differentiable Real-to-Sim-to-Real Engine for Learning Dexterous Grasping
by: Lou, Haozhe, et al.
Published: (2026)
by: Lou, Haozhe, et al.
Published: (2026)
Unsupervised Discovery of Object-Centric Neural Fields
by: Luo, Rundong, et al.
Published: (2024)
by: Luo, Rundong, et al.
Published: (2024)
Seeing a Rose in Five Thousand Ways
by: Zhang, Yunzhi, et al.
Published: (2022)
by: Zhang, Yunzhi, et al.
Published: (2022)
SurfPhase: 3D Interfacial Dynamics in Two-Phase Flows from Sparse Videos
by: Gao, Yue, et al.
Published: (2026)
by: Gao, Yue, et al.
Published: (2026)
A Language Agent for Autonomous Driving
by: Mao, Jiageng, et al.
Published: (2023)
by: Mao, Jiageng, et al.
Published: (2023)
RealWonder: Real-Time Physical Action-Conditioned Video Generation
by: Liu, Wei, et al.
Published: (2026)
by: Liu, Wei, et al.
Published: (2026)
Seeing Before Reasoning: A Unified Framework for Generalizable and Explainable Fake Image Detection
by: Lin, Kaiqing, et al.
Published: (2025)
by: Lin, Kaiqing, et al.
Published: (2025)
PerpetualWonder: Long-Horizon Action-Conditioned 4D Scene Generation
by: Zhan, Jiahao, et al.
Published: (2026)
by: Zhan, Jiahao, et al.
Published: (2026)
Reconstruction and Simulation of Elastic Objects with Spring-Mass 3D Gaussians
by: Zhong, Licheng, et al.
Published: (2024)
by: Zhong, Licheng, et al.
Published: (2024)
Oracle Noise: Faster Semantic Spherical Alignment for Interpretable Latent Optimization
by: Li, Haosen, et al.
Published: (2026)
by: Li, Haosen, et al.
Published: (2026)
Guided Path Sampling: Steering Diffusion Models Back on Track with Principled Path Guidance
by: Li, Haosen, et al.
Published: (2025)
by: Li, Haosen, et al.
Published: (2025)
RAM: Retrieval-Based Affordance Transfer for Generalizable Zero-Shot Robotic Manipulation
by: Kuang, Yuxuan, et al.
Published: (2024)
by: Kuang, Yuxuan, et al.
Published: (2024)
WonderZoom: Multi-Scale 3D World Generation
by: Cao, Jin, et al.
Published: (2025)
by: Cao, Jin, et al.
Published: (2025)
ZeroHSI: Zero-Shot 4D Human-Scene Interaction by Video Generation
by: Li, Hongjie, et al.
Published: (2024)
by: Li, Hongjie, et al.
Published: (2024)
ROVER: Benchmarking Reciprocal Cross-Modal Reasoning for Omnimodal Generation
by: Liang, Yongyuan, et al.
Published: (2025)
by: Liang, Yongyuan, et al.
Published: (2025)
Fall Leaf Adversarial Attack on Traffic Sign Classification
by: Etim, Anthony, et al.
Published: (2024)
by: Etim, Anthony, et al.
Published: (2024)
NeuROK: Generative 4D Neural Object Kinematics
by: Geng, Chen, et al.
Published: (2026)
by: Geng, Chen, et al.
Published: (2026)
Physics-Informed Representation Alignment for Sparse Radio-Map Reconstruction
by: Jia, Haozhe, et al.
Published: (2025)
by: Jia, Haozhe, et al.
Published: (2025)
POLARIS: Projection-Orthogonal Least Squares for Robust and Adaptive Inversion in Diffusion Models
by: Chen, Wenshuo, et al.
Published: (2025)
by: Chen, Wenshuo, et al.
Published: (2025)
WonderWorld: Interactive 3D Scene Generation from a Single Image
by: Yu, Hong-Xing, et al.
Published: (2024)
by: Yu, Hong-Xing, et al.
Published: (2024)
PhysBench: Benchmarking and Enhancing Vision-Language Models for Physical World Understanding
by: Chow, Wei, et al.
Published: (2025)
by: Chow, Wei, et al.
Published: (2025)
Free-T2M: Robust Text-to-Motion Generation for Humanoid Robots via Frequency-Domain
by: Chen, Wenshuo, et al.
Published: (2025)
by: Chen, Wenshuo, et al.
Published: (2025)
Bad Seeing or Bad Thinking? Rewarding Perception for Vision-Language Reasoning
by: Wang, Haozhe, et al.
Published: (2026)
by: Wang, Haozhe, et al.
Published: (2026)
NeuraLeaf: Neural Parametric Leaf Models with Shape and Deformation Disentanglement
by: Yang, Yang, et al.
Published: (2025)
by: Yang, Yang, et al.
Published: (2025)
RadioFlow: Efficient Radio Map Construction Framework with Flow Matching
by: Jia, Haozhe, et al.
Published: (2025)
by: Jia, Haozhe, et al.
Published: (2025)
Seeing is Believing? Enhancing Vision-Language Navigation using Visual Perturbations
by: Zhang, Xuesong, et al.
Published: (2024)
by: Zhang, Xuesong, et al.
Published: (2024)
Stanford-ORB: A Real-World 3D Object Inverse Rendering Benchmark
by: Kuang, Zhengfei, et al.
Published: (2023)
by: Kuang, Zhengfei, et al.
Published: (2023)
Do MLLMs Really See It: Reinforcing Visual Attention in Multimodal LLMs
by: Ou, Siqu, et al.
Published: (2026)
by: Ou, Siqu, et al.
Published: (2026)
DiT4Edit: Diffusion Transformer for Image Editing
by: Feng, Kunyu, et al.
Published: (2024)
by: Feng, Kunyu, et al.
Published: (2024)
ECHO: Edge-Cloud Humanoid Orchestration for Language-to-Motion Control
by: Jia, Haozhe, et al.
Published: (2026)
by: Jia, Haozhe, et al.
Published: (2026)
A Holistically Point-guided Text Framework for Weakly-Supervised Camouflaged Object Detection
by: Mok, Tsui Qin, et al.
Published: (2025)
by: Mok, Tsui Qin, et al.
Published: (2025)
SedarEval: Automated Evaluation using Self-Adaptive Rubrics
by: Fan, Zhiyuan, et al.
Published: (2025)
by: Fan, Zhiyuan, et al.
Published: (2025)
AnyLift: Scaling Motion Reconstruction from Internet Videos via 2D Diffusion
by: Li, Hongjie, et al.
Published: (2026)
by: Li, Hongjie, et al.
Published: (2026)
Seeing and Seeing Through the Glass: Real and Synthetic Data for Multi-Layer Depth Estimation
by: Wen, Hongyu, et al.
Published: (2025)
by: Wen, Hongyu, et al.
Published: (2025)
Learning to Think in Physics: Breaking Shortcut Learning in Scientific Diffusion via Representation Alignment
by: Jia, Haozhe, et al.
Published: (2026)
by: Jia, Haozhe, et al.
Published: (2026)
Unified Long Video Inpainting and Outpainting via Overlapping High-Order Co-Denoising
by: Lyu, Shuangquan, et al.
Published: (2025)
by: Lyu, Shuangquan, et al.
Published: (2025)
Similar Items
-
Learning an Implicit Physics Model for Image-based Fluid Simulation
by: Jia, Emily Yue-Ting, et al.
Published: (2025) -
Volume Rendering of Human Hand Anatomy
by: Huang, Jingtao, et al.
Published: (2024) -
InstantSfM: Towards GPU-Native SfM for the Deep Learning Era
by: Zhong, Jiankun, et al.
Published: (2025) -
FluidNexus: 3D Fluid Reconstruction and Prediction from a Single Video
by: Gao, Yue, et al.
Published: (2025) -
D-REX: Differentiable Real-to-Sim-to-Real Engine for Learning Dexterous Grasping
by: Lou, Haozhe, et al.
Published: (2026)