VisPhyWorld: Probing Physical Reasoning via Code-Driven Video Reconstruction
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Liang, Jiarong, Ku, Max, Hui, Ka-Hei, Nie, Ping, Chen, Wenhu |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
PhyMix: Towards Physically Consistent Single-Image 3D Indoor Scene Generation with Implicit--Explicit Optimization
von: Wu, Dongli, et al.
Veröffentlicht: (2026)
von: Wu, Dongli, et al.
Veröffentlicht: (2026)
EditReward: A Human-Aligned Reward Model for Instruction-Guided Image Editing
von: Wu, Keming, et al.
Veröffentlicht: (2025)
von: Wu, Keming, et al.
Veröffentlicht: (2025)
AnyV2V: A Tuning-Free Framework For Any Video-to-Video Editing Tasks
von: Ku, Max, et al.
Veröffentlicht: (2024)
von: Ku, Max, et al.
Veröffentlicht: (2024)
ContPhy: Continuum Physical Concept Learning and Reasoning from Videos
von: Zheng, Zhicheng, et al.
Veröffentlicht: (2024)
von: Zheng, Zhicheng, et al.
Veröffentlicht: (2024)
WorldReasonBench: Human-Aligned Stress Testing of Video Generators as Future World-State Predictors
von: Wu, Keming, et al.
Veröffentlicht: (2026)
von: Wu, Keming, et al.
Veröffentlicht: (2026)
Thinking with Spatial Code for Physical-World Video Reasoning
von: Chen, Jieneng, et al.
Veröffentlicht: (2026)
von: Chen, Jieneng, et al.
Veröffentlicht: (2026)
ProPhy: Progressive Physical Alignment for Dynamic World Simulation
von: Wang, Zijun, et al.
Veröffentlicht: (2025)
von: Wang, Zijun, et al.
Veröffentlicht: (2025)
PhyRecon: Physically Plausible Neural Scene Reconstruction
von: Ni, Junfeng, et al.
Veröffentlicht: (2024)
von: Ni, Junfeng, et al.
Veröffentlicht: (2024)
VideoEval-Pro: Robust and Realistic Long Video Understanding Evaluation
von: Ma, Wentao, et al.
Veröffentlicht: (2025)
von: Ma, Wentao, et al.
Veröffentlicht: (2025)
PhyGround: Benchmarking Physical Reasoning in Generative World Models
von: Lin, Juyi, et al.
Veröffentlicht: (2026)
von: Lin, Juyi, et al.
Veröffentlicht: (2026)
PhyWorld: Physics-Faithful World Model for Video Generation
von: Zhao, Pu, et al.
Veröffentlicht: (2026)
von: Zhao, Pu, et al.
Veröffentlicht: (2026)
PhyRPR: Training-Free Physics-Constrained Video Generation
von: Zhao, Yibo, et al.
Veröffentlicht: (2026)
von: Zhao, Yibo, et al.
Veröffentlicht: (2026)
"PhyWorldBench": A Comprehensive Evaluation of Physical Realism in Text-to-Video Models
von: Gu, Jing, et al.
Veröffentlicht: (2025)
von: Gu, Jing, et al.
Veröffentlicht: (2025)
PhyEdit: Towards Real-World Object Manipulation via Physically-Grounded Image Editing
von: Xu, Ruihang, et al.
Veröffentlicht: (2026)
von: Xu, Ruihang, et al.
Veröffentlicht: (2026)
WonderVerse: Extendable 3D Scene Generation with Video Generative Models
von: Feng, Hao, et al.
Veröffentlicht: (2025)
von: Feng, Hao, et al.
Veröffentlicht: (2025)
PhyEduVideo: A Benchmark for Evaluating Text-to-Video Models for Physics Education
von: M, Megha Mariam K., et al.
Veröffentlicht: (2026)
von: M, Megha Mariam K., et al.
Veröffentlicht: (2026)
VisCoP: Visual Probing for Video Domain Adaptation of Vision Language Models
von: Reilly, Dominick, et al.
Veröffentlicht: (2025)
von: Reilly, Dominick, et al.
Veröffentlicht: (2025)
PhyMAGIC: Physical Motion-Aware Generative Inference with Confidence-guided LLM
von: Meng, Siwei, et al.
Veröffentlicht: (2025)
von: Meng, Siwei, et al.
Veröffentlicht: (2025)
PhyGDPO: Physics-Aware Groupwise Direct Preference Optimization for Physically Consistent Text-to-Video Generation
von: Cai, Yuanhao, et al.
Veröffentlicht: (2025)
von: Cai, Yuanhao, et al.
Veröffentlicht: (2025)
TheoremExplainAgent: Towards Video-based Multimodal Explanations for LLM Theorem Understanding
von: Ku, Max, et al.
Veröffentlicht: (2025)
von: Ku, Max, et al.
Veröffentlicht: (2025)
VisRL: Intention-Driven Visual Perception via Reinforced Reasoning
von: Chen, Zhangquan, et al.
Veröffentlicht: (2025)
von: Chen, Zhangquan, et al.
Veröffentlicht: (2025)
VideoPhy-2: A Challenging Action-Centric Physical Commonsense Evaluation in Video Generation
von: Bansal, Hritik, et al.
Veröffentlicht: (2025)
von: Bansal, Hritik, et al.
Veröffentlicht: (2025)
VideoPhy: Evaluating Physical Commonsense for Video Generation
von: Bansal, Hritik, et al.
Veröffentlicht: (2024)
von: Bansal, Hritik, et al.
Veröffentlicht: (2024)
Context Forcing: Consistent Autoregressive Video Generation with Long Context
von: Chen, Shuo, et al.
Veröffentlicht: (2026)
von: Chen, Shuo, et al.
Veröffentlicht: (2026)
CNS-Edit: 3D Shape Editing via Coupled Neural Shape Optimization
von: Hu, Jingyu, et al.
Veröffentlicht: (2024)
von: Hu, Jingyu, et al.
Veröffentlicht: (2024)
VIEScore: Towards Explainable Metrics for Conditional Image Synthesis Evaluation
von: Ku, Max, et al.
Veröffentlicht: (2023)
von: Ku, Max, et al.
Veröffentlicht: (2023)
PhyBench: A Physical Commonsense Benchmark for Evaluating Text-to-Image Models
von: Meng, Fanqing, et al.
Veröffentlicht: (2024)
von: Meng, Fanqing, et al.
Veröffentlicht: (2024)
Phy124: Fast Physics-Driven 4D Content Generation from a Single Image
von: Lin, Jiajing, et al.
Veröffentlicht: (2024)
von: Lin, Jiajing, et al.
Veröffentlicht: (2024)
Phy-CoSF: Physics-Guided Continuous Spectral Fields Reconstruction and Super-Resolution for Snapshot Compressive Imaging
von: Chen, Wudi, et al.
Veröffentlicht: (2026)
von: Chen, Wudi, et al.
Veröffentlicht: (2026)
Physion-Eval: Evaluating Physical Realism in Generated Video via Human Reasoning
von: Zhang, Qin, et al.
Veröffentlicht: (2026)
von: Zhang, Qin, et al.
Veröffentlicht: (2026)
PhyVLLM: Physics-Guided Video Language Model with Motion-Appearance Disentanglement
von: Zhan, Yu-Wei, et al.
Veröffentlicht: (2025)
von: Zhan, Yu-Wei, et al.
Veröffentlicht: (2025)
M-PhyGs: Multi-Material Object Dynamics from Video
von: Wada, Norika, et al.
Veröffentlicht: (2025)
von: Wada, Norika, et al.
Veröffentlicht: (2025)
PhyCritic: Multimodal Critic Models for Physical AI
von: Xiong, Tianyi, et al.
Veröffentlicht: (2026)
von: Xiong, Tianyi, et al.
Veröffentlicht: (2026)
GenAI Arena: An Open Evaluation Platform for Generative Models
von: Jiang, Dongfu, et al.
Veröffentlicht: (2024)
von: Jiang, Dongfu, et al.
Veröffentlicht: (2024)
ImagenHub: Standardizing the evaluation of conditional image generation models
von: Ku, Max, et al.
Veröffentlicht: (2023)
von: Ku, Max, et al.
Veröffentlicht: (2023)
Moodifier: MLLM-Enhanced Emotion-Driven Image Editing
von: Ye, Jiarong, et al.
Veröffentlicht: (2025)
von: Ye, Jiarong, et al.
Veröffentlicht: (2025)
ViRectify: A Challenging Benchmark for Video Reasoning Correction with Multimodal Large Language Models
von: Hei, Xusen, et al.
Veröffentlicht: (2025)
von: Hei, Xusen, et al.
Veröffentlicht: (2025)
PhyPrompt: RL-based Prompt Refinement for Physically Plausible Text-to-Video Generation
von: Wu, Shang, et al.
Veröffentlicht: (2026)
von: Wu, Shang, et al.
Veröffentlicht: (2026)
VisRes Bench: On Evaluating the Visual Reasoning Capabilities of VLMs
von: Törtei, Brigitta Malagurski, et al.
Veröffentlicht: (2025)
von: Törtei, Brigitta Malagurski, et al.
Veröffentlicht: (2025)
Pixel Reasoner: Incentivizing Pixel-Space Reasoning with Curiosity-Driven Reinforcement Learning
von: Wang, Haozhe, et al.
Veröffentlicht: (2025)
von: Wang, Haozhe, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
PhyMix: Towards Physically Consistent Single-Image 3D Indoor Scene Generation with Implicit--Explicit Optimization
von: Wu, Dongli, et al.
Veröffentlicht: (2026) -
EditReward: A Human-Aligned Reward Model for Instruction-Guided Image Editing
von: Wu, Keming, et al.
Veröffentlicht: (2025) -
AnyV2V: A Tuning-Free Framework For Any Video-to-Video Editing Tasks
von: Ku, Max, et al.
Veröffentlicht: (2024) -
ContPhy: Continuum Physical Concept Learning and Reasoning from Videos
von: Zheng, Zhicheng, et al.
Veröffentlicht: (2024) -
WorldReasonBench: Human-Aligned Stress Testing of Video Generators as Future World-State Predictors
von: Wu, Keming, et al.
Veröffentlicht: (2026)