PhysGame: Uncovering Physical Commonsense Violations in Gameplay Videos
Fuente:
arXiv
Salvato in:
| Autori principali: | Cao, Meng, Tang, Haoran, Zhao, Haoze, Guo, Hangyu, Liu, Jiaheng, Zhang, Ge, Liu, Ruyang, Sun, Qiang, Reid, Ian, Liang, Xiaodan |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Order from Chaos: Physical World Understanding from Glitchy Gameplay Videos
di: Cao, Meng, et al.
Pubblicazione: (2026)
di: Cao, Meng, et al.
Pubblicazione: (2026)
Video Spatial Reasoning with Object-Centric 3D Rollout
di: Tang, Haoran, et al.
Pubblicazione: (2025)
di: Tang, Haoran, et al.
Pubblicazione: (2025)
ING-VP: MLLMs cannot Play Easy Vision-based Games Yet
di: Zhang, Haoran, et al.
Pubblicazione: (2024)
di: Zhang, Haoran, et al.
Pubblicazione: (2024)
MUSE: Mamba is Efficient Multi-scale Learner for Text-video Retrieval
di: Tang, Haoran, et al.
Pubblicazione: (2024)
di: Tang, Haoran, et al.
Pubblicazione: (2024)
RAP: Efficient Text-Video Retrieval with Sparse-and-Correlated Adapter
di: Cao, Meng, et al.
Pubblicazione: (2024)
di: Cao, Meng, et al.
Pubblicazione: (2024)
Ground-R1: Incentivizing Grounded Visual Reasoning via Reinforcement Learning
di: Cao, Meng, et al.
Pubblicazione: (2025)
di: Cao, Meng, et al.
Pubblicazione: (2025)
Flow4Agent: Long-form Video Understanding via Motion Prior from Optical Flow
di: Liu, Ruyang, et al.
Pubblicazione: (2025)
di: Liu, Ruyang, et al.
Pubblicazione: (2025)
Video SimpleQA: Towards Factuality Evaluation in Large Video Language Models
di: Cao, Meng, et al.
Pubblicazione: (2025)
di: Cao, Meng, et al.
Pubblicazione: (2025)
SpatialDreamer: Incentivizing Spatial Reasoning via Active Mental Imagery
di: Cao, Meng, et al.
Pubblicazione: (2025)
di: Cao, Meng, et al.
Pubblicazione: (2025)
Seeing through Imagination: Learning Scene Geometry via Implicit Spatial World Modeling
di: Cao, Meng, et al.
Pubblicazione: (2025)
di: Cao, Meng, et al.
Pubblicazione: (2025)
ST-LLM: Large Language Models Are Effective Temporal Learners
di: Liu, Ruyang, et al.
Pubblicazione: (2024)
di: Liu, Ruyang, et al.
Pubblicazione: (2024)
PPLLaVA: Varied Video Sequence Understanding With Prompt Guidance
di: Sun, Shangkun, et al.
Pubblicazione: (2024)
di: Sun, Shangkun, et al.
Pubblicazione: (2024)
BT-Adapter: Video Conversation is Feasible Without Video Instruction Tuning
di: Liu, Ruyang, et al.
Pubblicazione: (2023)
di: Liu, Ruyang, et al.
Pubblicazione: (2023)
PhysChoreo: Physics-Controllable Video Generation with Part-Aware Semantic Grounding
di: Zhang, Haoze, et al.
Pubblicazione: (2025)
di: Zhang, Haoze, et al.
Pubblicazione: (2025)
The NES Video-Music Database: A Dataset of Symbolic Video Game Music Paired with Gameplay Videos
di: Cardoso, Igor, et al.
Pubblicazione: (2024)
di: Cardoso, Igor, et al.
Pubblicazione: (2024)
Fighting Game Adaptive Background Music for Improved Gameplay
di: Khan, Ibrahim, et al.
Pubblicazione: (2024)
di: Khan, Ibrahim, et al.
Pubblicazione: (2024)
Continual LLaVA: Continual Instruction Tuning in Large Vision-Language Models
di: Cao, Meng, et al.
Pubblicazione: (2024)
di: Cao, Meng, et al.
Pubblicazione: (2024)
The Garden of Forking Paths: Narrative Arc-Conditioned Gameplay Planning
di: Wen, Yunge, et al.
Pubblicazione: (2026)
di: Wen, Yunge, et al.
Pubblicazione: (2026)
PhysX-3D: Physical-Grounded 3D Asset Generation
di: Cao, Ziang, et al.
Pubblicazione: (2025)
di: Cao, Ziang, et al.
Pubblicazione: (2025)
VideoPhy: Evaluating Physical Commonsense for Video Generation
di: Bansal, Hritik, et al.
Pubblicazione: (2024)
di: Bansal, Hritik, et al.
Pubblicazione: (2024)
PhysCtrl: Generative Physics for Controllable and Physics-Grounded Video Generation
di: Wang, Chen, et al.
Pubblicazione: (2025)
di: Wang, Chen, et al.
Pubblicazione: (2025)
Grammar and Gameplay-aligned RL for Game Description Generation with LLMs
di: Tanaka, Tsunehiko, et al.
Pubblicazione: (2025)
di: Tanaka, Tsunehiko, et al.
Pubblicazione: (2025)
Semantic, Orthographic, and Phonological Biases in Humans' Wordle Gameplay
di: Liang, Jiadong, et al.
Pubblicazione: (2024)
di: Liang, Jiadong, et al.
Pubblicazione: (2024)
Phys4D: Fine-Grained Physics-Consistent 4D Modeling from Video Diffusion
di: Lu, Haoran, et al.
Pubblicazione: (2026)
di: Lu, Haoran, et al.
Pubblicazione: (2026)
Commonsense Scene Graph-based Target Localization for Object Search
di: Ge, Wenqi, et al.
Pubblicazione: (2024)
di: Ge, Wenqi, et al.
Pubblicazione: (2024)
JEUX Gameplay
di: Kukkonen, Karin, et al.
Pubblicazione: (2025)
di: Kukkonen, Karin, et al.
Pubblicazione: (2025)
PhysGen: Rigid-Body Physics-Grounded Image-to-Video Generation
di: Liu, Shaowei, et al.
Pubblicazione: (2024)
di: Liu, Shaowei, et al.
Pubblicazione: (2024)
Towards World Simulator: Crafting Physical Commonsense-Based Benchmark for Video Generation
di: Meng, Fanqing, et al.
Pubblicazione: (2024)
di: Meng, Fanqing, et al.
Pubblicazione: (2024)
PhysX-Anything: Simulation-Ready Physical 3D Assets from Single Image
di: Cao, Ziang, et al.
Pubblicazione: (2025)
di: Cao, Ziang, et al.
Pubblicazione: (2025)
Exploring Gender and Racial/Ethnic Bias Against Video Game Streamers: Comparing Perceived Gameplay Skill and Viewer Engagement
di: Nguyen, David V., et al.
Pubblicazione: (2023)
di: Nguyen, David V., et al.
Pubblicazione: (2023)
Commonsense Video Question Answering through Video-Grounded Entailment Tree Reasoning
di: Liu, Huabin, et al.
Pubblicazione: (2025)
di: Liu, Huabin, et al.
Pubblicazione: (2025)
LLM-Hanabi: Evaluating Multi-Agent Gameplays with Theory-of-Mind and Rationale Inference in Imperfect Information Collaboration Game
di: Liang, Fangzhou, et al.
Pubblicazione: (2025)
di: Liang, Fangzhou, et al.
Pubblicazione: (2025)
Large Language Models in Game Development: Implications for Gameplay, Playability, and Player Experience
di: Johnson, Keeryn, et al.
Pubblicazione: (2026)
di: Johnson, Keeryn, et al.
Pubblicazione: (2026)
From Gameplay Traces to Game Mechanics: Causal Induction with Large Language Models
di: Jiwatode, Mohit, et al.
Pubblicazione: (2026)
di: Jiwatode, Mohit, et al.
Pubblicazione: (2026)
SeePhys Pro: Diagnosing Modality Transfer and Blind-Training Effects in Multimodal RLVR for Physics Reasoning
di: Xiang, Kun, et al.
Pubblicazione: (2026)
di: Xiang, Kun, et al.
Pubblicazione: (2026)
Gameplay Highlights Generation
di: Edithal, Vignesh, et al.
Pubblicazione: (2025)
di: Edithal, Vignesh, et al.
Pubblicazione: (2025)
PhysPT: Physics-aware Pretrained Transformer for Estimating Human Dynamics from Monocular Videos
di: Zhang, Yufei, et al.
Pubblicazione: (2024)
di: Zhang, Yufei, et al.
Pubblicazione: (2024)
ConsentDiff at Scale: Longitudinal Audits of Web Privacy Policy Changes and UI Frictions
di: Guo, Haoze
Pubblicazione: (2025)
di: Guo, Haoze
Pubblicazione: (2025)
Art Card Game (ACG): Embedding Illustration in Gameplay to Mitigate Artist Self-Criticism
di: Mullings, Catherine, et al.
Pubblicazione: (2026)
di: Mullings, Catherine, et al.
Pubblicazione: (2026)
Droplet3D: Commonsense Priors from Videos Facilitate 3D Generation
di: Li, Xiaochuan, et al.
Pubblicazione: (2025)
di: Li, Xiaochuan, et al.
Pubblicazione: (2025)
Documenti analoghi
-
Order from Chaos: Physical World Understanding from Glitchy Gameplay Videos
di: Cao, Meng, et al.
Pubblicazione: (2026) -
Video Spatial Reasoning with Object-Centric 3D Rollout
di: Tang, Haoran, et al.
Pubblicazione: (2025) -
ING-VP: MLLMs cannot Play Easy Vision-based Games Yet
di: Zhang, Haoran, et al.
Pubblicazione: (2024) -
MUSE: Mamba is Efficient Multi-scale Learner for Text-video Retrieval
di: Tang, Haoran, et al.
Pubblicazione: (2024) -
RAP: Efficient Text-Video Retrieval with Sparse-and-Correlated Adapter
di: Cao, Meng, et al.
Pubblicazione: (2024)