PhysGame: Uncovering Physical Commonsense Violations in Gameplay Videos
Fuente:
arXiv
Enregistré dans:
| Auteurs principaux: | Cao, Meng, Tang, Haoran, Zhao, Haoze, Guo, Hangyu, Liu, Jiaheng, Zhang, Ge, Liu, Ruyang, Sun, Qiang, Reid, Ian, Liang, Xiaodan |
|---|---|
| Format: | Preprint |
| Publié: |
2024
|
| Sujets: | |
| Accès en ligne: | |
| Tags: |
Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
|
Documents similaires
Order from Chaos: Physical World Understanding from Glitchy Gameplay Videos
par: Cao, Meng, et autres
Publié: (2026)
par: Cao, Meng, et autres
Publié: (2026)
Video Spatial Reasoning with Object-Centric 3D Rollout
par: Tang, Haoran, et autres
Publié: (2025)
par: Tang, Haoran, et autres
Publié: (2025)
ING-VP: MLLMs cannot Play Easy Vision-based Games Yet
par: Zhang, Haoran, et autres
Publié: (2024)
par: Zhang, Haoran, et autres
Publié: (2024)
MUSE: Mamba is Efficient Multi-scale Learner for Text-video Retrieval
par: Tang, Haoran, et autres
Publié: (2024)
par: Tang, Haoran, et autres
Publié: (2024)
RAP: Efficient Text-Video Retrieval with Sparse-and-Correlated Adapter
par: Cao, Meng, et autres
Publié: (2024)
par: Cao, Meng, et autres
Publié: (2024)
Ground-R1: Incentivizing Grounded Visual Reasoning via Reinforcement Learning
par: Cao, Meng, et autres
Publié: (2025)
par: Cao, Meng, et autres
Publié: (2025)
Flow4Agent: Long-form Video Understanding via Motion Prior from Optical Flow
par: Liu, Ruyang, et autres
Publié: (2025)
par: Liu, Ruyang, et autres
Publié: (2025)
Video SimpleQA: Towards Factuality Evaluation in Large Video Language Models
par: Cao, Meng, et autres
Publié: (2025)
par: Cao, Meng, et autres
Publié: (2025)
SpatialDreamer: Incentivizing Spatial Reasoning via Active Mental Imagery
par: Cao, Meng, et autres
Publié: (2025)
par: Cao, Meng, et autres
Publié: (2025)
Seeing through Imagination: Learning Scene Geometry via Implicit Spatial World Modeling
par: Cao, Meng, et autres
Publié: (2025)
par: Cao, Meng, et autres
Publié: (2025)
ST-LLM: Large Language Models Are Effective Temporal Learners
par: Liu, Ruyang, et autres
Publié: (2024)
par: Liu, Ruyang, et autres
Publié: (2024)
PPLLaVA: Varied Video Sequence Understanding With Prompt Guidance
par: Sun, Shangkun, et autres
Publié: (2024)
par: Sun, Shangkun, et autres
Publié: (2024)
BT-Adapter: Video Conversation is Feasible Without Video Instruction Tuning
par: Liu, Ruyang, et autres
Publié: (2023)
par: Liu, Ruyang, et autres
Publié: (2023)
PhysChoreo: Physics-Controllable Video Generation with Part-Aware Semantic Grounding
par: Zhang, Haoze, et autres
Publié: (2025)
par: Zhang, Haoze, et autres
Publié: (2025)
The NES Video-Music Database: A Dataset of Symbolic Video Game Music Paired with Gameplay Videos
par: Cardoso, Igor, et autres
Publié: (2024)
par: Cardoso, Igor, et autres
Publié: (2024)
Fighting Game Adaptive Background Music for Improved Gameplay
par: Khan, Ibrahim, et autres
Publié: (2024)
par: Khan, Ibrahim, et autres
Publié: (2024)
Continual LLaVA: Continual Instruction Tuning in Large Vision-Language Models
par: Cao, Meng, et autres
Publié: (2024)
par: Cao, Meng, et autres
Publié: (2024)
The Garden of Forking Paths: Narrative Arc-Conditioned Gameplay Planning
par: Wen, Yunge, et autres
Publié: (2026)
par: Wen, Yunge, et autres
Publié: (2026)
PhysX-3D: Physical-Grounded 3D Asset Generation
par: Cao, Ziang, et autres
Publié: (2025)
par: Cao, Ziang, et autres
Publié: (2025)
VideoPhy: Evaluating Physical Commonsense for Video Generation
par: Bansal, Hritik, et autres
Publié: (2024)
par: Bansal, Hritik, et autres
Publié: (2024)
PhysCtrl: Generative Physics for Controllable and Physics-Grounded Video Generation
par: Wang, Chen, et autres
Publié: (2025)
par: Wang, Chen, et autres
Publié: (2025)
Grammar and Gameplay-aligned RL for Game Description Generation with LLMs
par: Tanaka, Tsunehiko, et autres
Publié: (2025)
par: Tanaka, Tsunehiko, et autres
Publié: (2025)
Semantic, Orthographic, and Phonological Biases in Humans' Wordle Gameplay
par: Liang, Jiadong, et autres
Publié: (2024)
par: Liang, Jiadong, et autres
Publié: (2024)
Phys4D: Fine-Grained Physics-Consistent 4D Modeling from Video Diffusion
par: Lu, Haoran, et autres
Publié: (2026)
par: Lu, Haoran, et autres
Publié: (2026)
Commonsense Scene Graph-based Target Localization for Object Search
par: Ge, Wenqi, et autres
Publié: (2024)
par: Ge, Wenqi, et autres
Publié: (2024)
JEUX Gameplay
par: Kukkonen, Karin, et autres
Publié: (2025)
par: Kukkonen, Karin, et autres
Publié: (2025)
PhysGen: Rigid-Body Physics-Grounded Image-to-Video Generation
par: Liu, Shaowei, et autres
Publié: (2024)
par: Liu, Shaowei, et autres
Publié: (2024)
Towards World Simulator: Crafting Physical Commonsense-Based Benchmark for Video Generation
par: Meng, Fanqing, et autres
Publié: (2024)
par: Meng, Fanqing, et autres
Publié: (2024)
PhysX-Anything: Simulation-Ready Physical 3D Assets from Single Image
par: Cao, Ziang, et autres
Publié: (2025)
par: Cao, Ziang, et autres
Publié: (2025)
Exploring Gender and Racial/Ethnic Bias Against Video Game Streamers: Comparing Perceived Gameplay Skill and Viewer Engagement
par: Nguyen, David V., et autres
Publié: (2023)
par: Nguyen, David V., et autres
Publié: (2023)
Commonsense Video Question Answering through Video-Grounded Entailment Tree Reasoning
par: Liu, Huabin, et autres
Publié: (2025)
par: Liu, Huabin, et autres
Publié: (2025)
LLM-Hanabi: Evaluating Multi-Agent Gameplays with Theory-of-Mind and Rationale Inference in Imperfect Information Collaboration Game
par: Liang, Fangzhou, et autres
Publié: (2025)
par: Liang, Fangzhou, et autres
Publié: (2025)
Large Language Models in Game Development: Implications for Gameplay, Playability, and Player Experience
par: Johnson, Keeryn, et autres
Publié: (2026)
par: Johnson, Keeryn, et autres
Publié: (2026)
From Gameplay Traces to Game Mechanics: Causal Induction with Large Language Models
par: Jiwatode, Mohit, et autres
Publié: (2026)
par: Jiwatode, Mohit, et autres
Publié: (2026)
SeePhys Pro: Diagnosing Modality Transfer and Blind-Training Effects in Multimodal RLVR for Physics Reasoning
par: Xiang, Kun, et autres
Publié: (2026)
par: Xiang, Kun, et autres
Publié: (2026)
Gameplay Highlights Generation
par: Edithal, Vignesh, et autres
Publié: (2025)
par: Edithal, Vignesh, et autres
Publié: (2025)
PhysPT: Physics-aware Pretrained Transformer for Estimating Human Dynamics from Monocular Videos
par: Zhang, Yufei, et autres
Publié: (2024)
par: Zhang, Yufei, et autres
Publié: (2024)
ConsentDiff at Scale: Longitudinal Audits of Web Privacy Policy Changes and UI Frictions
par: Guo, Haoze
Publié: (2025)
par: Guo, Haoze
Publié: (2025)
Art Card Game (ACG): Embedding Illustration in Gameplay to Mitigate Artist Self-Criticism
par: Mullings, Catherine, et autres
Publié: (2026)
par: Mullings, Catherine, et autres
Publié: (2026)
Droplet3D: Commonsense Priors from Videos Facilitate 3D Generation
par: Li, Xiaochuan, et autres
Publié: (2025)
par: Li, Xiaochuan, et autres
Publié: (2025)
Documents similaires
-
Order from Chaos: Physical World Understanding from Glitchy Gameplay Videos
par: Cao, Meng, et autres
Publié: (2026) -
Video Spatial Reasoning with Object-Centric 3D Rollout
par: Tang, Haoran, et autres
Publié: (2025) -
ING-VP: MLLMs cannot Play Easy Vision-based Games Yet
par: Zhang, Haoran, et autres
Publié: (2024) -
MUSE: Mamba is Efficient Multi-scale Learner for Text-video Retrieval
par: Tang, Haoran, et autres
Publié: (2024) -
RAP: Efficient Text-Video Retrieval with Sparse-and-Correlated Adapter
par: Cao, Meng, et autres
Publié: (2024)