From Gameplay Traces to Game Mechanics: Causal Induction with Large Language Models
Fuente:
arXiv
Saved in:
| Main Authors: | Jiwatode, Mohit, Dockhorn, Alexander, Rosenhahn, Bodo |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Interpretable Decision-Making for End-to-End Autonomous Driving
by: Mirzaie, Mona, et al.
Published: (2025)
by: Mirzaie, Mona, et al.
Published: (2025)
Pixels to Play: A Foundation Model for 3D Gameplay
by: Yue, Yuguang, et al.
Published: (2025)
by: Yue, Yuguang, et al.
Published: (2025)
QPM: Discrete Optimization for Globally Interpretable Image Classification
by: Norrenbrock, Thomas, et al.
Published: (2025)
by: Norrenbrock, Thomas, et al.
Published: (2025)
Multi-Flow: Multi-View-Enriched Normalizing Flows for Industrial Anomaly Detection
by: Kruse, Mathis, et al.
Published: (2025)
by: Kruse, Mathis, et al.
Published: (2025)
Circuit Tracing in Vision-Language Models: Understanding the Internal Mechanisms of Multimodal Thinking
by: Yang, Jingcheng, et al.
Published: (2026)
by: Yang, Jingcheng, et al.
Published: (2026)
Causal Decoding for Hallucination-Resistant Multimodal Large Language Models
by: Tan, Shiwei, et al.
Published: (2026)
by: Tan, Shiwei, et al.
Published: (2026)
Online Optimization of Curriculum Learning Schedules using Evolutionary Optimization
by: Jiwatode, Mohit, et al.
Published: (2024)
by: Jiwatode, Mohit, et al.
Published: (2024)
Q-SENN: Quantized Self-Explaining Neural Networks
by: Norrenbrock, Thomas, et al.
Published: (2023)
by: Norrenbrock, Thomas, et al.
Published: (2023)
DINO-QPM: Adapting Visual Foundation Models for Globally Interpretable Image Classification
by: Zimmermann, Robert, et al.
Published: (2026)
by: Zimmermann, Robert, et al.
Published: (2026)
Rephrase, Augment, Reason: Visual Grounding of Questions for Vision-Language Models
by: Prasad, Archiki, et al.
Published: (2023)
by: Prasad, Archiki, et al.
Published: (2023)
ECoFLaP: Efficient Coarse-to-Fine Layer-Wise Pruning for Vision-Language Models
by: Sung, Yi-Lin, et al.
Published: (2023)
by: Sung, Yi-Lin, et al.
Published: (2023)
Perceptual Quality-based Model Training under Annotator Label Uncertainty
by: Zhou, Chen, et al.
Published: (2024)
by: Zhou, Chen, et al.
Published: (2024)
SplatPose & Detect: Pose-Agnostic 3D Anomaly Detection
by: Kruse, Mathis, et al.
Published: (2024)
by: Kruse, Mathis, et al.
Published: (2024)
Contrastive Region Guidance: Improving Grounding in Vision-Language Models without Training
by: Wan, David, et al.
Published: (2024)
by: Wan, David, et al.
Published: (2024)
Who Brings the Frisbee: Probing Hidden Hallucination Factors in Large Vision-Language Model via Causality Analysis
by: Huang, Po-Hsuan, et al.
Published: (2024)
by: Huang, Po-Hsuan, et al.
Published: (2024)
Exposing and Addressing Cross-Task Inconsistency in Unified Vision-Language Models
by: Maharana, Adyasha, et al.
Published: (2023)
by: Maharana, Adyasha, et al.
Published: (2023)
Ctrl-Adapter: An Efficient and Versatile Framework for Adapting Diverse Controls to Any Diffusion Model
by: Lin, Han, et al.
Published: (2024)
by: Lin, Han, et al.
Published: (2024)
Leveraging Foundation Models for Causal Generative Modeling
by: Komanduri, Aneesh, et al.
Published: (2026)
by: Komanduri, Aneesh, et al.
Published: (2026)
From Theory to Decision Rule: Calibrating the Noisy-Label Crossover for Vision-Language Model Weak Supervision Across Three Medical-Imaging Benchmarks
by: Xu, Bruce Changlong, et al.
Published: (2026)
by: Xu, Bruce Changlong, et al.
Published: (2026)
Counterfactual Gradients-based Quantification of Prediction Trust in Neural Networks
by: Prabhushankar, Mohit, et al.
Published: (2024)
by: Prabhushankar, Mohit, et al.
Published: (2024)
Mementos: A Comprehensive Benchmark for Multimodal Large Language Model Reasoning over Image Sequences
by: Wang, Xiyao, et al.
Published: (2024)
by: Wang, Xiyao, et al.
Published: (2024)
Causality-Driven Audits of Model Robustness
by: Drenkow, Nathan, et al.
Published: (2024)
by: Drenkow, Nathan, et al.
Published: (2024)
Where Reliability Lives in Vision-Language Models: A Mechanistic Study of Attention, Hidden States, and Causal Circuits
by: Mann, Logan, et al.
Published: (2026)
by: Mann, Logan, et al.
Published: (2026)
Parallel In-context Learning for Large Vision Language Models
by: Yamaguchi, Shin'ya, et al.
Published: (2026)
by: Yamaguchi, Shin'ya, et al.
Published: (2026)
PAVE: Patching and Adapting Video Large Language Models
by: Liu, Zhuoming, et al.
Published: (2025)
by: Liu, Zhuoming, et al.
Published: (2025)
Visual Hallucinations of Multi-modal Large Language Models
by: Huang, Wen, et al.
Published: (2024)
by: Huang, Wen, et al.
Published: (2024)
Effectiveness Assessment of Recent Large Vision-Language Models
by: Jiang, Yao, et al.
Published: (2024)
by: Jiang, Yao, et al.
Published: (2024)
Exploring Perceptual Limitation of Multimodal Large Language Models
by: Zhang, Jiarui, et al.
Published: (2024)
by: Zhang, Jiarui, et al.
Published: (2024)
Programmatic Video Prediction Using Large Language Models
by: Tang, Hao, et al.
Published: (2025)
by: Tang, Hao, et al.
Published: (2025)
Learning to Inference Adaptively for Multimodal Large Language Models
by: Xu, Zhuoyan, et al.
Published: (2025)
by: Xu, Zhuoyan, et al.
Published: (2025)
Diffusion Models Are Real-Time Game Engines
by: Valevski, Dani, et al.
Published: (2024)
by: Valevski, Dani, et al.
Published: (2024)
A Multimodal Architecture for Endpoint Position Prediction in Team-based Multiplayer Games
by: Peche, Jonas, et al.
Published: (2025)
by: Peche, Jonas, et al.
Published: (2025)
Exploring the Transferability of Visual Prompting for Multimodal Large Language Models
by: Zhang, Yichi, et al.
Published: (2024)
by: Zhang, Yichi, et al.
Published: (2024)
Diagnosing and Mitigating Modality Interference in Multimodal Large Language Models
by: Cai, Rui, et al.
Published: (2025)
by: Cai, Rui, et al.
Published: (2025)
Representing Online Handwriting for Recognition in Large Vision-Language Models
by: Fadeeva, Anastasiia, et al.
Published: (2024)
by: Fadeeva, Anastasiia, et al.
Published: (2024)
Rethinking Post-Unlearning Behavior of Large Vision-Language Models
by: Kim, Minsung, et al.
Published: (2025)
by: Kim, Minsung, et al.
Published: (2025)
Semi-Supervised Learning for Deep Causal Generative Models
by: Ibrahim, Yasin, et al.
Published: (2024)
by: Ibrahim, Yasin, et al.
Published: (2024)
Learning Truncated Causal History Model for Video Restoration
by: Ghasemabadi, Amirhosein, et al.
Published: (2024)
by: Ghasemabadi, Amirhosein, et al.
Published: (2024)
CLEFT: Language-Image Contrastive Learning with Efficient Large Language Model and Prompt Fine-Tuning
by: Du, Yuexi, et al.
Published: (2024)
by: Du, Yuexi, et al.
Published: (2024)
From Images to Signals: Are Large Vision Models Useful for Time Series Analysis?
by: Zhao, Ziming, et al.
Published: (2025)
by: Zhao, Ziming, et al.
Published: (2025)
Similar Items
-
Interpretable Decision-Making for End-to-End Autonomous Driving
by: Mirzaie, Mona, et al.
Published: (2025) -
Pixels to Play: A Foundation Model for 3D Gameplay
by: Yue, Yuguang, et al.
Published: (2025) -
QPM: Discrete Optimization for Globally Interpretable Image Classification
by: Norrenbrock, Thomas, et al.
Published: (2025) -
Multi-Flow: Multi-View-Enriched Normalizing Flows for Industrial Anomaly Detection
by: Kruse, Mathis, et al.
Published: (2025) -
Circuit Tracing in Vision-Language Models: Understanding the Internal Mechanisms of Multimodal Thinking
by: Yang, Jingcheng, et al.
Published: (2026)