MetaOthello: A Controlled Study of Multiple World Models in Transformers
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Chawla, Aviral, Hall, Galen, Lovato, Juniper |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Automatically Finding Rule-Based Neurons in OthelloGPT
von: Singh, Aditya, et al.
Veröffentlicht: (2025)
von: Singh, Aditya, et al.
Veröffentlicht: (2025)
Emergent Stack Representations in Modeling Counter Languages Using Transformers
von: Tiwari, Utkarsh, et al.
Veröffentlicht: (2025)
von: Tiwari, Utkarsh, et al.
Veröffentlicht: (2025)
Dictionary Learning Improves Patch-Free Circuit Discovery in Mechanistic Interpretability: A Case Study on Othello-GPT
von: He, Zhengfu, et al.
Veröffentlicht: (2024)
von: He, Zhengfu, et al.
Veröffentlicht: (2024)
Provable Generalization in Overparameterized Neural Nets
von: Dhingra, Aviral
Veröffentlicht: (2025)
von: Dhingra, Aviral
Veröffentlicht: (2025)
A Comparative Study of CNN, ResNet, and Vision Transformers for Multi-Classification of Chest Diseases
von: Jain, Ananya, et al.
Veröffentlicht: (2024)
von: Jain, Ananya, et al.
Veröffentlicht: (2024)
Gradient Descent Efficiency Index
von: Dhingra, Aviral
Veröffentlicht: (2024)
von: Dhingra, Aviral
Veröffentlicht: (2024)
MiniZero: Comparative Analysis of AlphaZero and MuZero on Go, Othello, and Atari Games
von: Wu, Ti-Rong, et al.
Veröffentlicht: (2023)
von: Wu, Ti-Rong, et al.
Veröffentlicht: (2023)
Dimension-Free Bounds for Generalized First-Order Methods via Gaussian Coupling
von: Reeves, Galen
Veröffentlicht: (2025)
von: Reeves, Galen
Veröffentlicht: (2025)
AnyLoss: Transforming Classification Metrics into Loss Functions
von: Han, Doheon, et al.
Veröffentlicht: (2024)
von: Han, Doheon, et al.
Veröffentlicht: (2024)
phepy: Visual Benchmarks and Improvements for Out-of-Distribution Detectors
von: Tyree, Juniper, et al.
Veröffentlicht: (2025)
von: Tyree, Juniper, et al.
Veröffentlicht: (2025)
A Formal Framework for Assessing and Mitigating Emergent Security Risks in Generative AI Models: Bridging Theory and Dynamic Risk Mitigation
von: Srivastava, Aviral, et al.
Veröffentlicht: (2024)
von: Srivastava, Aviral, et al.
Veröffentlicht: (2024)
Human-AI Co-Mentorship in Project-Based Learning: A Case Study in Financial Forecasting
von: Chawla, Freyaa, et al.
Veröffentlicht: (2026)
von: Chawla, Freyaa, et al.
Veröffentlicht: (2026)
Rapid Bayesian identification of sparse nonlinear dynamics from scarce and noisy data
von: Fung, Lloyd, et al.
Veröffentlicht: (2024)
von: Fung, Lloyd, et al.
Veröffentlicht: (2024)
Meta-DT: Offline Meta-RL as Conditional Sequence Modeling with World Model Disentanglement
von: Wang, Zhi, et al.
Veröffentlicht: (2024)
von: Wang, Zhi, et al.
Veröffentlicht: (2024)
Unfamiliar Finetuning Examples Control How Language Models Hallucinate
von: Kang, Katie, et al.
Veröffentlicht: (2024)
von: Kang, Katie, et al.
Veröffentlicht: (2024)
Digi-Q: Learning Q-Value Functions for Training Device-Control Agents
von: Bai, Hao, et al.
Veröffentlicht: (2025)
von: Bai, Hao, et al.
Veröffentlicht: (2025)
TabPFN Through The Looking Glass: An interpretability study of TabPFN and its internal representations
von: Gupta, Aviral, et al.
Veröffentlicht: (2026)
von: Gupta, Aviral, et al.
Veröffentlicht: (2026)
BIRD: Behavior Induction via Representation-structure Distillation
von: Pogoncheff, Galen, et al.
Veröffentlicht: (2025)
von: Pogoncheff, Galen, et al.
Veröffentlicht: (2025)
Steering Your Generalists: Improving Robotic Foundation Models via Value Guidance
von: Nakamoto, Mitsuhiko, et al.
Veröffentlicht: (2024)
von: Nakamoto, Mitsuhiko, et al.
Veröffentlicht: (2024)
Transformation-Augmented GRPO for Enhancing Exploration in Reasoning of Large Language Models
von: Le, Khiem, et al.
Veröffentlicht: (2026)
von: Le, Khiem, et al.
Veröffentlicht: (2026)
Behavior-Invariant Task Representation Learning with Transformer-based World Models for Offline Meta-Reinforcement Learning
von: Qian, Fuyuan, et al.
Veröffentlicht: (2026)
von: Qian, Fuyuan, et al.
Veröffentlicht: (2026)
Contextual Latent World Models for Offline Meta Reinforcement Learning
von: Nakheai, Mohammadreza, et al.
Veröffentlicht: (2026)
von: Nakheai, Mohammadreza, et al.
Veröffentlicht: (2026)
MAMBA: an Effective World Model Approach for Meta-Reinforcement Learning
von: Rimon, Zohar, et al.
Veröffentlicht: (2024)
von: Rimon, Zohar, et al.
Veröffentlicht: (2024)
DigiRL: Training In-The-Wild Device-Control Agents with Autonomous Reinforcement Learning
von: Bai, Hao, et al.
Veröffentlicht: (2024)
von: Bai, Hao, et al.
Veröffentlicht: (2024)
Scaling LLM Test-Time Compute Optimally can be More Effective than Scaling Model Parameters
von: Snell, Charlie, et al.
Veröffentlicht: (2024)
von: Snell, Charlie, et al.
Veröffentlicht: (2024)
Information-Theoretic Proofs for Diffusion Sampling
von: Reeves, Galen, et al.
Veröffentlicht: (2025)
von: Reeves, Galen, et al.
Veröffentlicht: (2025)
What Does Flow Matching Bring To TD Learning?
von: Agrawalla, Bhavya, et al.
Veröffentlicht: (2026)
von: Agrawalla, Bhavya, et al.
Veröffentlicht: (2026)
Generative Verifiers: Reward Modeling as Next-Token Prediction
von: Zhang, Lunjun, et al.
Veröffentlicht: (2024)
von: Zhang, Lunjun, et al.
Veröffentlicht: (2024)
TransDreamer: Reinforcement Learning with Transformer World Models
von: Chen, Chang, et al.
Veröffentlicht: (2022)
von: Chen, Chang, et al.
Veröffentlicht: (2022)
Reasoning Cache: Continual Improvement Over Long Horizons via Short-Horizon RL
von: Wu, Ian, et al.
Veröffentlicht: (2026)
von: Wu, Ian, et al.
Veröffentlicht: (2026)
SPARTAN: A Sparse Transformer World Model Attending to What Matters
von: Lei, Anson, et al.
Veröffentlicht: (2024)
von: Lei, Anson, et al.
Veröffentlicht: (2024)
Fourier Circuits in Neural Networks and Transformers: A Case Study of Modular Arithmetic with Multiple Inputs
von: Li, Chenyang, et al.
Veröffentlicht: (2024)
von: Li, Chenyang, et al.
Veröffentlicht: (2024)
Meta-learning Representations for Learning from Multiple Annotators
von: Kumagai, Atsutoshi, et al.
Veröffentlicht: (2025)
von: Kumagai, Atsutoshi, et al.
Veröffentlicht: (2025)
Differentially-Private Data Synthetisation for Efficient Re-Identification Risk Control
von: Carvalho, Tânia, et al.
Veröffentlicht: (2022)
von: Carvalho, Tânia, et al.
Veröffentlicht: (2022)
Discrete Codebook World Models for Continuous Control
von: Scannell, Aidan, et al.
Veröffentlicht: (2025)
von: Scannell, Aidan, et al.
Veröffentlicht: (2025)
Next-Latent Prediction Transformers Learn Compact World Models
von: Teoh, Jayden, et al.
Veröffentlicht: (2025)
von: Teoh, Jayden, et al.
Veröffentlicht: (2025)
Recursive Introspection: Teaching Language Model Agents How to Self-Improve
von: Qu, Yuxiao, et al.
Veröffentlicht: (2024)
von: Qu, Yuxiao, et al.
Veröffentlicht: (2024)
Latent Action World Models for Control with Unlabeled Trajectories
von: Alles, Marvin, et al.
Veröffentlicht: (2025)
von: Alles, Marvin, et al.
Veröffentlicht: (2025)
Optimizing Test-Time Compute via Meta Reinforcement Fine-Tuning
von: Qu, Yuxiao, et al.
Veröffentlicht: (2025)
von: Qu, Yuxiao, et al.
Veröffentlicht: (2025)
Ethical Frameworks for Conducting Social Challenge Studies
von: Sen, Protiva, et al.
Veröffentlicht: (2025)
von: Sen, Protiva, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Automatically Finding Rule-Based Neurons in OthelloGPT
von: Singh, Aditya, et al.
Veröffentlicht: (2025) -
Emergent Stack Representations in Modeling Counter Languages Using Transformers
von: Tiwari, Utkarsh, et al.
Veröffentlicht: (2025) -
Dictionary Learning Improves Patch-Free Circuit Discovery in Mechanistic Interpretability: A Case Study on Othello-GPT
von: He, Zhengfu, et al.
Veröffentlicht: (2024) -
Provable Generalization in Overparameterized Neural Nets
von: Dhingra, Aviral
Veröffentlicht: (2025) -
A Comparative Study of CNN, ResNet, and Vision Transformers for Multi-Classification of Chest Diseases
von: Jain, Ananya, et al.
Veröffentlicht: (2024)