What if Othello-Playing Language Models Could See?
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Chen, Xinyi, Yuan, Yifei, Li, Jiaang, Belongie, Serge, de Rijke, Maarten, Søgaard, Anders |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Revisiting the Othello World Model Hypothesis
von: Yuan, Yifei, et al.
Veröffentlicht: (2025)
von: Yuan, Yifei, et al.
Veröffentlicht: (2025)
Do Vision and Language Models Share Concepts? A Vector Space Alignment Study
von: Li, Jiaang, et al.
Veröffentlicht: (2023)
von: Li, Jiaang, et al.
Veröffentlicht: (2023)
Can Small Agents Collaborate to Beat a Single Large Language Model?
von: Żywot, Agata, et al.
Veröffentlicht: (2026)
von: Żywot, Agata, et al.
Veröffentlicht: (2026)
ChatMotion: A Multimodal Multi-Agent for Human Motion Analysis
von: Li, Lei, et al.
Veröffentlicht: (2025)
von: Li, Lei, et al.
Veröffentlicht: (2025)
AGENT-CQ: Automatic Generation and Evaluation of Clarifying Questions for Conversational Search with LLMs
von: Siro, Clemencia, et al.
Veröffentlicht: (2024)
von: Siro, Clemencia, et al.
Veröffentlicht: (2024)
Othello is Solved
von: Takizawa, Hiroki
Veröffentlicht: (2023)
von: Takizawa, Hiroki
Veröffentlicht: (2023)
Vision Language Models See What You Want but not What You See
von: Gao, Qingying, et al.
Veröffentlicht: (2024)
von: Gao, Qingying, et al.
Veröffentlicht: (2024)
Understanding Subword Compositionality of Large Language Models
von: Peng, Qiwei, et al.
Veröffentlicht: (2025)
von: Peng, Qiwei, et al.
Veröffentlicht: (2025)
Cognitive Biases in Large Language Models for News Recommendation
von: Lyu, Yougang, et al.
Veröffentlicht: (2024)
von: Lyu, Yougang, et al.
Veröffentlicht: (2024)
Better Language Models Exhibit Higher Visual Alignment
von: Ruthardt, Jona, et al.
Veröffentlicht: (2024)
von: Ruthardt, Jona, et al.
Veröffentlicht: (2024)
Sparse Autoencoders Learn Monosemantic Features in Vision-Language Models
von: Pach, Mateusz, et al.
Veröffentlicht: (2025)
von: Pach, Mateusz, et al.
Veröffentlicht: (2025)
EvalCards: A Framework for Standardized Evaluation Reporting
von: Dhar, Ruchira, et al.
Veröffentlicht: (2025)
von: Dhar, Ruchira, et al.
Veröffentlicht: (2025)
Empirical Evaluation of Progressive Coding for Sparse Autoencoders
von: Peter, Hans, et al.
Veröffentlicht: (2025)
von: Peter, Hans, et al.
Veröffentlicht: (2025)
Goal-Directedness is in the Eye of the Beholder
von: Rajcic, Nina, et al.
Veröffentlicht: (2025)
von: Rajcic, Nina, et al.
Veröffentlicht: (2025)
Brainrot: Deskilling and Addiction are Overlooked AI Risks
von: Chalkidis, Ilias, et al.
Veröffentlicht: (2026)
von: Chalkidis, Ilias, et al.
Veröffentlicht: (2026)
Real-Time Progress Prediction in Reasoning Language Models
von: Raaschou-Jensen, Hans Peter Lyngsøe, et al.
Veröffentlicht: (2025)
von: Raaschou-Jensen, Hans Peter Lyngsøe, et al.
Veröffentlicht: (2025)
Orcheo: A Modular Full-Stack Platform for Conversational Search
von: Jiang, Shaojie, et al.
Veröffentlicht: (2026)
von: Jiang, Shaojie, et al.
Veröffentlicht: (2026)
Beyond Technocratic XAI: The Who, What & How in Explanation Design
von: Dhar, Ruchira, et al.
Veröffentlicht: (2025)
von: Dhar, Ruchira, et al.
Veröffentlicht: (2025)
Realist and Pluralist Conceptions of Intelligence and Their Implications on AI Research
von: Oldenburg, Ninell, et al.
Veröffentlicht: (2025)
von: Oldenburg, Ninell, et al.
Veröffentlicht: (2025)
Demonstrating and Reducing Shortcuts in Vision-Language Representation Learning
von: Bleeker, Maurits, et al.
Veröffentlicht: (2024)
von: Bleeker, Maurits, et al.
Veröffentlicht: (2024)
Unlearning-based Neural Interpretations
von: Choi, Ching Lam, et al.
Veröffentlicht: (2024)
von: Choi, Ching Lam, et al.
Veröffentlicht: (2024)
Large Language Models for Next Point-of-Interest Recommendation
von: Li, Peibo, et al.
Veröffentlicht: (2024)
von: Li, Peibo, et al.
Veröffentlicht: (2024)
SubSearch: Intermediate Rewards for Unsupervised Guided Reasoning in Complex Retrieval
von: Petcu, Roxana, et al.
Veröffentlicht: (2026)
von: Petcu, Roxana, et al.
Veröffentlicht: (2026)
Does Instruction Tuning Make LLMs More Consistent?
von: Fierro, Constanza, et al.
Veröffentlicht: (2024)
von: Fierro, Constanza, et al.
Veröffentlicht: (2024)
Evaluating Adjective-Noun Compositionality in LLMs: Functional vs Representational Perspectives
von: Dhar, Ruchira, et al.
Veröffentlicht: (2026)
von: Dhar, Ruchira, et al.
Veröffentlicht: (2026)
Automatically Finding Rule-Based Neurons in OthelloGPT
von: Singh, Aditya, et al.
Veröffentlicht: (2025)
von: Singh, Aditya, et al.
Veröffentlicht: (2025)
Thinking in Frames: How Visual Context and Test-Time Scaling Empower Video Reasoning
von: Li, Chengzu, et al.
Veröffentlicht: (2026)
von: Li, Chengzu, et al.
Veröffentlicht: (2026)
From Words to Worlds: Compositionality for Cognitive Architectures
von: Dhar, Ruchira, et al.
Veröffentlicht: (2024)
von: Dhar, Ruchira, et al.
Veröffentlicht: (2024)
MEG: Medical Knowledge-Augmented Large Language Models for Question Answering
von: Cabello, Laura, et al.
Veröffentlicht: (2024)
von: Cabello, Laura, et al.
Veröffentlicht: (2024)
KnowTuning: Knowledge-aware Fine-tuning for Large Language Models
von: Lyu, Yougang, et al.
Veröffentlicht: (2024)
von: Lyu, Yougang, et al.
Veröffentlicht: (2024)
Ranked List Truncation for Large Language Model-based Re-Ranking
von: Meng, Chuan, et al.
Veröffentlicht: (2024)
von: Meng, Chuan, et al.
Veröffentlicht: (2024)
Comprehensive Reassessment of Large-Scale Evaluation Outcomes in LLMs: A Multifaceted Statistical Approach
von: Sun, Kun, et al.
Veröffentlicht: (2024)
von: Sun, Kun, et al.
Veröffentlicht: (2024)
Query Performance Prediction using Relevance Judgments Generated by Large Language Models
von: Meng, Chuan, et al.
Veröffentlicht: (2024)
von: Meng, Chuan, et al.
Veröffentlicht: (2024)
RAVENEA: A Benchmark for Multimodal Retrieval-Augmented Visual Culture Understanding
von: Li, Jiaang, et al.
Veröffentlicht: (2025)
von: Li, Jiaang, et al.
Veröffentlicht: (2025)
Large Language Models Could Be Rote Learners
von: Xu, Yuyang, et al.
Veröffentlicht: (2025)
von: Xu, Yuyang, et al.
Veröffentlicht: (2025)
What Would an LLM Do? Evaluating Large Language Models for Policymaking to Alleviate Homelessness
von: Coz, Pierre Le, et al.
Veröffentlicht: (2025)
von: Coz, Pierre Le, et al.
Veröffentlicht: (2025)
"I See What You Did There": Can Large Vision-Language Models Understand Multimodal Puns?
von: Xu, Naen, et al.
Veröffentlicht: (2026)
von: Xu, Naen, et al.
Veröffentlicht: (2026)
Word Order and World Knowledge
von: Zhao, Qinghua, et al.
Veröffentlicht: (2024)
von: Zhao, Qinghua, et al.
Veröffentlicht: (2024)
The Othello AI Arena: Evaluating Intelligent Systems Through Limited-Time Adaptation to Unseen Boards
von: Kim, Sundong
Veröffentlicht: (2025)
von: Kim, Sundong
Veröffentlicht: (2025)
Tutorial on Reasoning for IR & IR for Reasoning
von: Hoveyda, Mohanna, et al.
Veröffentlicht: (2026)
von: Hoveyda, Mohanna, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
Revisiting the Othello World Model Hypothesis
von: Yuan, Yifei, et al.
Veröffentlicht: (2025) -
Do Vision and Language Models Share Concepts? A Vector Space Alignment Study
von: Li, Jiaang, et al.
Veröffentlicht: (2023) -
Can Small Agents Collaborate to Beat a Single Large Language Model?
von: Żywot, Agata, et al.
Veröffentlicht: (2026) -
ChatMotion: A Multimodal Multi-Agent for Human Motion Analysis
von: Li, Lei, et al.
Veröffentlicht: (2025) -
AGENT-CQ: Automatic Generation and Evaluation of Clarifying Questions for Conversational Search with LLMs
von: Siro, Clemencia, et al.
Veröffentlicht: (2024)