Can LLM Teams Play What? Where? When?
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Kotelnikova, Anastasia, Byzov, Viktor, Dolzhenkova, Maria, Kotelnikov, Evgeny |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Optimizing Multimodal Language Models through Attention-based Interpretability
von: Sergeev, Alexander, et al.
Veröffentlicht: (2025)
von: Sergeev, Alexander, et al.
Veröffentlicht: (2025)
Improving Small Language Models for Code Generation with Reinforcement Learning from Verification Feedback
von: Skopin, Egor, et al.
Veröffentlicht: (2026)
von: Skopin, Egor, et al.
Veröffentlicht: (2026)
I've got the "Answer"! Interpretation of LLMs Hidden States in Question Answering
von: Goloviznina, Valeriya, et al.
Veröffentlicht: (2024)
von: Goloviznina, Valeriya, et al.
Veröffentlicht: (2024)
Talking to Data: Designing Smart Assistants for Humanities Databases
von: Sergeev, Alexander, et al.
Veröffentlicht: (2025)
von: Sergeev, Alexander, et al.
Veröffentlicht: (2025)
Do LLMs Understand Why We Write Diaries? A Method for Purpose Extraction and Clustering
von: Goloviznina, Valeriya, et al.
Veröffentlicht: (2025)
von: Goloviznina, Valeriya, et al.
Veröffentlicht: (2025)
Beamforming-LLM: What, Where and When Did I Miss?
von: Choudhari, Vishal
Veröffentlicht: (2025)
von: Choudhari, Vishal
Veröffentlicht: (2025)
STALE: Can LLM Agents Know When Their Memories Are No Longer Valid?
von: Chao, Hanxiang, et al.
Veröffentlicht: (2026)
von: Chao, Hanxiang, et al.
Veröffentlicht: (2026)
Is Reference Necessary in the Evaluation of NLG Systems? When and Where?
von: Sheng, Shuqian, et al.
Veröffentlicht: (2024)
von: Sheng, Shuqian, et al.
Veröffentlicht: (2024)
GLoRe: When, Where, and How to Improve LLM Reasoning via Global and Local Refinements
von: Havrilla, Alex, et al.
Veröffentlicht: (2024)
von: Havrilla, Alex, et al.
Veröffentlicht: (2024)
The Persuasion Paradox: When LLM Explanations Fail to Improve Human-AI Team Performance
von: Cohen, Ruth, et al.
Veröffentlicht: (2026)
von: Cohen, Ruth, et al.
Veröffentlicht: (2026)
When Can We Trust LLM Graders? Calibrating Confidence for Automated Assessment
von: Ferrer, Robinson, et al.
Veröffentlicht: (2026)
von: Ferrer, Robinson, et al.
Veröffentlicht: (2026)
LLM-as-a-Judge & Reward Model: What They Can and Cannot Do
von: Son, Guijin, et al.
Veröffentlicht: (2024)
von: Son, Guijin, et al.
Veröffentlicht: (2024)
Where does an LLM begin computing an instruction?
von: Pola, Aditya, et al.
Veröffentlicht: (2025)
von: Pola, Aditya, et al.
Veröffentlicht: (2025)
From Unstructured Data to In-Context Learning: Exploring What Tasks Can Be Learned and When
von: Wibisono, Kevin Christian, et al.
Veröffentlicht: (2024)
von: Wibisono, Kevin Christian, et al.
Veröffentlicht: (2024)
When Parts Are Greater Than Sums: Individual LLM Components Can Outperform Full Models
von: Chang, Ting-Yun, et al.
Veröffentlicht: (2024)
von: Chang, Ting-Yun, et al.
Veröffentlicht: (2024)
Plan-Then-Execute: An Empirical Study of User Trust and Team Performance When Using LLM Agents As A Daily Assistant
von: He, Gaole, et al.
Veröffentlicht: (2025)
von: He, Gaole, et al.
Veröffentlicht: (2025)
Input Matters: Evaluating Input Structure's Impact on LLM Summaries of Sports Play-by-Play
von: Sundararajan, Barkavi, et al.
Veröffentlicht: (2025)
von: Sundararajan, Barkavi, et al.
Veröffentlicht: (2025)
Team-Based Self-Play With Dual Adaptive Weighting for Fine-Tuning LLMs
von: Li, Wu, et al.
Veröffentlicht: (2026)
von: Li, Wu, et al.
Veröffentlicht: (2026)
Where Are We? Evaluating LLM Performance on African Languages
von: Adebara, Ife, et al.
Veröffentlicht: (2025)
von: Adebara, Ife, et al.
Veröffentlicht: (2025)
STAR-Teaming: A Strategy-Response Multiplex Network Approach to Automated LLM Red Teaming
von: Jung, MinJae, et al.
Veröffentlicht: (2026)
von: Jung, MinJae, et al.
Veröffentlicht: (2026)
When and What to Ask: AskBench and Rubric-Guided RLVR for LLM Clarification
von: Zhao, Jiale, et al.
Veröffentlicht: (2026)
von: Zhao, Jiale, et al.
Veröffentlicht: (2026)
What an Elegant Bridge: Multilingual LLMs are Biased Similarly in Different Languages
von: Mihaylov, Viktor, et al.
Veröffentlicht: (2024)
von: Mihaylov, Viktor, et al.
Veröffentlicht: (2024)
When LLMs Team Up: The Emergence of Collaborative Affective Computing
von: Lai, Wenna, et al.
Veröffentlicht: (2025)
von: Lai, Wenna, et al.
Veröffentlicht: (2025)
Fantastic Biases (What are They) and Where to Find Them
von: Barriere, Valentin
Veröffentlicht: (2024)
von: Barriere, Valentin
Veröffentlicht: (2024)
Text Classification in the LLM Era -- Where do we stand?
von: Vajjala, Sowmya, et al.
Veröffentlicht: (2025)
von: Vajjala, Sowmya, et al.
Veröffentlicht: (2025)
From Role-Play to Drama-Interaction: An LLM Solution
von: Wu, Weiqi, et al.
Veröffentlicht: (2024)
von: Wu, Weiqi, et al.
Veröffentlicht: (2024)
Pay What LLM Wants: Can LLM Simulate Economics Experiment with 522 Real-human Persona?
von: Choi, Junhyuk, et al.
Veröffentlicht: (2025)
von: Choi, Junhyuk, et al.
Veröffentlicht: (2025)
Where is the Mind? Persona Vectors and LLM Individuation
von: Beckmann, Pierre, et al.
Veröffentlicht: (2026)
von: Beckmann, Pierre, et al.
Veröffentlicht: (2026)
Do Proactive Agents Really Need an LLM to Decide When to Wake and What to Anchor?
von: Liu, Xiaoze, et al.
Veröffentlicht: (2026)
von: Liu, Xiaoze, et al.
Veröffentlicht: (2026)
When Can We Trust LLMs in Mental Health? Large-Scale Benchmarks for Reliable LLM Evaluation
von: Badawi, Abeer, et al.
Veröffentlicht: (2025)
von: Badawi, Abeer, et al.
Veröffentlicht: (2025)
When and Where Did it Happen? An Encoder-Decoder Model to Identify Scenario Context
von: Noriega-Atala, Enrique, et al.
Veröffentlicht: (2024)
von: Noriega-Atala, Enrique, et al.
Veröffentlicht: (2024)
When to Think, When to Speak: Learning Disclosure Policies for LLM Reasoning
von: Wei, Jiaqi, et al.
Veröffentlicht: (2026)
von: Wei, Jiaqi, et al.
Veröffentlicht: (2026)
TroubleLLM: Align to Red Team Expert
von: Xu, Zhuoer, et al.
Veröffentlicht: (2024)
von: Xu, Zhuoer, et al.
Veröffentlicht: (2024)
LLMs Can't Play Hangman: On the Necessity of a Private Working Memory for Language Agents
von: Baldelli, Davide, et al.
Veröffentlicht: (2026)
von: Baldelli, Davide, et al.
Veröffentlicht: (2026)
Empathy Is Not What Changed: Clinical Assessment of Psychological Safety Across GPT Model Generations
von: Keeman, Michael, et al.
Veröffentlicht: (2026)
von: Keeman, Michael, et al.
Veröffentlicht: (2026)
When or What? Understanding Consumer Engagement on Digital Platforms
von: Wu, Jingyi, et al.
Veröffentlicht: (2025)
von: Wu, Jingyi, et al.
Veröffentlicht: (2025)
Decoupling the "What" and "Where" With Polar Coordinate Positional Embeddings
von: Gopalakrishnan, Anand, et al.
Veröffentlicht: (2025)
von: Gopalakrishnan, Anand, et al.
Veröffentlicht: (2025)
When One LLM Drools, Multi-LLM Collaboration Rules
von: Feng, Shangbin, et al.
Veröffentlicht: (2025)
von: Feng, Shangbin, et al.
Veröffentlicht: (2025)
Maximizing Local Entropy Where It Matters: Prefix-Aware Localized LLM Unlearning
von: Zhai, Naixin, et al.
Veröffentlicht: (2026)
von: Zhai, Naixin, et al.
Veröffentlicht: (2026)
Semantic Deception: When Reasoning Models Can't Compute an Addition
von: de Leeuw, Nathaniël, et al.
Veröffentlicht: (2025)
von: de Leeuw, Nathaniël, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Optimizing Multimodal Language Models through Attention-based Interpretability
von: Sergeev, Alexander, et al.
Veröffentlicht: (2025) -
Improving Small Language Models for Code Generation with Reinforcement Learning from Verification Feedback
von: Skopin, Egor, et al.
Veröffentlicht: (2026) -
I've got the "Answer"! Interpretation of LLMs Hidden States in Question Answering
von: Goloviznina, Valeriya, et al.
Veröffentlicht: (2024) -
Talking to Data: Designing Smart Assistants for Humanities Databases
von: Sergeev, Alexander, et al.
Veröffentlicht: (2025) -
Do LLMs Understand Why We Write Diaries? A Method for Purpose Extraction and Clustering
von: Goloviznina, Valeriya, et al.
Veröffentlicht: (2025)