How to Interpret Agent Behavior
Fuente:
arXiv
Guardado en:
| Autores principales: | Gao, Jie, Sun, Kaiser, Huang, Jen-tse, Van Koevering, Katherine, Ji, Sijie, Huang, Heyuan, Shi, Weiyan, Lu, Zhuoran, Xiao, Ziang, Khashabi, Daniel, Dredze, Mark |
|---|---|
| Formato: | Preprint |
| Publicado: |
2026
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
On the Failure of Latent State Persistence in Large Language Models
por: Huang, Jen-tse, et al.
Publicado: (2025)
por: Huang, Jen-tse, et al.
Publicado: (2025)
How Random is Random? Evaluating the Randomness and Humaness of LLMs' Coin Flips
por: Van Koevering, Katherine, et al.
Publicado: (2024)
por: Van Koevering, Katherine, et al.
Publicado: (2024)
Evaluating the Evaluators: Are readability metrics good measures of readability?
por: Cachola, Isabel, et al.
Publicado: (2025)
por: Cachola, Isabel, et al.
Publicado: (2025)
Knowing But Not Doing: Convergent Morality and Divergent Action in LLMs
por: Huang, Jen-tse, et al.
Publicado: (2026)
por: Huang, Jen-tse, et al.
Publicado: (2026)
Probing Multimodal Large Language Models on Cognitive Biases in Chinese Short-Video Misinformation
por: Huang, Jen-tse, et al.
Publicado: (2026)
por: Huang, Jen-tse, et al.
Publicado: (2026)
MedScore: Generalizable Factuality Evaluation of Free-Form Medical Answers by Domain-adapted Claim Decomposition and Verification
por: Huang, Heyuan, et al.
Publicado: (2025)
por: Huang, Heyuan, et al.
Publicado: (2025)
Amuro and Char: Analyzing the Relationship between Pre-Training and Fine-Tuning of Large Language Models
por: Sun, Kaiser, et al.
Publicado: (2024)
por: Sun, Kaiser, et al.
Publicado: (2024)
Efficiency with Rigor! A Trustworthy LLM-powered Workflow for Qualitative Data Analysis
por: Gao, Jie, et al.
Publicado: (2025)
por: Gao, Jie, et al.
Publicado: (2025)
What's in a Niche? Migration Patterns in Online Communities
por: Van Koevering, Katherine, et al.
Publicado: (2024)
por: Van Koevering, Katherine, et al.
Publicado: (2024)
Same Verdict, Different Reasons: LLM-as-a-Judge and Clinician Disagreement on Medical Chatbot Completeness
por: DeLucia, Alexandra, et al.
Publicado: (2026)
por: DeLucia, Alexandra, et al.
Publicado: (2026)
Task Matters: Knowledge Requirements Shape LLM Responses to Context-Memory Conflict
por: Sun, Kaiser, et al.
Publicado: (2025)
por: Sun, Kaiser, et al.
Publicado: (2025)
Artificial Intolerance: Stigmatizing Language in Clinical Documentation Skews Large Language Model Decision-Making
por: Huang, Jen-tse, et al.
Publicado: (2026)
por: Huang, Jen-tse, et al.
Publicado: (2026)
Safe and Interpretable Multimodal Path Planning for Multi-Agent Cooperation
por: Shi, Haojun, et al.
Publicado: (2026)
por: Shi, Haojun, et al.
Publicado: (2026)
FairCoder: Evaluating Social Bias of LLMs in Code Generation
por: Du, Yongkang, et al.
Publicado: (2025)
por: Du, Yongkang, et al.
Publicado: (2025)
BIASINSPECTOR: Detecting Bias in Structured Data through LLM Agents
por: Li, Haoxuan, et al.
Publicado: (2025)
por: Li, Haoxuan, et al.
Publicado: (2025)
What do Language Models Learn and When? The Implicit Curriculum Hypothesis
por: Liu, Emmy, et al.
Publicado: (2026)
por: Liu, Emmy, et al.
Publicado: (2026)
DnDScore: Decontextualization and Decomposition for Factuality Verification in Long-Form Text Generation
por: Wanner, Miriam, et al.
Publicado: (2024)
por: Wanner, Miriam, et al.
Publicado: (2024)
InCharacter: Evaluating Personality Fidelity in Role-Playing Agents through Psychological Interviews
por: Wang, Xintao, et al.
Publicado: (2023)
por: Wang, Xintao, et al.
Publicado: (2023)
Weighted GKAT: Completeness and Complexity
por: Van Koevering, Spencer, et al.
Publicado: (2025)
por: Van Koevering, Spencer, et al.
Publicado: (2025)
Transferring Fairness using Multi-Task Learning with Limited Demographic Information
por: Aguirre, Carlos, et al.
Publicado: (2023)
por: Aguirre, Carlos, et al.
Publicado: (2023)
InterIntent: Investigating Social Intelligence of LLMs via Intention Understanding in an Interactive Game Context
por: Liu, Ziyi, et al.
Publicado: (2024)
por: Liu, Ziyi, et al.
Publicado: (2024)
CodeCrash: Exposing LLM Fragility to Misleading Natural Language in Code Reasoning
por: Lam, Man Ho, et al.
Publicado: (2025)
por: Lam, Man Ho, et al.
Publicado: (2025)
AI Sees Your Location, But With A Bias Toward The Wealthy World
por: Huang, Jingyuan, et al.
Publicado: (2025)
por: Huang, Jingyuan, et al.
Publicado: (2025)
A Survey on the Safety and Security Threats of Computer-Using Agents: JARVIS or Ultron?
por: Chen, Ada, et al.
Publicado: (2025)
por: Chen, Ada, et al.
Publicado: (2025)
The PIMMUR Principles: Ensuring Validity in Collective Behavior of LLM Societies
por: Zhou, Jiaxu, et al.
Publicado: (2025)
por: Zhou, Jiaxu, et al.
Publicado: (2025)
How Well Can LLMs Echo Us? Evaluating AI Chatbots' Role-Play Ability with ECHO
por: Ng, Man Tik, et al.
Publicado: (2024)
por: Ng, Man Tik, et al.
Publicado: (2024)
The Chameleon's Limit: Investigating Persona Collapse and Homogenization in Large Language Models
por: Xiao, Yunze, et al.
Publicado: (2026)
por: Xiao, Yunze, et al.
Publicado: (2026)
Can one size fit all?: Measuring Failure in Multi-Document Summarization Domain Transfer
por: DeLucia, Alexandra, et al.
Publicado: (2025)
por: DeLucia, Alexandra, et al.
Publicado: (2025)
RAG LLMs are Not Safer: A Safety Analysis of Retrieval-Augmented Generation for Large Language Models
por: An, Bang, et al.
Publicado: (2025)
por: An, Bang, et al.
Publicado: (2025)
From Text to Trust: Empowering AI-assisted Decision Making with Adaptive LLM-powered Analysis
por: Li, Zhuoyan, et al.
Publicado: (2025)
por: Li, Zhuoyan, et al.
Publicado: (2025)
UniDebugger: Hierarchical Multi-Agent Framework for Unified Software Debugging
por: Lee, Cheryl, et al.
Publicado: (2024)
por: Lee, Cheryl, et al.
Publicado: (2024)
On the Resilience of LLM-Based Multi-Agent Collaboration with Faulty Agents
por: Huang, Jen-tse, et al.
Publicado: (2024)
por: Huang, Jen-tse, et al.
Publicado: (2024)
How Far Are We on the Decision-Making of LLMs? Evaluating LLMs' Gaming Ability in Multi-Agent Environments
por: Huang, Jen-tse, et al.
Publicado: (2024)
por: Huang, Jen-tse, et al.
Publicado: (2024)
On the Shortcut Learning in Multilingual Neural Machine Translation
por: Wang, Wenxuan, et al.
Publicado: (2024)
por: Wang, Wenxuan, et al.
Publicado: (2024)
Diversity-Enhanced Reasoning for Subjective Questions
por: Wang, Yumeng, et al.
Publicado: (2025)
por: Wang, Yumeng, et al.
Publicado: (2025)
GPT-4 Is Too Smart To Be Safe: Stealthy Chat with LLMs via Cipher
por: Yuan, Youliang, et al.
Publicado: (2023)
por: Yuan, Youliang, et al.
Publicado: (2023)
SWE-Chain: Benchmarking Coding Agents on Chained Release-Level Package Upgrades
por: Lam, Man Ho, et al.
Publicado: (2026)
por: Lam, Man Ho, et al.
Publicado: (2026)
WorldAPIs: The World Is Worth How Many APIs? A Thought Experiment
por: Ou, Jiefu, et al.
Publicado: (2024)
por: Ou, Jiefu, et al.
Publicado: (2024)
Can Optimization Trajectories Explain Multi-Task Transfer?
por: Mueller, David, et al.
Publicado: (2024)
por: Mueller, David, et al.
Publicado: (2024)
Evaluating Implicit Biases in LLM Reasoning through Logic Grid Puzzles
por: Jahara, Fatima, et al.
Publicado: (2025)
por: Jahara, Fatima, et al.
Publicado: (2025)
Ejemplares similares
-
On the Failure of Latent State Persistence in Large Language Models
por: Huang, Jen-tse, et al.
Publicado: (2025) -
How Random is Random? Evaluating the Randomness and Humaness of LLMs' Coin Flips
por: Van Koevering, Katherine, et al.
Publicado: (2024) -
Evaluating the Evaluators: Are readability metrics good measures of readability?
por: Cachola, Isabel, et al.
Publicado: (2025) -
Knowing But Not Doing: Convergent Morality and Divergent Action in LLMs
por: Huang, Jen-tse, et al.
Publicado: (2026) -
Probing Multimodal Large Language Models on Cognitive Biases in Chinese Short-Video Misinformation
por: Huang, Jen-tse, et al.
Publicado: (2026)