Probing the Multi-turn Planning Capabilities of LLMs via 20 Question Games
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Zhang, Yizhe, Lu, Jiarui, Jaitly, Navdeep |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2023
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
MotionTeller: Multi-modal Integration of Wearable Time-Series with LLMs for Health and Behavioral Understanding
von: Zhang, Aiwei, et al.
Veröffentlicht: (2025)
von: Zhang, Aiwei, et al.
Veröffentlicht: (2025)
HybridQuestion: Human-AI Collaboration for Identifying High-Impact Research Questions
von: Zhao, Keyu, et al.
Veröffentlicht: (2025)
von: Zhao, Keyu, et al.
Veröffentlicht: (2025)
Comparing Exploration-Exploitation Strategies of LLMs and Humans: Insights from Standard Multi-armed Bandit Experiments
von: Zhang, Ziyuan, et al.
Veröffentlicht: (2025)
von: Zhang, Ziyuan, et al.
Veröffentlicht: (2025)
Conversational Planning for Personal Plans
von: Christakopoulou, Konstantina, et al.
Veröffentlicht: (2025)
von: Christakopoulou, Konstantina, et al.
Veröffentlicht: (2025)
Teaching According to Students' Aptitude: Personalized Mathematics Tutoring via Persona-, Memory-, and Forgetting-Aware LLMs
von: Wu, Yang, et al.
Veröffentlicht: (2025)
von: Wu, Yang, et al.
Veröffentlicht: (2025)
Playing 20 Question Game with Policy-Based Reinforcement Learning
von: Hu, Huang, et al.
Veröffentlicht: (2018)
von: Hu, Huang, et al.
Veröffentlicht: (2018)
Beyond Semantic Similarity: A Component-Wise Evaluation Framework for Medical Question Answering Systems with Health Equity Implications
von: Sakib, Abu Noman Md, et al.
Veröffentlicht: (2026)
von: Sakib, Abu Noman Md, et al.
Veröffentlicht: (2026)
Harmonic LLMs are Trustworthy
von: Kersting, Nicholas S., et al.
Veröffentlicht: (2024)
von: Kersting, Nicholas S., et al.
Veröffentlicht: (2024)
KGLens: Towards Efficient and Effective Knowledge Probing of Large Language Models with Knowledge Graphs
von: Zheng, Shangshang, et al.
Veröffentlicht: (2023)
von: Zheng, Shangshang, et al.
Veröffentlicht: (2023)
Model-in-the-Loop (MILO): Accelerating Multimodal AI Data Annotation with LLMs
von: Wang, Yifan, et al.
Veröffentlicht: (2024)
von: Wang, Yifan, et al.
Veröffentlicht: (2024)
Interaction Dynamics as a Reward Signal for LLMs
von: Gooding, Sian, et al.
Veröffentlicht: (2025)
von: Gooding, Sian, et al.
Veröffentlicht: (2025)
LLMs for XAI: Future Directions for Explaining Explanations
von: Zytek, Alexandra, et al.
Veröffentlicht: (2024)
von: Zytek, Alexandra, et al.
Veröffentlicht: (2024)
Generative UI: LLMs are Effective UI Generators
von: Leviathan, Yaniv, et al.
Veröffentlicht: (2026)
von: Leviathan, Yaniv, et al.
Veröffentlicht: (2026)
Multimodal Fusion with LLMs for Engagement Prediction in Natural Conversation
von: Ma, Cheng Charles, et al.
Veröffentlicht: (2024)
von: Ma, Cheng Charles, et al.
Veröffentlicht: (2024)
PREF: Reference-Free Evaluation of Personalised Text Generation in LLMs
von: Fu, Xiao, et al.
Veröffentlicht: (2025)
von: Fu, Xiao, et al.
Veröffentlicht: (2025)
AIRepr: An Analyst-Inspector Framework for Evaluating Reproducibility of LLMs in Data Science
von: Zeng, Qiuhai, et al.
Veröffentlicht: (2025)
von: Zeng, Qiuhai, et al.
Veröffentlicht: (2025)
Are You Being Tracked? Discover the Power of Zero-Shot Trajectory Tracing with LLMs!
von: Yang, Huanqi, et al.
Veröffentlicht: (2024)
von: Yang, Huanqi, et al.
Veröffentlicht: (2024)
Demo: Statistically Significant Results On Biases and Errors of LLMs Do Not Guarantee Generalizable Results
von: Liu, Jonathan, et al.
Veröffentlicht: (2025)
von: Liu, Jonathan, et al.
Veröffentlicht: (2025)
Observations on LLMs for Telecom Domain: Capabilities and Limitations
von: Soman, Sumit, et al.
Veröffentlicht: (2023)
von: Soman, Sumit, et al.
Veröffentlicht: (2023)
Evaluation of LLMs-based Hidden States as Author Representations for Psychological Human-Centered NLP Tasks
von: Soni, Nikita, et al.
Veröffentlicht: (2025)
von: Soni, Nikita, et al.
Veröffentlicht: (2025)
Olapa-MCoT: Enhancing the Chinese Mathematical Reasoning Capability of LLMs
von: Zhu, Shaojie, et al.
Veröffentlicht: (2023)
von: Zhu, Shaojie, et al.
Veröffentlicht: (2023)
LingVarBench: Benchmarking LLMs on Entity Recognitions and Linguistic Verbalization Patterns in Phone-Call Transcripts
von: Mohammadi, Seyedali, et al.
Veröffentlicht: (2025)
von: Mohammadi, Seyedali, et al.
Veröffentlicht: (2025)
Can LLM feedback enhance review quality? A randomized study of 20K reviews at ICLR 2025
von: Thakkar, Nitya, et al.
Veröffentlicht: (2025)
von: Thakkar, Nitya, et al.
Veröffentlicht: (2025)
SAGE: Steering Dialog Generation with Future-Aware State-Action Augmentation
von: Zhang, Yizhe, et al.
Veröffentlicht: (2025)
von: Zhang, Yizhe, et al.
Veröffentlicht: (2025)
Asking For It: Question-Answering for Predicting Rule Infractions in Online Content Moderation
von: Samory, Mattia, et al.
Veröffentlicht: (2025)
von: Samory, Mattia, et al.
Veröffentlicht: (2025)
Think Like a Person Before Responding: A Multi-Faceted Evaluation of Persona-Guided LLMs for Countering Hate
von: Ngueajio, Mikel K., et al.
Veröffentlicht: (2025)
von: Ngueajio, Mikel K., et al.
Veröffentlicht: (2025)
AutoAct: Automatic Agent Learning from Scratch for QA via Self-Planning
von: Qiao, Shuofei, et al.
Veröffentlicht: (2024)
von: Qiao, Shuofei, et al.
Veröffentlicht: (2024)
Towards Actionable Pedagogical Feedback: A Multi-Perspective Analysis of Mathematics Teaching and Tutoring Dialogue
von: Naim, Jannatun, et al.
Veröffentlicht: (2025)
von: Naim, Jannatun, et al.
Veröffentlicht: (2025)
Llms, Virtual Users, and Bias: Predicting Any Survey Question Without Human Data
von: Sinacola, Enzo, et al.
Veröffentlicht: (2025)
von: Sinacola, Enzo, et al.
Veröffentlicht: (2025)
ProVox: Personalization and Proactive Planning for Situated Human-Robot Collaboration
von: Grannen, Jennifer, et al.
Veröffentlicht: (2025)
von: Grannen, Jennifer, et al.
Veröffentlicht: (2025)
Implicit Personalization in Language Models: A Systematic Study
von: Jin, Zhijing, et al.
Veröffentlicht: (2024)
von: Jin, Zhijing, et al.
Veröffentlicht: (2024)
Never Start from Scratch: Expediting On-Device LLM Personalization via Explainable Model Selection
von: Wang, Haoming, et al.
Veröffentlicht: (2025)
von: Wang, Haoming, et al.
Veröffentlicht: (2025)
Can LLMs Replace Economic Choice Prediction Labs? The Case of Language-based Persuasion Games
von: Shapira, Eilam, et al.
Veröffentlicht: (2024)
von: Shapira, Eilam, et al.
Veröffentlicht: (2024)
LLMs as Writing Assistants: Exploring Perspectives on Sense of Ownership and Reasoning
von: Wasi, Azmine Toushik, et al.
Veröffentlicht: (2024)
von: Wasi, Azmine Toushik, et al.
Veröffentlicht: (2024)
Sociodemographic Prompting is Not Yet an Effective Approach for Simulating Subjective Judgments with LLMs
von: Sun, Huaman, et al.
Veröffentlicht: (2023)
von: Sun, Huaman, et al.
Veröffentlicht: (2023)
Heterogeneous Value Alignment Evaluation for Large Language Models
von: Zhang, Zhaowei, et al.
Veröffentlicht: (2023)
von: Zhang, Zhaowei, et al.
Veröffentlicht: (2023)
LaDiR: Latent Diffusion Enhances LLMs for Text Reasoning
von: Kang, Haoqiang, et al.
Veröffentlicht: (2025)
von: Kang, Haoqiang, et al.
Veröffentlicht: (2025)
LLMs May Not Be Human-Level Players, But They Can Be Testers: Measuring Game Difficulty with LLM Agents
von: Xiao, Chang, et al.
Veröffentlicht: (2024)
von: Xiao, Chang, et al.
Veröffentlicht: (2024)
SPRIG: Improving Large Language Model Performance by System Prompt Optimization
von: Zhang, Lechen, et al.
Veröffentlicht: (2024)
von: Zhang, Lechen, et al.
Veröffentlicht: (2024)
KnowAgent: Knowledge-Augmented Planning for LLM-Based Agents
von: Zhu, Yuqi, et al.
Veröffentlicht: (2024)
von: Zhu, Yuqi, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
MotionTeller: Multi-modal Integration of Wearable Time-Series with LLMs for Health and Behavioral Understanding
von: Zhang, Aiwei, et al.
Veröffentlicht: (2025) -
HybridQuestion: Human-AI Collaboration for Identifying High-Impact Research Questions
von: Zhao, Keyu, et al.
Veröffentlicht: (2025) -
Comparing Exploration-Exploitation Strategies of LLMs and Humans: Insights from Standard Multi-armed Bandit Experiments
von: Zhang, Ziyuan, et al.
Veröffentlicht: (2025) -
Conversational Planning for Personal Plans
von: Christakopoulou, Konstantina, et al.
Veröffentlicht: (2025) -
Teaching According to Students' Aptitude: Personalized Mathematics Tutoring via Persona-, Memory-, and Forgetting-Aware LLMs
von: Wu, Yang, et al.
Veröffentlicht: (2025)