MT-OSC: Path for LLMs that Get Lost in Multi-Turn Conversation
Fuente:
arXiv
Saved in:
| Main Authors: | Singh, Jyotika, Tu, Fang, Ballesteros, Miguel, Sun, Weiyi, Ghoshal, Sandip, Yuan, Michelle, Benajiba, Yassine, Ravi, Sujith, Roth, Dan |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
JTPRO: A Joint Tool-Prompt Reflective Optimization Framework for Language Agents
by: Ghoshal, Sandip, et al.
Published: (2026)
by: Ghoshal, Sandip, et al.
Published: (2026)
Can LLMs Narrate Tabular Data? An Evaluation Framework for Natural Language Representations of Text-to-SQL System Outputs
by: Singh, Jyotika, et al.
Published: (2025)
by: Singh, Jyotika, et al.
Published: (2025)
Barriers to Discrete Reasoning with Transformers: A Survey Across Depth, Exactness, and Bandwidth
by: Yuan, Michelle, et al.
Published: (2026)
by: Yuan, Michelle, et al.
Published: (2026)
DiffuMask: Diffusion Language Model for Token-level Prompt Pruning
by: Zheng, Caleb, et al.
Published: (2026)
by: Zheng, Caleb, et al.
Published: (2026)
GSM-SEM: Benchmark and Framework for Generating Semantically Variant Augmentations
by: Singh, Jyotika, et al.
Published: (2026)
by: Singh, Jyotika, et al.
Published: (2026)
LLMs Get Lost In Multi-Turn Conversation
by: Laban, Philippe, et al.
Published: (2025)
by: Laban, Philippe, et al.
Published: (2025)
Intent Mismatch Causes LLMs to Get Lost in Multi-Turn Conversation
by: Liu, Geng, et al.
Published: (2026)
by: Liu, Geng, et al.
Published: (2026)
Arabic Named Entity Recognition
by: Yassine Benajiba
Published: (2010)
by: Yassine Benajiba
Published: (2010)
Active Evaluation Acquisition for Efficient LLM Benchmarking
by: Li, Yang, et al.
Published: (2024)
by: Li, Yang, et al.
Published: (2024)
RCI: A Score for Evaluating Global and Local Reasoning in Multimodal Benchmarks
by: Agarwal, Amit, et al.
Published: (2025)
by: Agarwal, Amit, et al.
Published: (2025)
Aligning LLMs for Multilingual Consistency in Enterprise Applications
by: Agarwal, Amit, et al.
Published: (2025)
by: Agarwal, Amit, et al.
Published: (2025)
Open Agent Specification (Agent Spec): A Unified Representation for AI Agents
by: Amini, Soufiane, et al.
Published: (2025)
by: Amini, Soufiane, et al.
Published: (2025)
Self-supervised Analogical Learning using Language Models
by: Zhou, Ben, et al.
Published: (2025)
by: Zhou, Ben, et al.
Published: (2025)
Do Image-Text Metrics Respect Semantic Invariances?
by: Agarwal, Amit, et al.
Published: (2026)
by: Agarwal, Amit, et al.
Published: (2026)
MemInsight: Autonomous Memory Augmentation for LLM Agents
by: Salama, Rana, et al.
Published: (2025)
by: Salama, Rana, et al.
Published: (2025)
MedMT-Bench: Can LLMs Memorize and Understand Long Multi-Turn Conversations in Medical Scenarios?
by: Yang, Lin, et al.
Published: (2026)
by: Yang, Lin, et al.
Published: (2026)
MCP2OSC: Parametric Control by Natural Language
by: Fan, Yuan-Yi
Published: (2025)
by: Fan, Yuan-Yi
Published: (2025)
NewsQs: Multi-Source Question Generation for the Inquiring Mind
by: Hwang, Alyssa, et al.
Published: (2024)
by: Hwang, Alyssa, et al.
Published: (2024)
Retracing the Past: LLMs Emit Training Data When They Get Lost
by: Ko, Myeongseob, et al.
Published: (2025)
by: Ko, Myeongseob, et al.
Published: (2025)
From Instructions to Constraints: Language Model Alignment with Automatic Constraint Verification
by: Wang, Fei, et al.
Published: (2024)
by: Wang, Fei, et al.
Published: (2024)
PARL-MT: Learning to Call Functions in Multi-Turn Conversation with Progress Awareness
by: Chai, Huacan, et al.
Published: (2025)
by: Chai, Huacan, et al.
Published: (2025)
Asking Forever: Universal Activations Behind Turn Amplification in Conversational LLMs
by: Coalson, Zachary, et al.
Published: (2026)
by: Coalson, Zachary, et al.
Published: (2026)
General Purpose Verification for Chain of Thought Prompting
by: Vacareanu, Robert, et al.
Published: (2024)
by: Vacareanu, Robert, et al.
Published: (2024)
GraphER: An Efficient Graph-Based Enrichment and Reranking Method for Retrieval-Augmented Generation
by: Miao, Ruizhong, et al.
Published: (2026)
by: Miao, Ruizhong, et al.
Published: (2026)
OraPlan-SQL: A Planning-Centric Framework for Complex Bilingual NL2SQL Reasoning
by: Liu, Marianne Menglin, et al.
Published: (2025)
by: Liu, Marianne Menglin, et al.
Published: (2025)
Towards Long Context Hallucination Detection
by: Liu, Siyi, et al.
Published: (2025)
by: Liu, Siyi, et al.
Published: (2025)
Trifluoromethylation Strategies of Alcohols and Phenols
by: Ravi Pratap Singh, et al.
Published: (2024)
by: Ravi Pratap Singh, et al.
Published: (2024)
Texture or Semantics? Vision-Language Models Get Lost in Font Recognition
by: Li, Zhecheng, et al.
Published: (2025)
by: Li, Zhecheng, et al.
Published: (2025)
BrenneckeLab/OSC-genome_purge-haplotigs: v1.0_Handler_Brennecke.2026 - OSC-genome publication
by: dominik-handler
Published: (2026)
by: dominik-handler
Published: (2026)
ClarifyMT-Bench: Benchmarking and Improving Multi-Turn Clarification for Conversational Large Language Models
by: Luo, Sichun, et al.
Published: (2025)
by: Luo, Sichun, et al.
Published: (2025)
A Study on Leveraging Search and Self-Feedback for Agent Reasoning
by: K, Karthikeyan, et al.
Published: (2025)
by: K, Karthikeyan, et al.
Published: (2025)
Rethinking LLM Uncertainty: A Multi-Agent Approach to Estimating Black-Box Model Uncertainty
by: Feng, Yu, et al.
Published: (2024)
by: Feng, Yu, et al.
Published: (2024)
LLM NL2SQL Robustness: Surface Noise vs. Linguistic Variation in Traditional and Agentic Settings
by: Tu, Lifu, et al.
Published: (2026)
by: Tu, Lifu, et al.
Published: (2026)
BrenneckeLab/OSC-genome_insert-heterozygous-TEs: v1.0_Handler_Brennecke.2026 - OSC-genome publication
by: dominik-handler
Published: (2026)
by: dominik-handler
Published: (2026)
Unraveling and Mitigating Safety Alignment Degradation of Vision-Language Models
by: Liu, Qin, et al.
Published: (2024)
by: Liu, Qin, et al.
Published: (2024)
Don't Get Lost in the Trees: Streamlining LLM Reasoning by Overcoming Tree Search Exploration Pitfalls
by: Wang, Ante, et al.
Published: (2025)
by: Wang, Ante, et al.
Published: (2025)
Large Vision-Language Models Get Lost in Attention
by: Xi, Gongli, et al.
Published: (2026)
by: Xi, Gongli, et al.
Published: (2026)
PAR$^2$-RAG: Planned Active Retrieval and Reasoning for Multi-Hop Question Answering
by: Li, Xingyu, et al.
Published: (2026)
by: Li, Xingyu, et al.
Published: (2026)
BrenneckeLab/OSC-genome_fill-genome-scaffold-gaps: v1.0_Handler_Brennecke.2026 - OSC-genome publication
by: dominik-handler
Published: (2026)
by: dominik-handler
Published: (2026)
Reasoning Is Not All You Need: Examining LLMs for Multi-Turn Mental Health Conversations
by: Chandra, Mohit, et al.
Published: (2025)
by: Chandra, Mohit, et al.
Published: (2025)
Similar Items
-
JTPRO: A Joint Tool-Prompt Reflective Optimization Framework for Language Agents
by: Ghoshal, Sandip, et al.
Published: (2026) -
Can LLMs Narrate Tabular Data? An Evaluation Framework for Natural Language Representations of Text-to-SQL System Outputs
by: Singh, Jyotika, et al.
Published: (2025) -
Barriers to Discrete Reasoning with Transformers: A Survey Across Depth, Exactness, and Bandwidth
by: Yuan, Michelle, et al.
Published: (2026) -
DiffuMask: Diffusion Language Model for Token-level Prompt Pruning
by: Zheng, Caleb, et al.
Published: (2026) -
GSM-SEM: Benchmark and Framework for Generating Semantically Variant Augmentations
by: Singh, Jyotika, et al.
Published: (2026)