What's the plan? Metrics for implicit planning in LLMs and their application to rhyme generation and question answering
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Maar, Jim, Paperno, Denis, McDougall, Callum Stuart, Nanda, Neel |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Internal states before wait modulate reasoning patterns
von: Troitskii, Dmitrii, et al.
Veröffentlicht: (2025)
von: Troitskii, Dmitrii, et al.
Veröffentlicht: (2025)
Towards Best Practices of Activation Patching in Language Models: Metrics and Methods
von: Zhang, Fred, et al.
Veröffentlicht: (2023)
von: Zhang, Fred, et al.
Veröffentlicht: (2023)
Explorations of Self-Repair in Language Models
von: Rushing, Cody, et al.
Veröffentlicht: (2024)
von: Rushing, Cody, et al.
Veröffentlicht: (2024)
Agribot: agriculture-specific question answer system
von: Jain, Naman, et al.
Veröffentlicht: (2025)
von: Jain, Naman, et al.
Veröffentlicht: (2025)
Censored LLMs as a Natural Testbed for Secret Knowledge Elicitation
von: Casademunt, Helena, et al.
Veröffentlicht: (2026)
von: Casademunt, Helena, et al.
Veröffentlicht: (2026)
Multi-step retrieval and reasoning improves radiology question answering with large language models
von: Wind, Sebastian, et al.
Veröffentlicht: (2025)
von: Wind, Sebastian, et al.
Veröffentlicht: (2025)
Do I Know This Entity? Knowledge Awareness and Hallucinations in Language Models
von: Ferrando, Javier, et al.
Veröffentlicht: (2024)
von: Ferrando, Javier, et al.
Veröffentlicht: (2024)
CaLMQA: Exploring culturally specific long-form question answering across 23 languages
von: Arora, Shane, et al.
Veröffentlicht: (2024)
von: Arora, Shane, et al.
Veröffentlicht: (2024)
From text to multimodal: a survey of adversarial example generation in question answering systems
von: Yigit, Gulsum, et al.
Veröffentlicht: (2023)
von: Yigit, Gulsum, et al.
Veröffentlicht: (2023)
Enhancing textual textbook question answering with large language models and retrieval augmented generation
von: Alawwad, Hessa Abdulrahman, et al.
Veröffentlicht: (2024)
von: Alawwad, Hessa Abdulrahman, et al.
Veröffentlicht: (2024)
Thought Branches: Interpreting LLM Reasoning Requires Resampling
von: Macar, Uzay, et al.
Veröffentlicht: (2025)
von: Macar, Uzay, et al.
Veröffentlicht: (2025)
Thought Anchors: Which LLM Reasoning Steps Matter?
von: Bogdan, Paul C., et al.
Veröffentlicht: (2025)
von: Bogdan, Paul C., et al.
Veröffentlicht: (2025)
Scaling sparse feature circuit finding for in-context learning
von: Kharlapenko, Dmitrii, et al.
Veröffentlicht: (2025)
von: Kharlapenko, Dmitrii, et al.
Veröffentlicht: (2025)
Overcoming Sparsity Artifacts in Crosscoders to Interpret Chat-Tuning
von: Minder, Julian, et al.
Veröffentlicht: (2025)
von: Minder, Julian, et al.
Veröffentlicht: (2025)
Understanding the planning of LLM agents: A survey
von: Huang, Xu, et al.
Veröffentlicht: (2024)
von: Huang, Xu, et al.
Veröffentlicht: (2024)
Steering Out-of-Distribution Generalization with Concept Ablation Fine-Tuning
von: Casademunt, Helena, et al.
Veröffentlicht: (2025)
von: Casademunt, Helena, et al.
Veröffentlicht: (2025)
Chain-of-Thought Reasoning In The Wild Is Not Always Faithful
von: Arcuschin, Iván, et al.
Veröffentlicht: (2025)
von: Arcuschin, Iván, et al.
Veröffentlicht: (2025)
Real-Time Detection of Hallucinated Entities in Long-Form Generation
von: Obeso, Oscar, et al.
Veröffentlicht: (2025)
von: Obeso, Oscar, et al.
Veröffentlicht: (2025)
RealMedQA: A pilot biomedical question answering dataset containing realistic clinical questions
von: Kell, Gregory, et al.
Veröffentlicht: (2024)
von: Kell, Gregory, et al.
Veröffentlicht: (2024)
What type of inference is planning?
von: Lázaro-Gredilla, Miguel, et al.
Veröffentlicht: (2024)
von: Lázaro-Gredilla, Miguel, et al.
Veröffentlicht: (2024)
The Impact of Inference Acceleration on Bias of LLMs
von: Kirsten, Elisabeth, et al.
Veröffentlicht: (2024)
von: Kirsten, Elisabeth, et al.
Veröffentlicht: (2024)
ACL-Verbatim: hallucination-free question answering for research
von: Recski, Gábor, et al.
Veröffentlicht: (2026)
von: Recski, Gábor, et al.
Veröffentlicht: (2026)
Confidence Regulation Neurons in Language Models
von: Stolfo, Alessandro, et al.
Veröffentlicht: (2024)
von: Stolfo, Alessandro, et al.
Veröffentlicht: (2024)
Building Production-Ready Probes For Gemini
von: Kramár, János, et al.
Veröffentlicht: (2026)
von: Kramár, János, et al.
Veröffentlicht: (2026)
Refusal in Language Models Is Mediated by a Single Direction
von: Arditi, Andy, et al.
Veröffentlicht: (2024)
von: Arditi, Andy, et al.
Veröffentlicht: (2024)
keqing: knowledge-based question answering is a nature chain-of-thought mentor of LLM
von: Wang, Chaojie, et al.
Veröffentlicht: (2023)
von: Wang, Chaojie, et al.
Veröffentlicht: (2023)
QuestBench: Can LLMs ask the right question to acquire information in reasoning tasks?
von: Li, Belinda Z., et al.
Veröffentlicht: (2025)
von: Li, Belinda Z., et al.
Veröffentlicht: (2025)
Universal Neurons in GPT2 Language Models
von: Gurnee, Wes, et al.
Veröffentlicht: (2024)
von: Gurnee, Wes, et al.
Veröffentlicht: (2024)
Emergent Misalignment is Easy, Narrow Misalignment is Hard
von: Soligo, Anna, et al.
Veröffentlicht: (2026)
von: Soligo, Anna, et al.
Veröffentlicht: (2026)
Can LLMs plan paths with extra hints from solvers?
von: Wu, Erik, et al.
Veröffentlicht: (2024)
von: Wu, Erik, et al.
Veröffentlicht: (2024)
Data Cartography for Detecting Memorization Hotspots and Guiding Data Interventions in Generative Models
von: Patel, Laksh, et al.
Veröffentlicht: (2025)
von: Patel, Laksh, et al.
Veröffentlicht: (2025)
Distilling LLMs' Decomposition Abilities into Compact Language Models
von: Tarasov, Denis, et al.
Veröffentlicht: (2024)
von: Tarasov, Denis, et al.
Veröffentlicht: (2024)
BatchTopK Sparse Autoencoders
von: Bussmann, Bart, et al.
Veröffentlicht: (2024)
von: Bussmann, Bart, et al.
Veröffentlicht: (2024)
Steering Evaluation-Aware Language Models to Act Like They Are Deployed
von: Hua, Tim Tian, et al.
Veröffentlicht: (2025)
von: Hua, Tim Tian, et al.
Veröffentlicht: (2025)
Gemma Scope: Open Sparse Autoencoders Everywhere All At Once on Gemma 2
von: Lieberum, Tom, et al.
Veröffentlicht: (2024)
von: Lieberum, Tom, et al.
Veröffentlicht: (2024)
KeyKnowledgeRAG (K^2RAG): An Enhanced RAG method for improved LLM question-answering capabilities
von: Markondapatnaikuni, Hruday, et al.
Veröffentlicht: (2025)
von: Markondapatnaikuni, Hruday, et al.
Veröffentlicht: (2025)
TIC: Translate-Infer-Compile for accurate "text to plan" using LLMs and Logical Representations
von: Agarwal, Sudhir, et al.
Veröffentlicht: (2024)
von: Agarwal, Sudhir, et al.
Veröffentlicht: (2024)
Forget What You Know about LLMs Evaluations -- LLMs are Like a Chameleon
von: Cohen-Inger, Nurit, et al.
Veröffentlicht: (2025)
von: Cohen-Inger, Nurit, et al.
Veröffentlicht: (2025)
A Unified Framework with Novel Metrics for Evaluating the Effectiveness of XAI Techniques in LLMs
von: Mersha, Melkamu Abay, et al.
Veröffentlicht: (2025)
von: Mersha, Melkamu Abay, et al.
Veröffentlicht: (2025)
Is Conformal Factuality for RAG-based LLMs Robust? Novel Metrics and Systematic Insights
von: Chen, Yi, et al.
Veröffentlicht: (2026)
von: Chen, Yi, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
Internal states before wait modulate reasoning patterns
von: Troitskii, Dmitrii, et al.
Veröffentlicht: (2025) -
Towards Best Practices of Activation Patching in Language Models: Metrics and Methods
von: Zhang, Fred, et al.
Veröffentlicht: (2023) -
Explorations of Self-Repair in Language Models
von: Rushing, Cody, et al.
Veröffentlicht: (2024) -
Agribot: agriculture-specific question answer system
von: Jain, Naman, et al.
Veröffentlicht: (2025) -
Censored LLMs as a Natural Testbed for Secret Knowledge Elicitation
von: Casademunt, Helena, et al.
Veröffentlicht: (2026)