An Empirical Study of Conformal Prediction in LLM with ASP Scaffolds for Robust Reasoning
Fuente:
arXiv
Salvato in:
| Autori principali: | Kaur, Navdeep, McPheat, Lachlan, Russo, Alessandra, Cohn, Anthony G, Madhyastha, Pranava |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
DecompSR: A dataset for decomposed analyses of compositional multihop spatial reasoning
di: McPheat, Lachlan, et al.
Pubblicazione: (2025)
di: McPheat, Lachlan, et al.
Pubblicazione: (2025)
$\texttt{SEM-CTRL}$: Semantically Controlled Decoding
di: Albinhassan, Mohammad, et al.
Pubblicazione: (2025)
di: Albinhassan, Mohammad, et al.
Pubblicazione: (2025)
Working Memory Constraints Scaffold Learning in Transformers under Data Scarcity
di: Madhyastha, Pranava, et al.
Pubblicazione: (2026)
di: Madhyastha, Pranava, et al.
Pubblicazione: (2026)
A Cognitively Grounded Bayesian Framework for Misinformation Susceptibility
di: Madhyastha, Pranava
Pubblicazione: (2026)
di: Madhyastha, Pranava
Pubblicazione: (2026)
Learning and Enforcing Context-Sensitive Control for LLMs
di: Albinhassan, Mohammad, et al.
Pubblicazione: (2026)
di: Albinhassan, Mohammad, et al.
Pubblicazione: (2026)
Categorical Vector Space Semantics for Lambek Calculus with a Relevant Modality
di: McPheat, Lachlan, et al.
Pubblicazione: (2020)
di: McPheat, Lachlan, et al.
Pubblicazione: (2020)
LLM-Assisted Visual Analytics: Opportunities and Challenges
di: Hutchinson, Maeve, et al.
Pubblicazione: (2024)
di: Hutchinson, Maeve, et al.
Pubblicazione: (2024)
Steamroller Problems: An Evaluation of LLM Reasoning Capability with Automated Theorem Prover Strategies
di: McGinness, Lachlan, et al.
Pubblicazione: (2024)
di: McGinness, Lachlan, et al.
Pubblicazione: (2024)
Exploring Spatial Representations in the Historical Lake District Texts with LLM-based Relation Extraction
di: Haris, Erum, et al.
Pubblicazione: (2024)
di: Haris, Erum, et al.
Pubblicazione: (2024)
Highlighting Case Studies in LLM Literature Review of Interdisciplinary System Science
di: McGinness, Lachlan, et al.
Pubblicazione: (2025)
di: McGinness, Lachlan, et al.
Pubblicazione: (2025)
Reframing Spatial Reasoning Evaluation in Language Models: A Real-World Simulation Benchmark for Qualitative Reasoning
di: Li, Fangjun, et al.
Pubblicazione: (2024)
di: Li, Fangjun, et al.
Pubblicazione: (2024)
Validating Political Position Predictions of Arguments
di: Robinson, Jordan, et al.
Pubblicazione: (2026)
di: Robinson, Jordan, et al.
Pubblicazione: (2026)
An LLM + ASP Workflow for Joint Entity-Relation Extraction
di: Tran, Trang, et al.
Pubblicazione: (2025)
di: Tran, Trang, et al.
Pubblicazione: (2025)
Noise or Nuance: An Investigation Into Useful Information and Filtering For LLM Driven AKBC
di: Clay, Alex, et al.
Pubblicazione: (2025)
di: Clay, Alex, et al.
Pubblicazione: (2025)
Simple Augmentations of Logical Rules for Neuro-Symbolic Knowledge Graph Completion
di: Nandi, Ananjan, et al.
Pubblicazione: (2024)
di: Nandi, Ananjan, et al.
Pubblicazione: (2024)
Advancing Spatial Reasoning in Large Language Models: An In-Depth Evaluation and Enhancement Using the StepGame Benchmark
di: Li, Fangjun, et al.
Pubblicazione: (2024)
di: Li, Fangjun, et al.
Pubblicazione: (2024)
Large Reasoning Models Struggle to Transfer Parametric Knowledge Across Scripts
di: Bandarkar, Lucas, et al.
Pubblicazione: (2026)
di: Bandarkar, Lucas, et al.
Pubblicazione: (2026)
An Empirical Study on Reinforcement Learning for Reasoning-Search Interleaved LLM Agents
di: Jin, Bowen, et al.
Pubblicazione: (2025)
di: Jin, Bowen, et al.
Pubblicazione: (2025)
Safety Through Reasoning: An Empirical Study of Reasoning Guardrail Models
di: Sreedhar, Makesh Narsimhan, et al.
Pubblicazione: (2025)
di: Sreedhar, Makesh Narsimhan, et al.
Pubblicazione: (2025)
Quantization Hurts Reasoning? An Empirical Study on Quantized Reasoning Models
di: Liu, Ruikang, et al.
Pubblicazione: (2025)
di: Liu, Ruikang, et al.
Pubblicazione: (2025)
Chasing Progress, Not Perfection: Revisiting Strategies for End-to-End LLM Plan Generation
di: Huang, Sukai, et al.
Pubblicazione: (2024)
di: Huang, Sukai, et al.
Pubblicazione: (2024)
An Empirical Study of Group Conformity in Multi-Agent Systems
di: Choi, Min, et al.
Pubblicazione: (2025)
di: Choi, Min, et al.
Pubblicazione: (2025)
Dissecting Tool-Integrated Reasoning: An Empirical Study and Analysis
di: Zhao, Yufeng, et al.
Pubblicazione: (2025)
di: Zhao, Yufeng, et al.
Pubblicazione: (2025)
DynaSemble: Dynamic Ensembling of Textual and Structure-Based Models for Knowledge Graph Completion
di: Nandi, Ananjan, et al.
Pubblicazione: (2023)
di: Nandi, Ananjan, et al.
Pubblicazione: (2023)
SAGE: Steering Dialog Generation with Future-Aware State-Action Augmentation
di: Zhang, Yizhe, et al.
Pubblicazione: (2025)
di: Zhang, Yizhe, et al.
Pubblicazione: (2025)
Beyond Surface Statistics: Robust Conformal Prediction for LLMs via Internal Representations
di: Wang, Yanli, et al.
Pubblicazione: (2026)
di: Wang, Yanli, et al.
Pubblicazione: (2026)
Scaf-GRPO: Scaffolded Group Relative Policy Optimization for Enhancing LLM Reasoning
di: Zhang, Xichen, et al.
Pubblicazione: (2025)
di: Zhang, Xichen, et al.
Pubblicazione: (2025)
Explainable Chain-of-Thought Reasoning: An Empirical Analysis on State-Aware Reasoning Dynamics
di: Yu, Sheldon, et al.
Pubblicazione: (2025)
di: Yu, Sheldon, et al.
Pubblicazione: (2025)
Between Underthinking and Overthinking: An Empirical Study of Reasoning Length and correctness in LLMs
di: Su, Jinyan, et al.
Pubblicazione: (2025)
di: Su, Jinyan, et al.
Pubblicazione: (2025)
Graph-enhanced Large Language Models in Asynchronous Plan Reasoning
di: Lin, Fangru, et al.
Pubblicazione: (2024)
di: Lin, Fangru, et al.
Pubblicazione: (2024)
A Reliable Common-Sense Reasoning Socialbot Built Using LLMs and Goal-Directed ASP
di: Zeng, Yankai, et al.
Pubblicazione: (2024)
di: Zeng, Yankai, et al.
Pubblicazione: (2024)
The Thinking Spectrum: An Empirical Study of Tunable Reasoning in LLMs through Model Merging
di: Lan, Xiaochong, et al.
Pubblicazione: (2025)
di: Lan, Xiaochong, et al.
Pubblicazione: (2025)
Online Reasoning Calibration: Test-Time Training Enables Generalizable Conformal LLM Reasoning
di: Zhou, Cai, et al.
Pubblicazione: (2026)
di: Zhou, Cai, et al.
Pubblicazione: (2026)
Polysemantic Dropout: Conformal OOD Detection for Specialized LLMs
di: Gupta, Ayush, et al.
Pubblicazione: (2025)
di: Gupta, Ayush, et al.
Pubblicazione: (2025)
Diagnosing LLM Judge Reliability: Conformal Prediction Sets and Transitivity Violations
di: Gupta, Manan, et al.
Pubblicazione: (2026)
di: Gupta, Manan, et al.
Pubblicazione: (2026)
QSTRBench: a New Benchmark to Evaluate the Ability of Language Models to Reason with Qualitative Spatial and Temporal Calculi
di: Cohn, Anthony G., et al.
Pubblicazione: (2026)
di: Cohn, Anthony G., et al.
Pubblicazione: (2026)
What Affects the Stability of Tool Learning? An Empirical Study on the Robustness of Tool Learning Frameworks
di: Huang, Chengrui, et al.
Pubblicazione: (2024)
di: Huang, Chengrui, et al.
Pubblicazione: (2024)
An Empirical Study on Large Language Models in Accuracy and Robustness under Chinese Industrial Scenarios
di: Li, Zongjie, et al.
Pubblicazione: (2024)
di: Li, Zongjie, et al.
Pubblicazione: (2024)
Can LLM-Generated Textual Explanations Enhance Model Classification Performance? An Empirical Study
di: Dhaini, Mahdi, et al.
Pubblicazione: (2025)
di: Dhaini, Mahdi, et al.
Pubblicazione: (2025)
The Zero-Step Thinking: An Empirical Study of Mode Selection as Harder Early Exit in Reasoning Models
di: Tan, Yuqiao, et al.
Pubblicazione: (2025)
di: Tan, Yuqiao, et al.
Pubblicazione: (2025)
Documenti analoghi
-
DecompSR: A dataset for decomposed analyses of compositional multihop spatial reasoning
di: McPheat, Lachlan, et al.
Pubblicazione: (2025) -
$\texttt{SEM-CTRL}$: Semantically Controlled Decoding
di: Albinhassan, Mohammad, et al.
Pubblicazione: (2025) -
Working Memory Constraints Scaffold Learning in Transformers under Data Scarcity
di: Madhyastha, Pranava, et al.
Pubblicazione: (2026) -
A Cognitively Grounded Bayesian Framework for Misinformation Susceptibility
di: Madhyastha, Pranava
Pubblicazione: (2026) -
Learning and Enforcing Context-Sensitive Control for LLMs
di: Albinhassan, Mohammad, et al.
Pubblicazione: (2026)