Assessing LLM Reasoning Steps via Principal Knowledge Grounding
Fuente:
arXiv
Saved in:
| Main Authors: | Hwang, Hyeon, Cho, Yewon, Yoon, Chanwoong, Park, Yein, Song, Minju, Lee, Kyungjae, Kim, Gangwoo, Kang, Jaewoo |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
The Curious Case of Analogies: Investigating Analogical Reasoning in Large Language Models
by: Lee, Taewhoo, et al.
Published: (2025)
by: Lee, Taewhoo, et al.
Published: (2025)
ChroKnowledge: Unveiling Chronological Knowledge of Language Models in Multiple Domains
by: Park, Yein, et al.
Published: (2024)
by: Park, Yein, et al.
Published: (2024)
Rationale-Guided Retrieval Augmented Generation for Medical Question Answering
by: Sohn, Jiwoong, et al.
Published: (2024)
by: Sohn, Jiwoong, et al.
Published: (2024)
Does Time Have Its Place? Temporal Heads: Where Language Models Recall Time-specific Information
by: Park, Yein, et al.
Published: (2025)
by: Park, Yein, et al.
Published: (2025)
Outlier-Safe Pre-Training for Robust 4-Bit Quantization of Large Language Models
by: Park, Jungwoo, et al.
Published: (2025)
by: Park, Jungwoo, et al.
Published: (2025)
Ask Optimal Questions: Aligning Large Language Models with Retriever's Preference in Conversation
by: Yoon, Chanwoong, et al.
Published: (2024)
by: Yoon, Chanwoong, et al.
Published: (2024)
CompAct: Compressing Retrieved Documents Actively for Question Answering
by: Yoon, Chanwoong, et al.
Published: (2024)
by: Yoon, Chanwoong, et al.
Published: (2024)
OLAPH: Improving Factuality in Biomedical Long-form Question Answering
by: Jeong, Minbyul, et al.
Published: (2024)
by: Jeong, Minbyul, et al.
Published: (2024)
ETHIC: Evaluating Large Language Models on Long-Context Tasks with High Information Coverage
by: Lee, Taewhoo, et al.
Published: (2024)
by: Lee, Taewhoo, et al.
Published: (2024)
Small Language Models Learn Enhanced Reasoning Skills from Medical Textbooks
by: Kim, Hyunjae, et al.
Published: (2024)
by: Kim, Hyunjae, et al.
Published: (2024)
KU-DMIS at EHRSQL 2024:Generating SQL query via question templatization in EHR
by: Kim, Hajung, et al.
Published: (2024)
by: Kim, Hajung, et al.
Published: (2024)
Learning to Explore and Select for Coverage-Conditioned Retrieval-Augmented Generation
by: Kim, Takyoung, et al.
Published: (2024)
by: Kim, Takyoung, et al.
Published: (2024)
Teaching Language Models to Think in Code
by: Hwang, Hyeon, et al.
Published: (2026)
by: Hwang, Hyeon, et al.
Published: (2026)
CoTox: Chain-of-Thought-Based Molecular Toxicity Reasoning and Prediction
by: Park, Jueon, et al.
Published: (2025)
by: Park, Jueon, et al.
Published: (2025)
LAPIS: Language Model-Augmented Police Investigation System
by: Kim, Heedou, et al.
Published: (2024)
by: Kim, Heedou, et al.
Published: (2024)
ToxReason: A Benchmark for Mechanistic Chemical Toxicity Reasoning via Adverse Outcome Pathway
by: Park, Jueon, et al.
Published: (2026)
by: Park, Jueon, et al.
Published: (2026)
Thinking Sparks!: Emergent Attention Heads in Reasoning Models During Post Training
by: Park, Yein, et al.
Published: (2025)
by: Park, Yein, et al.
Published: (2025)
Framing Matters: Addressing Framing Sensitivity in Decision-Making through Behaviorally-Grounded Value Alignment
by: Hwang, Seojin, et al.
Published: (2026)
by: Hwang, Seojin, et al.
Published: (2026)
ASGuard: Activation-Scaling Guard to Mitigate Targeted Jailbreaking Attack
by: Park, Yein, et al.
Published: (2025)
by: Park, Yein, et al.
Published: (2025)
Reinforcement Learning from Reflective Feedback (RLRF): Aligning and Improving LLMs via Fine-Grained Self-Reflection
by: Lee, Kyungjae, et al.
Published: (2024)
by: Lee, Kyungjae, et al.
Published: (2024)
Revisiting the UID Hypothesis in LLM Reasoning Traces
by: Gwak, Minju, et al.
Published: (2025)
by: Gwak, Minju, et al.
Published: (2025)
Revisiting the Uniform Information Density Hypothesis in LLM Reasoning
by: Gwak, Minju, et al.
Published: (2025)
by: Gwak, Minju, et al.
Published: (2025)
Pearl: A Review-driven Persona-Knowledge Grounded Conversational Recommendation Dataset
by: Kim, Minjin, et al.
Published: (2024)
by: Kim, Minjin, et al.
Published: (2024)
CONDESION-BENCH: Conditional Decision-Making of Large Language Models in Compositional Action Space
by: Hwang, Yeonjun, et al.
Published: (2026)
by: Hwang, Yeonjun, et al.
Published: (2026)
One Missing Piece for Open-Source Reasoning Models: A Dataset to Mitigate Cold-Starting Short CoT LLMs in RL
by: Chae, Hyungjoo, et al.
Published: (2025)
by: Chae, Hyungjoo, et al.
Published: (2025)
RAISE: Enhancing Scientific Reasoning in LLMs via Step-by-Step Retrieval
by: Oh, Minhae, et al.
Published: (2025)
by: Oh, Minhae, et al.
Published: (2025)
Grounding LLM Reasoning with Knowledge Graphs
by: Amayuelas, Alfonso, et al.
Published: (2025)
by: Amayuelas, Alfonso, et al.
Published: (2025)
MolDeTox: Evaluating Language Model's Stepwise Fragment Editing for Molecular Detoxification
by: Park, Jueon, et al.
Published: (2026)
by: Park, Jueon, et al.
Published: (2026)
Structure-Grounded Knowledge Retrieval via Code Dependencies for Multi-Step Data Reasoning
by: Huang, Xinyi
Published: (2026)
by: Huang, Xinyi
Published: (2026)
Multi-View Attention Multiple-Instance Learning Enhanced by LLM Reasoning for Cognitive Distortion Detection
by: Kim, Jun Seo, et al.
Published: (2025)
by: Kim, Jun Seo, et al.
Published: (2025)
Knowledge Integration Decay in Search-Augmented Reasoning of Large Language Models
by: Yu, Sangwon, et al.
Published: (2026)
by: Yu, Sangwon, et al.
Published: (2026)
ReGUIDE: Data Efficient GUI Grounding via Spatial Reasoning and Search
by: Lee, Hyunseok, et al.
Published: (2025)
by: Lee, Hyunseok, et al.
Published: (2025)
LLM-as-a-Judge & Reward Model: What They Can and Cannot Do
by: Son, Guijin, et al.
Published: (2024)
by: Son, Guijin, et al.
Published: (2024)
Improving Visual Token Reduction via Rectifying Distortions for Efficient Multimodal LLM Inference
by: Cho, Hyeonwoo, et al.
Published: (2026)
by: Cho, Hyeonwoo, et al.
Published: (2026)
TWICE: What Advantages Can Low-Resource Domain-Specific Embedding Model Bring? -- A Case Study on Korea Financial Texts
by: Hwang, Yewon, et al.
Published: (2025)
by: Hwang, Yewon, et al.
Published: (2025)
Leveraging Language Models and RAG for Efficient Knowledge Discovery in Clinical Environments
by: Ko, Seokhwan, et al.
Published: (2025)
by: Ko, Seokhwan, et al.
Published: (2025)
Zero2Text: Zero-Training Cross-Domain Inversion Attacks on Textual Embeddings
by: Kim, Doohyun, et al.
Published: (2026)
by: Kim, Doohyun, et al.
Published: (2026)
Learning from Negative Samples in Biomedical Generative Entity Linking
by: Kim, Chanhwi, et al.
Published: (2024)
by: Kim, Chanhwi, et al.
Published: (2024)
Latent Paraphrasing: Perturbation on Layers Improves Knowledge Injection in Language Models
by: Kang, Minki, et al.
Published: (2024)
by: Kang, Minki, et al.
Published: (2024)
K-BrowseComp: A Web Browsing Agent Benchmark Grounded in Korean Contexts
by: Lee, Nahyun, et al.
Published: (2026)
by: Lee, Nahyun, et al.
Published: (2026)
Similar Items
-
The Curious Case of Analogies: Investigating Analogical Reasoning in Large Language Models
by: Lee, Taewhoo, et al.
Published: (2025) -
ChroKnowledge: Unveiling Chronological Knowledge of Language Models in Multiple Domains
by: Park, Yein, et al.
Published: (2024) -
Rationale-Guided Retrieval Augmented Generation for Medical Question Answering
by: Sohn, Jiwoong, et al.
Published: (2024) -
Does Time Have Its Place? Temporal Heads: Where Language Models Recall Time-specific Information
by: Park, Yein, et al.
Published: (2025) -
Outlier-Safe Pre-Training for Robust 4-Bit Quantization of Large Language Models
by: Park, Jungwoo, et al.
Published: (2025)