Argument Reconstruction as Supervision for Critical Thinking in LLMs
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Ryu, Hyun, Chu, Gyouk, Betz, Gregor, Yang, Eunho, Rose, Carolyn, Welleck, Sean |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Every Expert Matters: Towards Effective Knowledge Distillation for Mixture-of-Experts Language Models
von: Kim, Gyeongman, et al.
Veröffentlicht: (2025)
von: Kim, Gyeongman, et al.
Veröffentlicht: (2025)
L1: Controlling How Long A Reasoning Model Thinks With Reinforcement Learning
von: Aggarwal, Pranjal, et al.
Veröffentlicht: (2025)
von: Aggarwal, Pranjal, et al.
Veröffentlicht: (2025)
Divide and Translate: Compositional First-Order Logic Translation and Verification for Complex Logical Reasoning
von: Ryu, Hyun, et al.
Veröffentlicht: (2024)
von: Ryu, Hyun, et al.
Veröffentlicht: (2024)
OptimalThinkingBench: Evaluating Over and Underthinking in LLMs
von: Aggarwal, Pranjal, et al.
Veröffentlicht: (2025)
von: Aggarwal, Pranjal, et al.
Veröffentlicht: (2025)
ReviewScore: Misinformed Peer Review Detection with Large Language Models
von: Ryu, Hyun, et al.
Veröffentlicht: (2025)
von: Ryu, Hyun, et al.
Veröffentlicht: (2025)
Optimizing Temperature for Language Models with Multi-Sample Inference
von: Du, Weihua, et al.
Veröffentlicht: (2025)
von: Du, Weihua, et al.
Veröffentlicht: (2025)
On Code-Induced Reasoning in LLMs
von: Waheed, Abdul, et al.
Veröffentlicht: (2025)
von: Waheed, Abdul, et al.
Veröffentlicht: (2025)
Easy-to-Hard Generalization: Scalable Alignment Beyond Human Supervision
von: Sun, Zhiqing, et al.
Veröffentlicht: (2024)
von: Sun, Zhiqing, et al.
Veröffentlicht: (2024)
miniCTX: Neural Theorem Proving with (Long-)Contexts
von: Hu, Jiewen, et al.
Veröffentlicht: (2024)
von: Hu, Jiewen, et al.
Veröffentlicht: (2024)
Agentic-R1: Distilled Dual-Strategy Reasoning
von: Du, Weihua, et al.
Veröffentlicht: (2025)
von: Du, Weihua, et al.
Veröffentlicht: (2025)
GradAlign: Gradient-Aligned Data Selection for LLM Reinforcement Learning
von: Yang, Ningyuan, et al.
Veröffentlicht: (2026)
von: Yang, Ningyuan, et al.
Veröffentlicht: (2026)
Token-Supervised Value Models for Enhancing Mathematical Problem-Solving Capabilities of Large Language Models
von: Lee, Jung Hyun, et al.
Veröffentlicht: (2024)
von: Lee, Jung Hyun, et al.
Veröffentlicht: (2024)
Think Together and Work Better: Combining Humans' and LLMs' Think-Aloud Outcomes for Effective Text Evaluation
von: Chu, SeongYeub, et al.
Veröffentlicht: (2024)
von: Chu, SeongYeub, et al.
Veröffentlicht: (2024)
Language Models as Critical Thinking Tools: A Case Study of Philosophers
von: Ye, Andre, et al.
Veröffentlicht: (2024)
von: Ye, Andre, et al.
Veröffentlicht: (2024)
ImProver 2: Iteratively Self-Improving LMs for Neurosymbolic Proof Optimization
von: Ahuja, Riyaz, et al.
Veröffentlicht: (2026)
von: Ahuja, Riyaz, et al.
Veröffentlicht: (2026)
Double-Checker: Enhancing Reasoning of Slow-Thinking LLMs via Self-Critical Fine-Tuning
von: Xu, Xin, et al.
Veröffentlicht: (2025)
von: Xu, Xin, et al.
Veröffentlicht: (2025)
What are They Thinking? Delineation, Probing and Tracking of Concepts in LLMs
von: Abdelwahab, Mohamed, et al.
Veröffentlicht: (2026)
von: Abdelwahab, Mohamed, et al.
Veröffentlicht: (2026)
Argument-Based Consistency in Toxicity Explanations of LLMs
von: Mothilal, Ramaravind Kommiya, et al.
Veröffentlicht: (2025)
von: Mothilal, Ramaravind Kommiya, et al.
Veröffentlicht: (2025)
Can LLMs Extract Frame-Semantic Arguments?
von: Devasier, Jacob, et al.
Veröffentlicht: (2025)
von: Devasier, Jacob, et al.
Veröffentlicht: (2025)
LLMs for Argument Mining: Detection, Extraction, and Relationship Classification of pre-defined Arguments in Online Comments
von: Guida, Matteo, et al.
Veröffentlicht: (2025)
von: Guida, Matteo, et al.
Veröffentlicht: (2025)
The CoT Encyclopedia: Analyzing, Predicting, and Controlling how a Reasoning Model will Think
von: Lee, Seongyun, et al.
Veröffentlicht: (2025)
von: Lee, Seongyun, et al.
Veröffentlicht: (2025)
Asking and Answering Questions to Extract Event-Argument Structures
von: Uddin, Md Nayem, et al.
Veröffentlicht: (2024)
von: Uddin, Md Nayem, et al.
Veröffentlicht: (2024)
Leveraging Small LLMs for Argument Mining in Education: Argument Component Identification, Classification, and Assessment
von: Favero, Lucile, et al.
Veröffentlicht: (2025)
von: Favero, Lucile, et al.
Veröffentlicht: (2025)
Explorable Theorems: Making Written Theorems Explorable by Grounding Them in Formal Representations
von: Kambhamettu, Hita, et al.
Veröffentlicht: (2026)
von: Kambhamettu, Hita, et al.
Veröffentlicht: (2026)
Generating Uncontextualized and Contextualized Questions for Document-Level Event Argument Extraction
von: Uddin, Md Nayem, et al.
Veröffentlicht: (2024)
von: Uddin, Md Nayem, et al.
Veröffentlicht: (2024)
ProCoT: Stimulating Critical Thinking and Writing of Students through Engagement with Large Language Models (LLMs)
von: Adewumi, Tosin, et al.
Veröffentlicht: (2023)
von: Adewumi, Tosin, et al.
Veröffentlicht: (2023)
KorMedMCQA: Multi-Choice Question Answering Benchmark for Korean Healthcare Professional Licensing Examinations
von: Kweon, Sunjun, et al.
Veröffentlicht: (2024)
von: Kweon, Sunjun, et al.
Veröffentlicht: (2024)
ThinkPatterns-21k: A Systematic Study on the Impact of Thinking Patterns in LLMs
von: Wen, Pengcheng, et al.
Veröffentlicht: (2025)
von: Wen, Pengcheng, et al.
Veröffentlicht: (2025)
Do LLMs Align Human Values Regarding Social Biases? Judging and Explaining Social Biases with LLMs
von: Liu, Yang, et al.
Veröffentlicht: (2025)
von: Liu, Yang, et al.
Veröffentlicht: (2025)
CURaTE: Continual Unlearning in Real Time with Ensured Preservation of LLM Knowledge
von: Bae, Seyun, et al.
Veröffentlicht: (2026)
von: Bae, Seyun, et al.
Veröffentlicht: (2026)
Critical-Questions-of-Thought: Steering LLM reasoning with Argumentative Querying
von: Castagna, Federico, et al.
Veröffentlicht: (2024)
von: Castagna, Federico, et al.
Veröffentlicht: (2024)
ArgBench: Benchmarking LLMs on Computational Argumentation Tasks
von: Ajjour, Yamen, et al.
Veröffentlicht: (2026)
von: Ajjour, Yamen, et al.
Veröffentlicht: (2026)
Argumentation for Explainable and Globally Contestable Decision Support with LLMs
von: Dejl, Adam, et al.
Veröffentlicht: (2026)
von: Dejl, Adam, et al.
Veröffentlicht: (2026)
Format Inertia: A Failure Mechanism of LLMs in Medical Pre-Consultation
von: Lim, Seungseop, et al.
Veröffentlicht: (2025)
von: Lim, Seungseop, et al.
Veröffentlicht: (2025)
Cognitive Agent Compilation for Explicit Problem Solver Modeling
von: Moon, Hyeongdon, et al.
Veröffentlicht: (2026)
von: Moon, Hyeongdon, et al.
Veröffentlicht: (2026)
Demystifying Hybrid Thinking: Can LLMs Truly Switch Between Think and No-Think?
von: Wang, Shouren, et al.
Veröffentlicht: (2025)
von: Wang, Shouren, et al.
Veröffentlicht: (2025)
Enhancing Rating Prediction with Off-the-Shelf LLMs Using In-Context User Reviews
von: Ryu, Koki, et al.
Veröffentlicht: (2025)
von: Ryu, Koki, et al.
Veröffentlicht: (2025)
Latent Reasoning with Supervised Thinking States
von: Amos, Ido, et al.
Veröffentlicht: (2026)
von: Amos, Ido, et al.
Veröffentlicht: (2026)
Teaching LLMs Human-Like Editing of Inappropriate Argumentation via Reinforcement Learning
von: Ziegenbein, Timon, et al.
Veröffentlicht: (2026)
von: Ziegenbein, Timon, et al.
Veröffentlicht: (2026)
Fearful Falcons and Angry Llamas: Emotion Category Annotations of Arguments by Humans and LLMs
von: Greschner, Lynn, et al.
Veröffentlicht: (2024)
von: Greschner, Lynn, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Every Expert Matters: Towards Effective Knowledge Distillation for Mixture-of-Experts Language Models
von: Kim, Gyeongman, et al.
Veröffentlicht: (2025) -
L1: Controlling How Long A Reasoning Model Thinks With Reinforcement Learning
von: Aggarwal, Pranjal, et al.
Veröffentlicht: (2025) -
Divide and Translate: Compositional First-Order Logic Translation and Verification for Complex Logical Reasoning
von: Ryu, Hyun, et al.
Veröffentlicht: (2024) -
OptimalThinkingBench: Evaluating Over and Underthinking in LLMs
von: Aggarwal, Pranjal, et al.
Veröffentlicht: (2025) -
ReviewScore: Misinformed Peer Review Detection with Large Language Models
von: Ryu, Hyun, et al.
Veröffentlicht: (2025)