Beyond Correctness: Learning Robust Reasoning via Transfer
Fuente:
arXiv
Guardado en:
| Autores principales: | Lee, Hyunseok, Abbasloo, Soheil, Tack, Jihoon, Shin, Jinwoo |
|---|---|
| Formato: | Preprint |
| Publicado: |
2026
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
ReMoDetect: Reward Models Recognize Aligned LLM's Generations
por: Lee, Hyunseok, et al.
Publicado: (2024)
por: Lee, Hyunseok, et al.
Publicado: (2024)
ReVISE: Learning to Refine at Test-Time via Intrinsic Self-Verification
por: Lee, Hyunseok, et al.
Publicado: (2025)
por: Lee, Hyunseok, et al.
Publicado: (2025)
ReGUIDE: Data Efficient GUI Grounding via Spatial Reasoning and Search
por: Lee, Hyunseok, et al.
Publicado: (2025)
por: Lee, Hyunseok, et al.
Publicado: (2025)
Tabular Transfer Learning via Prompting LLMs
por: Nam, Jaehyun, et al.
Publicado: (2024)
por: Nam, Jaehyun, et al.
Publicado: (2024)
Online Adaptation of Language Models with a Memory of Amortized Contexts
por: Tack, Jihoon, et al.
Publicado: (2024)
por: Tack, Jihoon, et al.
Publicado: (2024)
Think Clearly: Improving Reasoning via Redundant Token Pruning
por: Choi, Daewon, et al.
Publicado: (2025)
por: Choi, Daewon, et al.
Publicado: (2025)
Are Language Models Up to Sequential Optimization Problems? From Evaluation to a Hegelian-Inspired Enhancement
por: Abbasloo, Soheil
Publicado: (2025)
por: Abbasloo, Soheil
Publicado: (2025)
Evaluating LLM Reasoning Beyond Correctness and CoT
por: Abbasloo, Soheil
Publicado: (2025)
por: Abbasloo, Soheil
Publicado: (2025)
Optimized Feature Generation for Tabular Data via LLMs with Decision Tree Reasoning
por: Nam, Jaehyun, et al.
Publicado: (2024)
por: Nam, Jaehyun, et al.
Publicado: (2024)
Self-Refining Language Model Anonymizers via Adversarial Distillation
por: Kim, Kyuyoung, et al.
Publicado: (2025)
por: Kim, Kyuyoung, et al.
Publicado: (2025)
Learning to Contextualize Web Pages for Enhanced Decision Making by LLM Agents
por: Lee, Dongjun, et al.
Publicado: (2025)
por: Lee, Dongjun, et al.
Publicado: (2025)
Sparsified State-Space Models are Efficient Highway Networks
por: Song, Woomin, et al.
Publicado: (2025)
por: Song, Woomin, et al.
Publicado: (2025)
Chain-of-Defensive-Thought: Structured Reasoning Elicits Robustness in Large Language Models against Reference Corruption
por: Wang, Wenxiao, et al.
Publicado: (2025)
por: Wang, Wenxiao, et al.
Publicado: (2025)
Spread Preference Annotation: Direct Preference Judgment for Efficient LLM Alignment
por: Kim, Dongyoung, et al.
Publicado: (2024)
por: Kim, Dongyoung, et al.
Publicado: (2024)
Critique-Guided Distillation for Robust Reasoning via Refinement
por: Kapusuzoglu, Berkcan, et al.
Publicado: (2025)
por: Kapusuzoglu, Berkcan, et al.
Publicado: (2025)
Off-the-Shelf LLMs as Process Scorers: Training-Free Alternative to PRMs for Mathematical Reasoning
por: Chegini, Atoosa, et al.
Publicado: (2026)
por: Chegini, Atoosa, et al.
Publicado: (2026)
LLM Pretraining with Continuous Concepts
por: Tack, Jihoon, et al.
Publicado: (2025)
por: Tack, Jihoon, et al.
Publicado: (2025)
Early Stopping for Large Reasoning Models via Confidence Dynamics
por: Hosseini, Parsa, et al.
Publicado: (2026)
por: Hosseini, Parsa, et al.
Publicado: (2026)
Model Medicine: A Clinical Framework for Understanding, Diagnosing, and Treating AI Models
por: Jeong, Jihoon
Publicado: (2026)
por: Jeong, Jihoon
Publicado: (2026)
M-CARE: Standardized Clinical Case Reporting for AI Model Behavioral Disorders, with a 20-Case Atlas and Experimental Validation
por: Jeong, Jihoon
Publicado: (2026)
por: Jeong, Jihoon
Publicado: (2026)
Papez: Resource-Efficient Speech Separation with Auditory Working Memory
por: Oh, Hyunseok, et al.
Publicado: (2024)
por: Oh, Hyunseok, et al.
Publicado: (2024)
Beyond Preferences: Learning Alignment Principles Grounded in Human Reasons and Values
por: Bell, Henry, et al.
Publicado: (2026)
por: Bell, Henry, et al.
Publicado: (2026)
Schoenfeld's Anatomy of Mathematical Reasoning by Language Models
por: Li, Ming, et al.
Publicado: (2025)
por: Li, Ming, et al.
Publicado: (2025)
Sign Gradient Descent-based Neuronal Dynamics: ANN-to-SNN Conversion Beyond ReLU Network
por: Oh, Hyunseok, et al.
Publicado: (2024)
por: Oh, Hyunseok, et al.
Publicado: (2024)
Learning and Transferring Sparse Contextual Bigrams with Linear Transformers
por: Ren, Yunwei, et al.
Publicado: (2024)
por: Ren, Yunwei, et al.
Publicado: (2024)
Transferable Post-training via Inverse Value Learning
por: Lu, Xinyu, et al.
Publicado: (2024)
por: Lu, Xinyu, et al.
Publicado: (2024)
Beyond Answers: Transferring Reasoning Capabilities to Smaller LLMs Using Multi-Teacher Knowledge Distillation
por: Tian, Yijun, et al.
Publicado: (2024)
por: Tian, Yijun, et al.
Publicado: (2024)
Reasoning Beyond Literal: Cross-style Multimodal Reasoning for Figurative Language Understanding
por: Cheshmi, Seyyed Saeid, et al.
Publicado: (2026)
por: Cheshmi, Seyyed Saeid, et al.
Publicado: (2026)
Learning to Correct for QA Reasoning with Black-box LLMs
por: Kim, Jaehyung, et al.
Publicado: (2024)
por: Kim, Jaehyung, et al.
Publicado: (2024)
Reasoning Beyond Limits: Advances and Open Problems for LLMs
por: Ferrag, Mohamed Amine, et al.
Publicado: (2025)
por: Ferrag, Mohamed Amine, et al.
Publicado: (2025)
Beyond Autoregression: Discrete Diffusion for Complex Reasoning and Planning
por: Ye, Jiacheng, et al.
Publicado: (2024)
por: Ye, Jiacheng, et al.
Publicado: (2024)
Clarify: Improving Model Robustness With Natural Language Corrections
por: Lee, Yoonho, et al.
Publicado: (2024)
por: Lee, Yoonho, et al.
Publicado: (2024)
The Reasoning Trap: An Information-Theoretic Bound on Closed-System Multi-Step LLM Reasoning
por: Shin, Kwan Soo
Publicado: (2026)
por: Shin, Kwan Soo
Publicado: (2026)
Learning to Reason in LLMs by Expectation Maximization
por: Lee, Junghyun, et al.
Publicado: (2025)
por: Lee, Junghyun, et al.
Publicado: (2025)
Think Beyond Size: Adaptive Prompting for More Effective Reasoning
por: R, Kamesh
Publicado: (2024)
por: R, Kamesh
Publicado: (2024)
OncoReason: Structuring Clinical Reasoning in LLMs for Robust and Interpretable Survival Prediction
por: Hemadri, Raghu Vamshi, et al.
Publicado: (2025)
por: Hemadri, Raghu Vamshi, et al.
Publicado: (2025)
On the Robustness of Answer Formats in Medical Reasoning Models
por: Taveekitworachai, Pittawat, et al.
Publicado: (2025)
por: Taveekitworachai, Pittawat, et al.
Publicado: (2025)
FTFT: Efficient and Robust Fine-Tuning by Transferring Training Dynamics
por: Du, Yupei, et al.
Publicado: (2023)
por: Du, Yupei, et al.
Publicado: (2023)
Can A Gamer Train A Mathematical Reasoning Model?
por: Shin, Andrew
Publicado: (2025)
por: Shin, Andrew
Publicado: (2025)
Evaluating Robustness of Reward Models for Mathematical Reasoning
por: Kim, Sunghwan, et al.
Publicado: (2024)
por: Kim, Sunghwan, et al.
Publicado: (2024)
Ejemplares similares
-
ReMoDetect: Reward Models Recognize Aligned LLM's Generations
por: Lee, Hyunseok, et al.
Publicado: (2024) -
ReVISE: Learning to Refine at Test-Time via Intrinsic Self-Verification
por: Lee, Hyunseok, et al.
Publicado: (2025) -
ReGUIDE: Data Efficient GUI Grounding via Spatial Reasoning and Search
por: Lee, Hyunseok, et al.
Publicado: (2025) -
Tabular Transfer Learning via Prompting LLMs
por: Nam, Jaehyun, et al.
Publicado: (2024) -
Online Adaptation of Language Models with a Memory of Amortized Contexts
por: Tack, Jihoon, et al.
Publicado: (2024)