CoRefine: Confidence-Guided Self-Refinement for Adaptive Test-Time Compute
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Jin, Chen, Tanno, Ryutaro, Diethe, Tom, Teare, Philip |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
An Image is Worth Multiple Words: Discovering Object Level Concepts using Multi-Concept Prompt Learning
von: Jin, Chen, et al.
Veröffentlicht: (2023)
von: Jin, Chen, et al.
Veröffentlicht: (2023)
Diffusion Instruction Tuning
von: Jin, Chen, et al.
Veröffentlicht: (2025)
von: Jin, Chen, et al.
Veröffentlicht: (2025)
DeCoRe: Decoding by Contrasting Retrieval Heads to Mitigate Hallucinations
von: Gema, Aryo Pradipta, et al.
Veröffentlicht: (2024)
von: Gema, Aryo Pradipta, et al.
Veröffentlicht: (2024)
Specification Self-Correction: Mitigating In-Context Reward Hacking Through Test-Time Refinement
von: Gallego, Víctor
Veröffentlicht: (2025)
von: Gallego, Víctor
Veröffentlicht: (2025)
Low-Confidence Gold: Refining Low-Confidence Samples for Efficient Instruction Tuning
von: Cai, Hongyi, et al.
Veröffentlicht: (2025)
von: Cai, Hongyi, et al.
Veröffentlicht: (2025)
AdaRefiner: Refining Decisions of Language Models with Adaptive Feedback
von: Zhang, Wanpeng, et al.
Veröffentlicht: (2023)
von: Zhang, Wanpeng, et al.
Veröffentlicht: (2023)
Confidence-guided Refinement Reasoning for Zero-shot Question Answering
von: Jang, Youwon, et al.
Veröffentlicht: (2025)
von: Jang, Youwon, et al.
Veröffentlicht: (2025)
DSPy Assertions: Computational Constraints for Self-Refining Language Model Pipelines
von: Singhvi, Arnav, et al.
Veröffentlicht: (2023)
von: Singhvi, Arnav, et al.
Veröffentlicht: (2023)
RefineCoder: Iterative Improving of Large Language Models via Adaptive Critique Refinement for Code Generation
von: Zhou, Changzhi, et al.
Veröffentlicht: (2025)
von: Zhou, Changzhi, et al.
Veröffentlicht: (2025)
Learning to Refine: Self-Refinement of Parallel Reasoning in LLMs
von: Wang, Qibin, et al.
Veröffentlicht: (2025)
von: Wang, Qibin, et al.
Veröffentlicht: (2025)
RefineX: Learning to Refine Pre-training Data at Scale from Expert-Guided Programs
von: Bi, Baolong, et al.
Veröffentlicht: (2025)
von: Bi, Baolong, et al.
Veröffentlicht: (2025)
Guided by Gut: Efficient Test-Time Scaling with Reinforced Intrinsic Confidence
von: Ghasemabadi, Amirhosein, et al.
Veröffentlicht: (2025)
von: Ghasemabadi, Amirhosein, et al.
Veröffentlicht: (2025)
BrowseConf: Confidence-Guided Test-Time Scaling for Web Agents
von: Ou, Litu, et al.
Veröffentlicht: (2025)
von: Ou, Litu, et al.
Veröffentlicht: (2025)
Eigen-1: Adaptive Multi-Agent Refinement with Monitor-Based RAG for Scientific Reasoning
von: Tang, Xiangru, et al.
Veröffentlicht: (2025)
von: Tang, Xiangru, et al.
Veröffentlicht: (2025)
TimeRefine: Temporal Grounding with Time Refining Video LLM
von: Wang, Xizi, et al.
Veröffentlicht: (2024)
von: Wang, Xizi, et al.
Veröffentlicht: (2024)
A Stitch in Time Saves Nine: Proactive Self-Refinement for Language Models
von: Han, Jinyi, et al.
Veröffentlicht: (2025)
von: Han, Jinyi, et al.
Veröffentlicht: (2025)
Refine Knowledge of Large Language Models via Adaptive Contrastive Learning
von: Li, Yinghui, et al.
Veröffentlicht: (2025)
von: Li, Yinghui, et al.
Veröffentlicht: (2025)
Spontaneous Reward Hacking in Iterative Self-Refinement
von: Pan, Jane, et al.
Veröffentlicht: (2024)
von: Pan, Jane, et al.
Veröffentlicht: (2024)
ProRefine: Inference-Time Prompt Refinement with Textual Feedback
von: Pandita, Deepak, et al.
Veröffentlicht: (2025)
von: Pandita, Deepak, et al.
Veröffentlicht: (2025)
Pride and Prejudice: LLM Amplifies Self-Bias in Self-Refinement
von: Xu, Wenda, et al.
Veröffentlicht: (2024)
von: Xu, Wenda, et al.
Veröffentlicht: (2024)
OSCAR: Orchestrated Self-verification and Cross-path Refinement
von: Shah, Yash, et al.
Veröffentlicht: (2026)
von: Shah, Yash, et al.
Veröffentlicht: (2026)
Self-Polish: Enhance Reasoning in Large Language Models via Problem Refinement
von: Xi, Zhiheng, et al.
Veröffentlicht: (2023)
von: Xi, Zhiheng, et al.
Veröffentlicht: (2023)
Bootstrapping Language-Guided Navigation Learning with Self-Refining Data Flywheel
von: Wang, Zun, et al.
Veröffentlicht: (2024)
von: Wang, Zun, et al.
Veröffentlicht: (2024)
Evolutionary Guided Decoding: Iterative Value Refinement for LLMs
von: Liu, Zhenhua, et al.
Veröffentlicht: (2025)
von: Liu, Zhenhua, et al.
Veröffentlicht: (2025)
LIMOPro: Reasoning Refinement for Efficient and Effective Test-time Scaling
von: Xiao, Yang, et al.
Veröffentlicht: (2025)
von: Xiao, Yang, et al.
Veröffentlicht: (2025)
Refine Thought: A Test-Time Inference Method for Embedding Model Reasoning
von: Wang, Guangzhi, et al.
Veröffentlicht: (2025)
von: Wang, Guangzhi, et al.
Veröffentlicht: (2025)
DeepRefine: Agent-Compiled Knowledge Refinement via Reinforcement Learning
von: Huang, Haoyu, et al.
Veröffentlicht: (2026)
von: Huang, Haoyu, et al.
Veröffentlicht: (2026)
Causal-Adapter: Taming Text-to-Image Diffusion for Faithful Counterfactual Generation
von: Tong, Lei, et al.
Veröffentlicht: (2025)
von: Tong, Lei, et al.
Veröffentlicht: (2025)
Confidence Improves Self-Consistency in LLMs
von: Taubenfeld, Amir, et al.
Veröffentlicht: (2025)
von: Taubenfeld, Amir, et al.
Veröffentlicht: (2025)
Adaptive Rectification Sampling for Test-Time Compute Scaling
von: Tan, Zhendong, et al.
Veröffentlicht: (2025)
von: Tan, Zhendong, et al.
Veröffentlicht: (2025)
Database Normalization via Dual-LLM Self-Refinement
von: Jo, Eunjae, et al.
Veröffentlicht: (2025)
von: Jo, Eunjae, et al.
Veröffentlicht: (2025)
Direct Alignment of Language Models via Quality-Aware Self-Refinement
von: Yu, Runsheng, et al.
Veröffentlicht: (2024)
von: Yu, Runsheng, et al.
Veröffentlicht: (2024)
Enhancing the Medical Context-Awareness Ability of LLMs via Multifaceted Self-Refinement Learning
von: Zhou, Yuxuan, et al.
Veröffentlicht: (2025)
von: Zhou, Yuxuan, et al.
Veröffentlicht: (2025)
Boosting Chart-to-Code Generation in MLLM via Dual Preference-Guided Refinement
von: Zhang, Zhihan, et al.
Veröffentlicht: (2025)
von: Zhang, Zhihan, et al.
Veröffentlicht: (2025)
AgentRefine: Enhancing Agent Generalization through Refinement Tuning
von: Fu, Dayuan, et al.
Veröffentlicht: (2025)
von: Fu, Dayuan, et al.
Veröffentlicht: (2025)
Mitigating Attention Localization in Small Scale: Self-Attention Refinement via One-step Belief Propagation
von: Lee, Nakyung, et al.
Veröffentlicht: (2025)
von: Lee, Nakyung, et al.
Veröffentlicht: (2025)
Adaptive Graph Refinement and Label Propagation with LLMs for Cost-Effective Entity Resolution
von: Wang, Hongtao, et al.
Veröffentlicht: (2026)
von: Wang, Hongtao, et al.
Veröffentlicht: (2026)
Search and Refine During Think: Facilitating Knowledge Refinement for Improved Retrieval-Augmented Reasoning
von: Shi, Yaorui, et al.
Veröffentlicht: (2025)
von: Shi, Yaorui, et al.
Veröffentlicht: (2025)
mrCAD: Multimodal Refinement of Computer-aided Designs
von: McCarthy, William P., et al.
Veröffentlicht: (2025)
von: McCarthy, William P., et al.
Veröffentlicht: (2025)
Adaptive Multi-Agent Response Refinement in Conversational Systems
von: Jeong, Soyeong, et al.
Veröffentlicht: (2025)
von: Jeong, Soyeong, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
An Image is Worth Multiple Words: Discovering Object Level Concepts using Multi-Concept Prompt Learning
von: Jin, Chen, et al.
Veröffentlicht: (2023) -
Diffusion Instruction Tuning
von: Jin, Chen, et al.
Veröffentlicht: (2025) -
DeCoRe: Decoding by Contrasting Retrieval Heads to Mitigate Hallucinations
von: Gema, Aryo Pradipta, et al.
Veröffentlicht: (2024) -
Specification Self-Correction: Mitigating In-Context Reward Hacking Through Test-Time Refinement
von: Gallego, Víctor
Veröffentlicht: (2025) -
Low-Confidence Gold: Refining Low-Confidence Samples for Efficient Instruction Tuning
von: Cai, Hongyi, et al.
Veröffentlicht: (2025)