ATLAS: Adaptive Test-Time Latent Steering with External Verifiers for Enhancing LLMs Reasoning
Fuente:
arXiv
Salvato in:
| Autori principali: | Nguyen, Tuc, Le, Thai |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2026
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Adapters Mixup: Mixing Parameter-Efficient Adapters to Enhance the Adversarial Robustness of Fine-tuned Pre-trained Text Classifiers
di: Nguyen, Tuc, et al.
Pubblicazione: (2024)
di: Nguyen, Tuc, et al.
Pubblicazione: (2024)
Generalizability of Mixture of Domain-Specific Adapters from the Lens of Signed Weight Directions and its Application to Effective Model Pruning
di: Nguyen, Tuc, et al.
Pubblicazione: (2024)
di: Nguyen, Tuc, et al.
Pubblicazione: (2024)
Unraveling Interwoven Roles of Large Language Models in Authorship Privacy: Obfuscation, Mimicking, and Verification
di: Nguyen, Tuc, et al.
Pubblicazione: (2025)
di: Nguyen, Tuc, et al.
Pubblicazione: (2025)
VerifierQ: Enhancing LLM Test Time Compute with Q-Learning-based Verifiers
di: Qi, Jianing, et al.
Pubblicazione: (2024)
di: Qi, Jianing, et al.
Pubblicazione: (2024)
Understanding and Steering the Cognitive Behaviors of Reasoning Models at Test-Time
di: Zhang, Zhenyu, et al.
Pubblicazione: (2025)
di: Zhang, Zhenyu, et al.
Pubblicazione: (2025)
NoMatterXAI: Generating "No Matter What" Alterfactual Examples for Explaining Black-Box Text Classification Models
di: Nguyen, Tuc, et al.
Pubblicazione: (2024)
di: Nguyen, Tuc, et al.
Pubblicazione: (2024)
Thinking in Latents: Adaptive Anchor Refinement for Implicit Reasoning in LLMs
di: Sheshanarayana, Disha, et al.
Pubblicazione: (2026)
di: Sheshanarayana, Disha, et al.
Pubblicazione: (2026)
Parallel Test-Time Scaling for Latent Reasoning Models
di: You, Runyang, et al.
Pubblicazione: (2025)
di: You, Runyang, et al.
Pubblicazione: (2025)
Dynamic Experts Search: Enhancing Reasoning in Mixture-of-Experts LLMs at Test Time
di: Han, Yixuan, et al.
Pubblicazione: (2025)
di: Han, Yixuan, et al.
Pubblicazione: (2025)
SteerConf: Steering LLMs for Confidence Elicitation
di: Zhou, Ziang, et al.
Pubblicazione: (2025)
di: Zhou, Ziang, et al.
Pubblicazione: (2025)
Focus On This, Not That! Steering LLMs with Adaptive Feature Specification
di: Lamb, Tom A., et al.
Pubblicazione: (2024)
di: Lamb, Tom A., et al.
Pubblicazione: (2024)
Scaling up Test-Time Compute with Latent Reasoning: A Recurrent Depth Approach
di: Geiping, Jonas, et al.
Pubblicazione: (2025)
di: Geiping, Jonas, et al.
Pubblicazione: (2025)
Fractional Reasoning via Latent Steering Vectors Improves Inference Time Compute
di: Liu, Sheng, et al.
Pubblicazione: (2025)
di: Liu, Sheng, et al.
Pubblicazione: (2025)
The Hidden Signal of Verifier Strictness: Controlling and Improving Step-Wise Verification via Selective Latent Steering
di: Zhou, Yefan, et al.
Pubblicazione: (2026)
di: Zhou, Yefan, et al.
Pubblicazione: (2026)
LaDiR: Latent Diffusion Enhances LLMs for Text Reasoning
di: Kang, Haoqiang, et al.
Pubblicazione: (2025)
di: Kang, Haoqiang, et al.
Pubblicazione: (2025)
Feature Extraction and Steering for Enhanced Chain-of-Thought Reasoning in Language Models
di: Li, Zihao, et al.
Pubblicazione: (2025)
di: Li, Zihao, et al.
Pubblicazione: (2025)
Adaptive Rollout Allocation for Online Reinforcement Learning with Verifiable Rewards
di: Nguyen, Hieu Trung, et al.
Pubblicazione: (2026)
di: Nguyen, Hieu Trung, et al.
Pubblicazione: (2026)
Adaptive Two-Phase Finetuning LLMs for Japanese Legal Text Retrieval
di: Trung, Quang Hoang, et al.
Pubblicazione: (2024)
di: Trung, Quang Hoang, et al.
Pubblicazione: (2024)
Examining Reasoning LLMs-as-Judges in Non-Verifiable LLM Post-Training
di: Liu, Yixin, et al.
Pubblicazione: (2026)
di: Liu, Yixin, et al.
Pubblicazione: (2026)
Steer LLM Latents for Hallucination Detection
di: Park, Seongheon, et al.
Pubblicazione: (2025)
di: Park, Seongheon, et al.
Pubblicazione: (2025)
Learning to Reason without External Rewards
di: Zhao, Xuandong, et al.
Pubblicazione: (2025)
di: Zhao, Xuandong, et al.
Pubblicazione: (2025)
Reinforcing General Reasoning without Verifiers
di: Zhou, Xiangxin, et al.
Pubblicazione: (2025)
di: Zhou, Xiangxin, et al.
Pubblicazione: (2025)
Steering the Verifiability of Multimodal AI Hallucinations
di: Pang, Jianhong, et al.
Pubblicazione: (2026)
di: Pang, Jianhong, et al.
Pubblicazione: (2026)
Inference-Time Rethinking with Latent Thought Vectors for Math Reasoning
di: Kong, Deqian, et al.
Pubblicazione: (2026)
di: Kong, Deqian, et al.
Pubblicazione: (2026)
Seek in the Dark: Reasoning via Test-Time Instance-Level Policy Gradient in Latent Space
di: Li, Hengli, et al.
Pubblicazione: (2025)
di: Li, Hengli, et al.
Pubblicazione: (2025)
Beyond Steering Vector: Flow-based Activation Steering for Inference-Time Intervention
di: Jin, Zehao, et al.
Pubblicazione: (2026)
di: Jin, Zehao, et al.
Pubblicazione: (2026)
Adaptive Test-Time Reasoning via Reward-Guided Dual-Phase Search
di: Cui, Yingqian, et al.
Pubblicazione: (2025)
di: Cui, Yingqian, et al.
Pubblicazione: (2025)
AAPO: Enhancing the Reasoning Capabilities of LLMs with Advantage Margin
di: Xiong, Jian, et al.
Pubblicazione: (2025)
di: Xiong, Jian, et al.
Pubblicazione: (2025)
TDD-Bench Verified: Can LLMs Generate Tests for Issues Before They Get Resolved?
di: Ahmed, Toufique, et al.
Pubblicazione: (2024)
di: Ahmed, Toufique, et al.
Pubblicazione: (2024)
Guiding Giants: Lightweight Controllers for Weighted Activation Steering in LLMs
di: Hegazy, Amr, et al.
Pubblicazione: (2025)
di: Hegazy, Amr, et al.
Pubblicazione: (2025)
Steering MoE LLMs via Expert (De)Activation
di: Fayyaz, Mohsen, et al.
Pubblicazione: (2025)
di: Fayyaz, Mohsen, et al.
Pubblicazione: (2025)
ShareChat: A Dataset of Chatbot Conversations in the Wild
di: Yan, Yueru, et al.
Pubblicazione: (2025)
di: Yan, Yueru, et al.
Pubblicazione: (2025)
Silence the Judge: Reinforcement Learning with Self-Verifier via Latent Geometric Clustering
di: Zhang, Nonghai, et al.
Pubblicazione: (2026)
di: Zhang, Nonghai, et al.
Pubblicazione: (2026)
Enhancing LLM Reasoning via Critique Models with Test-Time and Training-Time Supervision
di: Xi, Zhiheng, et al.
Pubblicazione: (2024)
di: Xi, Zhiheng, et al.
Pubblicazione: (2024)
MM-Verify: Enhancing Multimodal Reasoning with Chain-of-Thought Verification
di: Sun, Linzhuang, et al.
Pubblicazione: (2025)
di: Sun, Linzhuang, et al.
Pubblicazione: (2025)
Dynamic Cheatsheet: Test-Time Learning with Adaptive Memory
di: Suzgun, Mirac, et al.
Pubblicazione: (2025)
di: Suzgun, Mirac, et al.
Pubblicazione: (2025)
Adaptive Prompting for Continual Relation Extraction: A Within-Task Variance Perspective
di: Le, Minh, et al.
Pubblicazione: (2024)
di: Le, Minh, et al.
Pubblicazione: (2024)
Understanding How CodeLLMs (Mis)Predict Types with Activation Steering
di: Lucchetti, Francesca, et al.
Pubblicazione: (2024)
di: Lucchetti, Francesca, et al.
Pubblicazione: (2024)
Steering LLMs for Formal Theorem Proving
di: Kirtania, Shashank, et al.
Pubblicazione: (2025)
di: Kirtania, Shashank, et al.
Pubblicazione: (2025)
Latent-GRPO: Group Relative Policy Optimization for Latent Reasoning
di: Deng, Jingcheng, et al.
Pubblicazione: (2026)
di: Deng, Jingcheng, et al.
Pubblicazione: (2026)
Documenti analoghi
-
Adapters Mixup: Mixing Parameter-Efficient Adapters to Enhance the Adversarial Robustness of Fine-tuned Pre-trained Text Classifiers
di: Nguyen, Tuc, et al.
Pubblicazione: (2024) -
Generalizability of Mixture of Domain-Specific Adapters from the Lens of Signed Weight Directions and its Application to Effective Model Pruning
di: Nguyen, Tuc, et al.
Pubblicazione: (2024) -
Unraveling Interwoven Roles of Large Language Models in Authorship Privacy: Obfuscation, Mimicking, and Verification
di: Nguyen, Tuc, et al.
Pubblicazione: (2025) -
VerifierQ: Enhancing LLM Test Time Compute with Q-Learning-based Verifiers
di: Qi, Jianing, et al.
Pubblicazione: (2024) -
Understanding and Steering the Cognitive Behaviors of Reasoning Models at Test-Time
di: Zhang, Zhenyu, et al.
Pubblicazione: (2025)