ATLAS: Adaptive Test-Time Latent Steering with External Verifiers for Enhancing LLMs Reasoning
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Nguyen, Tuc, Le, Thai |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Adapters Mixup: Mixing Parameter-Efficient Adapters to Enhance the Adversarial Robustness of Fine-tuned Pre-trained Text Classifiers
von: Nguyen, Tuc, et al.
Veröffentlicht: (2024)
von: Nguyen, Tuc, et al.
Veröffentlicht: (2024)
Generalizability of Mixture of Domain-Specific Adapters from the Lens of Signed Weight Directions and its Application to Effective Model Pruning
von: Nguyen, Tuc, et al.
Veröffentlicht: (2024)
von: Nguyen, Tuc, et al.
Veröffentlicht: (2024)
Unraveling Interwoven Roles of Large Language Models in Authorship Privacy: Obfuscation, Mimicking, and Verification
von: Nguyen, Tuc, et al.
Veröffentlicht: (2025)
von: Nguyen, Tuc, et al.
Veröffentlicht: (2025)
VerifierQ: Enhancing LLM Test Time Compute with Q-Learning-based Verifiers
von: Qi, Jianing, et al.
Veröffentlicht: (2024)
von: Qi, Jianing, et al.
Veröffentlicht: (2024)
Understanding and Steering the Cognitive Behaviors of Reasoning Models at Test-Time
von: Zhang, Zhenyu, et al.
Veröffentlicht: (2025)
von: Zhang, Zhenyu, et al.
Veröffentlicht: (2025)
NoMatterXAI: Generating "No Matter What" Alterfactual Examples for Explaining Black-Box Text Classification Models
von: Nguyen, Tuc, et al.
Veröffentlicht: (2024)
von: Nguyen, Tuc, et al.
Veröffentlicht: (2024)
Thinking in Latents: Adaptive Anchor Refinement for Implicit Reasoning in LLMs
von: Sheshanarayana, Disha, et al.
Veröffentlicht: (2026)
von: Sheshanarayana, Disha, et al.
Veröffentlicht: (2026)
Parallel Test-Time Scaling for Latent Reasoning Models
von: You, Runyang, et al.
Veröffentlicht: (2025)
von: You, Runyang, et al.
Veröffentlicht: (2025)
Dynamic Experts Search: Enhancing Reasoning in Mixture-of-Experts LLMs at Test Time
von: Han, Yixuan, et al.
Veröffentlicht: (2025)
von: Han, Yixuan, et al.
Veröffentlicht: (2025)
SteerConf: Steering LLMs for Confidence Elicitation
von: Zhou, Ziang, et al.
Veröffentlicht: (2025)
von: Zhou, Ziang, et al.
Veröffentlicht: (2025)
Focus On This, Not That! Steering LLMs with Adaptive Feature Specification
von: Lamb, Tom A., et al.
Veröffentlicht: (2024)
von: Lamb, Tom A., et al.
Veröffentlicht: (2024)
Scaling up Test-Time Compute with Latent Reasoning: A Recurrent Depth Approach
von: Geiping, Jonas, et al.
Veröffentlicht: (2025)
von: Geiping, Jonas, et al.
Veröffentlicht: (2025)
Fractional Reasoning via Latent Steering Vectors Improves Inference Time Compute
von: Liu, Sheng, et al.
Veröffentlicht: (2025)
von: Liu, Sheng, et al.
Veröffentlicht: (2025)
The Hidden Signal of Verifier Strictness: Controlling and Improving Step-Wise Verification via Selective Latent Steering
von: Zhou, Yefan, et al.
Veröffentlicht: (2026)
von: Zhou, Yefan, et al.
Veröffentlicht: (2026)
LaDiR: Latent Diffusion Enhances LLMs for Text Reasoning
von: Kang, Haoqiang, et al.
Veröffentlicht: (2025)
von: Kang, Haoqiang, et al.
Veröffentlicht: (2025)
Feature Extraction and Steering for Enhanced Chain-of-Thought Reasoning in Language Models
von: Li, Zihao, et al.
Veröffentlicht: (2025)
von: Li, Zihao, et al.
Veröffentlicht: (2025)
Adaptive Rollout Allocation for Online Reinforcement Learning with Verifiable Rewards
von: Nguyen, Hieu Trung, et al.
Veröffentlicht: (2026)
von: Nguyen, Hieu Trung, et al.
Veröffentlicht: (2026)
Adaptive Two-Phase Finetuning LLMs for Japanese Legal Text Retrieval
von: Trung, Quang Hoang, et al.
Veröffentlicht: (2024)
von: Trung, Quang Hoang, et al.
Veröffentlicht: (2024)
Examining Reasoning LLMs-as-Judges in Non-Verifiable LLM Post-Training
von: Liu, Yixin, et al.
Veröffentlicht: (2026)
von: Liu, Yixin, et al.
Veröffentlicht: (2026)
Steer LLM Latents for Hallucination Detection
von: Park, Seongheon, et al.
Veröffentlicht: (2025)
von: Park, Seongheon, et al.
Veröffentlicht: (2025)
Learning to Reason without External Rewards
von: Zhao, Xuandong, et al.
Veröffentlicht: (2025)
von: Zhao, Xuandong, et al.
Veröffentlicht: (2025)
Reinforcing General Reasoning without Verifiers
von: Zhou, Xiangxin, et al.
Veröffentlicht: (2025)
von: Zhou, Xiangxin, et al.
Veröffentlicht: (2025)
Steering the Verifiability of Multimodal AI Hallucinations
von: Pang, Jianhong, et al.
Veröffentlicht: (2026)
von: Pang, Jianhong, et al.
Veröffentlicht: (2026)
Inference-Time Rethinking with Latent Thought Vectors for Math Reasoning
von: Kong, Deqian, et al.
Veröffentlicht: (2026)
von: Kong, Deqian, et al.
Veröffentlicht: (2026)
Seek in the Dark: Reasoning via Test-Time Instance-Level Policy Gradient in Latent Space
von: Li, Hengli, et al.
Veröffentlicht: (2025)
von: Li, Hengli, et al.
Veröffentlicht: (2025)
Beyond Steering Vector: Flow-based Activation Steering for Inference-Time Intervention
von: Jin, Zehao, et al.
Veröffentlicht: (2026)
von: Jin, Zehao, et al.
Veröffentlicht: (2026)
Adaptive Test-Time Reasoning via Reward-Guided Dual-Phase Search
von: Cui, Yingqian, et al.
Veröffentlicht: (2025)
von: Cui, Yingqian, et al.
Veröffentlicht: (2025)
AAPO: Enhancing the Reasoning Capabilities of LLMs with Advantage Margin
von: Xiong, Jian, et al.
Veröffentlicht: (2025)
von: Xiong, Jian, et al.
Veröffentlicht: (2025)
TDD-Bench Verified: Can LLMs Generate Tests for Issues Before They Get Resolved?
von: Ahmed, Toufique, et al.
Veröffentlicht: (2024)
von: Ahmed, Toufique, et al.
Veröffentlicht: (2024)
Guiding Giants: Lightweight Controllers for Weighted Activation Steering in LLMs
von: Hegazy, Amr, et al.
Veröffentlicht: (2025)
von: Hegazy, Amr, et al.
Veröffentlicht: (2025)
Steering MoE LLMs via Expert (De)Activation
von: Fayyaz, Mohsen, et al.
Veröffentlicht: (2025)
von: Fayyaz, Mohsen, et al.
Veröffentlicht: (2025)
ShareChat: A Dataset of Chatbot Conversations in the Wild
von: Yan, Yueru, et al.
Veröffentlicht: (2025)
von: Yan, Yueru, et al.
Veröffentlicht: (2025)
Silence the Judge: Reinforcement Learning with Self-Verifier via Latent Geometric Clustering
von: Zhang, Nonghai, et al.
Veröffentlicht: (2026)
von: Zhang, Nonghai, et al.
Veröffentlicht: (2026)
Enhancing LLM Reasoning via Critique Models with Test-Time and Training-Time Supervision
von: Xi, Zhiheng, et al.
Veröffentlicht: (2024)
von: Xi, Zhiheng, et al.
Veröffentlicht: (2024)
MM-Verify: Enhancing Multimodal Reasoning with Chain-of-Thought Verification
von: Sun, Linzhuang, et al.
Veröffentlicht: (2025)
von: Sun, Linzhuang, et al.
Veröffentlicht: (2025)
Dynamic Cheatsheet: Test-Time Learning with Adaptive Memory
von: Suzgun, Mirac, et al.
Veröffentlicht: (2025)
von: Suzgun, Mirac, et al.
Veröffentlicht: (2025)
Adaptive Prompting for Continual Relation Extraction: A Within-Task Variance Perspective
von: Le, Minh, et al.
Veröffentlicht: (2024)
von: Le, Minh, et al.
Veröffentlicht: (2024)
Understanding How CodeLLMs (Mis)Predict Types with Activation Steering
von: Lucchetti, Francesca, et al.
Veröffentlicht: (2024)
von: Lucchetti, Francesca, et al.
Veröffentlicht: (2024)
Steering LLMs for Formal Theorem Proving
von: Kirtania, Shashank, et al.
Veröffentlicht: (2025)
von: Kirtania, Shashank, et al.
Veröffentlicht: (2025)
Latent-GRPO: Group Relative Policy Optimization for Latent Reasoning
von: Deng, Jingcheng, et al.
Veröffentlicht: (2026)
von: Deng, Jingcheng, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
Adapters Mixup: Mixing Parameter-Efficient Adapters to Enhance the Adversarial Robustness of Fine-tuned Pre-trained Text Classifiers
von: Nguyen, Tuc, et al.
Veröffentlicht: (2024) -
Generalizability of Mixture of Domain-Specific Adapters from the Lens of Signed Weight Directions and its Application to Effective Model Pruning
von: Nguyen, Tuc, et al.
Veröffentlicht: (2024) -
Unraveling Interwoven Roles of Large Language Models in Authorship Privacy: Obfuscation, Mimicking, and Verification
von: Nguyen, Tuc, et al.
Veröffentlicht: (2025) -
VerifierQ: Enhancing LLM Test Time Compute with Q-Learning-based Verifiers
von: Qi, Jianing, et al.
Veröffentlicht: (2024) -
Understanding and Steering the Cognitive Behaviors of Reasoning Models at Test-Time
von: Zhang, Zhenyu, et al.
Veröffentlicht: (2025)