ATLAS: Agentic Test-time Learning-to-Allocate Scaling
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Qin, Peijia, Cao, Qi, Xie, Pengtao |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
DAJ: Data-Reweighted LLM Judge for Test-Time Scaling in Code Generation
von: Qin, Peijia, et al.
Veröffentlicht: (2026)
von: Qin, Peijia, et al.
Veröffentlicht: (2026)
LLMs Know When They Know, but Do Not Act on It: A Metacognitive Harness for Test-time Scaling
von: Cao, Qi, et al.
Veröffentlicht: (2026)
von: Cao, Qi, et al.
Veröffentlicht: (2026)
DreamPRM-Code: Function-as-Step Process Reward Model with Label Correction for LLM Coding
von: Zhang, Ruiyi, et al.
Veröffentlicht: (2025)
von: Zhang, Ruiyi, et al.
Veröffentlicht: (2025)
BiDoRA: Bi-level Optimization-Based Weight-Decomposed Low-Rank Adaptation
von: Qin, Peijia, et al.
Veröffentlicht: (2024)
von: Qin, Peijia, et al.
Veröffentlicht: (2024)
FunPRM: Function-as-Step Process Reward Model with Meta Reward Correction for Code Generation
von: Zhang, Ruiyi, et al.
Veröffentlicht: (2026)
von: Zhang, Ruiyi, et al.
Veröffentlicht: (2026)
Models Under SCOPE: Scalable and Controllable Routing via Pre-hoc Reasoning
von: Cao, Qi, et al.
Veröffentlicht: (2026)
von: Cao, Qi, et al.
Veröffentlicht: (2026)
DreamPRM-1.5: Unlocking the Potential of Each Instance for Multimodal Process Reward Model Training
von: Cao, Qi, et al.
Veröffentlicht: (2025)
von: Cao, Qi, et al.
Veröffentlicht: (2025)
Send a SCOUT First: Pre-hoc Reasoning for Adaptive Detector Allocation in Prompt-Injection Defense
von: Zhang, Shuhao, et al.
Veröffentlicht: (2026)
von: Zhang, Shuhao, et al.
Veröffentlicht: (2026)
A Foundational Multi-Modal Model for Few-Shot Learning
von: Dang, Pengtao, et al.
Veröffentlicht: (2025)
von: Dang, Pengtao, et al.
Veröffentlicht: (2025)
DreamPRM: Domain-Reweighted Process Reward Model for Multimodal Reasoning
von: Cao, Qi, et al.
Veröffentlicht: (2025)
von: Cao, Qi, et al.
Veröffentlicht: (2025)
AIBuildAI: An AI Agent for Automatically Building AI Models
von: Zhang, Ruiyi, et al.
Veröffentlicht: (2026)
von: Zhang, Ruiyi, et al.
Veröffentlicht: (2026)
ATLAS: Adaptive Topology-based Learning at Scale for Homophilic and Heterophilic Graphs
von: Kundu, Turja, et al.
Veröffentlicht: (2025)
von: Kundu, Turja, et al.
Veröffentlicht: (2025)
TapWeight: Reweighting Pretraining Objectives for Task-Adaptive Pretraining
von: Zhang, Ruiyi, et al.
Veröffentlicht: (2024)
von: Zhang, Ruiyi, et al.
Veröffentlicht: (2024)
BiLoRA: A Bi-level Optimization Framework for Overfitting-Resilient Low-Rank Adaptation of Large Pre-trained Models
von: Qiang, Rushi, et al.
Veröffentlicht: (2024)
von: Qiang, Rushi, et al.
Veröffentlicht: (2024)
Towards Large-Scale In-Context Reinforcement Learning by Meta-Training in Randomized Worlds
von: Wang, Fan, et al.
Veröffentlicht: (2025)
von: Wang, Fan, et al.
Veröffentlicht: (2025)
Neural Operators for Predictor Feedback Control of Nonlinear Delay Systems
von: Bhan, Luke, et al.
Veröffentlicht: (2024)
von: Bhan, Luke, et al.
Veröffentlicht: (2024)
Scaling Test-Time Compute for Agentic Coding
von: Kim, Joongwon, et al.
Veröffentlicht: (2026)
von: Kim, Joongwon, et al.
Veröffentlicht: (2026)
Every Rollout Counts: Optimal Resource Allocation for Efficient Test-Time Scaling
von: Wang, Xinglin, et al.
Veröffentlicht: (2025)
von: Wang, Xinglin, et al.
Veröffentlicht: (2025)
SteganoBackdoor: Stealthy and Data-Efficient Backdoor Attacks on Language Models
von: Xue, Eric, et al.
Veröffentlicht: (2025)
von: Xue, Eric, et al.
Veröffentlicht: (2025)
Small-Scale-Fading-Aware Resource Allocation in Wireless Federated Learning
von: Wang, Jiacheng, et al.
Veröffentlicht: (2025)
von: Wang, Jiacheng, et al.
Veröffentlicht: (2025)
RL-Selector: Reinforcement Learning-Guided Data Selection via Redundancy Assessment
von: Yang, Suorong, et al.
Veröffentlicht: (2025)
von: Yang, Suorong, et al.
Veröffentlicht: (2025)
AutoLoRA: Automatically Tuning Matrix Ranks in Low-Rank Adaptation Based on Meta Learning
von: Zhang, Ruiyi, et al.
Veröffentlicht: (2024)
von: Zhang, Ruiyi, et al.
Veröffentlicht: (2024)
SSR: Speculative Parallel Scaling Reasoning in Test-time
von: Chu, Yuanlin, et al.
Veröffentlicht: (2025)
von: Chu, Yuanlin, et al.
Veröffentlicht: (2025)
Adaptive Reinforcement Learning for Dynamic Configuration Allocation in Pre-Production Testing
von: Zhu, Yu
Veröffentlicht: (2025)
von: Zhu, Yu
Veröffentlicht: (2025)
DynamicFL: Federated Learning with Dynamic Communication Resource Allocation
von: Le, Qi, et al.
Veröffentlicht: (2024)
von: Le, Qi, et al.
Veröffentlicht: (2024)
Equivariant Spherical Transformer for Efficient Molecular Modeling
von: An, Junyi, et al.
Veröffentlicht: (2025)
von: An, Junyi, et al.
Veröffentlicht: (2025)
TEMPO: Scaling Test-time Training for Large Reasoning Models
von: Zhang, Qingyang, et al.
Veröffentlicht: (2026)
von: Zhang, Qingyang, et al.
Veröffentlicht: (2026)
Efficient Privacy-Preserving KAN Inference Using Homomorphic Encryption
von: Lai, Zhizheng, et al.
Veröffentlicht: (2024)
von: Lai, Zhizheng, et al.
Veröffentlicht: (2024)
How to Allocate, How to Learn? Dynamic Rollout Allocation and Advantage Modulation for Policy Optimization
von: Fang, Yangyi, et al.
Veröffentlicht: (2026)
von: Fang, Yangyi, et al.
Veröffentlicht: (2026)
S*: Test Time Scaling for Code Generation
von: Li, Dacheng, et al.
Veröffentlicht: (2025)
von: Li, Dacheng, et al.
Veröffentlicht: (2025)
Dynamic Prompt Allocation and Tuning for Continual Test-Time Adaptation
von: Cui, Chaoran, et al.
Veröffentlicht: (2024)
von: Cui, Chaoran, et al.
Veröffentlicht: (2024)
PACEvolve++: Improving Test-time Learning for Evolutionary Search Agents
von: Yan, Minghao, et al.
Veröffentlicht: (2026)
von: Yan, Minghao, et al.
Veröffentlicht: (2026)
AM-SAM: Automated Prompting and Mask Calibration for Segment Anything Model
von: Li, Yuchen, et al.
Veröffentlicht: (2024)
von: Li, Yuchen, et al.
Veröffentlicht: (2024)
ScaleNet: Scale Invariance Learning in Directed Graphs
von: Jiang, Qin, et al.
Veröffentlicht: (2024)
von: Jiang, Qin, et al.
Veröffentlicht: (2024)
Multi-Objective Reinforcement Learning for Large-Scale Tote Allocation in Human-Robot Collaborative Fulfillment Centers
von: Sengupta, Sikata, et al.
Veröffentlicht: (2026)
von: Sengupta, Sikata, et al.
Veröffentlicht: (2026)
Exploring Test-time Scaling via Prediction Merging on Large-Scale Recommendation
von: Lyu, Fuyuan, et al.
Veröffentlicht: (2025)
von: Lyu, Fuyuan, et al.
Veröffentlicht: (2025)
EQA-RM: A Generative Embodied Reward Model with Test-time Scaling
von: Chen, Yuhang, et al.
Veröffentlicht: (2025)
von: Chen, Yuhang, et al.
Veröffentlicht: (2025)
SKILL0: In-Context Agentic Reinforcement Learning for Skill Internalization
von: Lu, Zhengxi, et al.
Veröffentlicht: (2026)
von: Lu, Zhengxi, et al.
Veröffentlicht: (2026)
ATLAS: Adaptive Test-Time Latent Steering with External Verifiers for Enhancing LLMs Reasoning
von: Nguyen, Tuc, et al.
Veröffentlicht: (2026)
von: Nguyen, Tuc, et al.
Veröffentlicht: (2026)
Agentic DDQN-Based Scheduling for Licensed and Unlicensed Band Allocation in Sidelink Networks
von: Chou, Po-Heng, et al.
Veröffentlicht: (2025)
von: Chou, Po-Heng, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
DAJ: Data-Reweighted LLM Judge for Test-Time Scaling in Code Generation
von: Qin, Peijia, et al.
Veröffentlicht: (2026) -
LLMs Know When They Know, but Do Not Act on It: A Metacognitive Harness for Test-time Scaling
von: Cao, Qi, et al.
Veröffentlicht: (2026) -
DreamPRM-Code: Function-as-Step Process Reward Model with Label Correction for LLM Coding
von: Zhang, Ruiyi, et al.
Veröffentlicht: (2025) -
BiDoRA: Bi-level Optimization-Based Weight-Decomposed Low-Rank Adaptation
von: Qin, Peijia, et al.
Veröffentlicht: (2024) -
FunPRM: Function-as-Step Process Reward Model with Meta Reward Correction for Code Generation
von: Zhang, Ruiyi, et al.
Veröffentlicht: (2026)