State Tuning: State-based Test-Time Scaling on RWKV-7
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Xiao, Liu, Zhiyuan, Li, Yueyu, Lin |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Millions of States: Designing a Scalable MoE Architecture with RWKV-7 Meta-learner
von: Xiao, Liu, et al.
Veröffentlicht: (2025)
von: Xiao, Liu, et al.
Veröffentlicht: (2025)
Cross-attention for State-based model RWKV-7
von: Xiao, Liu, et al.
Veröffentlicht: (2025)
von: Xiao, Liu, et al.
Veröffentlicht: (2025)
WuNeng: Hybrid State with Attention
von: Xiao, Liu, et al.
Veröffentlicht: (2025)
von: Xiao, Liu, et al.
Veröffentlicht: (2025)
RWKV-7 "Goose" with Expressive Dynamic State Evolution
von: Peng, Bo, et al.
Veröffentlicht: (2025)
von: Peng, Bo, et al.
Veröffentlicht: (2025)
TrimR: Verifier-based Training-Free Thinking Compression for Efficient Test-Time Scaling
von: Lin, Weizhe, et al.
Veröffentlicht: (2025)
von: Lin, Weizhe, et al.
Veröffentlicht: (2025)
Revisiting the Test-Time Scaling of o1-like Models: Do they Truly Possess Test-Time Scaling Capabilities?
von: Zeng, Zhiyuan, et al.
Veröffentlicht: (2025)
von: Zeng, Zhiyuan, et al.
Veröffentlicht: (2025)
Test-Time Scaling with Reflective Generative Model
von: Wang, Zixiao, et al.
Veröffentlicht: (2025)
von: Wang, Zixiao, et al.
Veröffentlicht: (2025)
Scaling Linear Attention with Sparse State Expansion
von: Pan, Yuqi, et al.
Veröffentlicht: (2025)
von: Pan, Yuqi, et al.
Veröffentlicht: (2025)
Learning to (Learn at Test Time): RNNs with Expressive Hidden States
von: Sun, Yu, et al.
Veröffentlicht: (2024)
von: Sun, Yu, et al.
Veröffentlicht: (2024)
RWKVTTS: Yet another TTS based on RWKV-7
von: yueyu, Lin, et al.
Veröffentlicht: (2025)
von: yueyu, Lin, et al.
Veröffentlicht: (2025)
DynaPrompt: Dynamic Test-Time Prompt Tuning
von: Xiao, Zehao, et al.
Veröffentlicht: (2025)
von: Xiao, Zehao, et al.
Veröffentlicht: (2025)
Associative-State Universal Transformers: Sparse Retrieval Meets Structured Recurrence
von: Xiao, Liu
Veröffentlicht: (2026)
von: Xiao, Liu
Veröffentlicht: (2026)
Parameter-Efficient Fine-Tuning of State Space Models
von: Galim, Kevin, et al.
Veröffentlicht: (2024)
von: Galim, Kevin, et al.
Veröffentlicht: (2024)
Parallel Test-Time Scaling for Latent Reasoning Models
von: You, Runyang, et al.
Veröffentlicht: (2025)
von: You, Runyang, et al.
Veröffentlicht: (2025)
MiniMax-M1: Scaling Test-Time Compute Efficiently with Lightning Attention
von: MiniMax, et al.
Veröffentlicht: (2025)
von: MiniMax, et al.
Veröffentlicht: (2025)
Kinetics: Rethinking Test-Time Scaling Laws
von: Sadhukhan, Ranajoy, et al.
Veröffentlicht: (2025)
von: Sadhukhan, Ranajoy, et al.
Veröffentlicht: (2025)
Belief-State RWKV for Reinforcement Learning under Partial Observability
von: Xiao, Liu
Veröffentlicht: (2026)
von: Xiao, Liu
Veröffentlicht: (2026)
ARWKV: Pretrain is not what we need, an RNN-Attention-Based Language Model Born from Transformer
von: Yueyu, Lin, et al.
Veröffentlicht: (2025)
von: Yueyu, Lin, et al.
Veröffentlicht: (2025)
Test-Time Scaling Makes Overtraining Compute-Optimal
von: Roberts, Nicholas, et al.
Veröffentlicht: (2026)
von: Roberts, Nicholas, et al.
Veröffentlicht: (2026)
TTRL: Test-Time Reinforcement Learning
von: Zuo, Yuxin, et al.
Veröffentlicht: (2025)
von: Zuo, Yuxin, et al.
Veröffentlicht: (2025)
StateX: Enhancing RNN Recall via Post-training State Expansion
von: Shen, Xingyu, et al.
Veröffentlicht: (2025)
von: Shen, Xingyu, et al.
Veröffentlicht: (2025)
Scaling over Scaling: Exploring Test-Time Scaling Plateau in Large Reasoning Models
von: Wang, Jian, et al.
Veröffentlicht: (2025)
von: Wang, Jian, et al.
Veröffentlicht: (2025)
MetaScale: Test-Time Scaling with Evolving Meta-Thoughts
von: Liu, Qin, et al.
Veröffentlicht: (2025)
von: Liu, Qin, et al.
Veröffentlicht: (2025)
Scaling Test-Time Compute Without Verification or RL is Suboptimal
von: Setlur, Amrith, et al.
Veröffentlicht: (2025)
von: Setlur, Amrith, et al.
Veröffentlicht: (2025)
HEART: Emotionally-Driven Test-Time Scaling of Language Models
von: Pinto, Gabriela, et al.
Veröffentlicht: (2025)
von: Pinto, Gabriela, et al.
Veröffentlicht: (2025)
Do Not Waste Your Rollouts: Recycling Search Experience for Efficient Test-Time Scaling
von: Wang, Xinglin, et al.
Veröffentlicht: (2026)
von: Wang, Xinglin, et al.
Veröffentlicht: (2026)
Efficient Test-Time Scaling via Self-Calibration
von: Huang, Chengsong, et al.
Veröffentlicht: (2025)
von: Huang, Chengsong, et al.
Veröffentlicht: (2025)
Learning a Continue-Thinking Token for Enhanced Test-Time Scaling
von: Ringel, Liran, et al.
Veröffentlicht: (2025)
von: Ringel, Liran, et al.
Veröffentlicht: (2025)
VisualRWKV: Exploring Recurrent Neural Networks for Visual Language Models
von: Hou, Haowen, et al.
Veröffentlicht: (2024)
von: Hou, Haowen, et al.
Veröffentlicht: (2024)
Scaling LLM Test-Time Compute Optimally can be More Effective than Scaling Model Parameters
von: Snell, Charlie, et al.
Veröffentlicht: (2024)
von: Snell, Charlie, et al.
Veröffentlicht: (2024)
Stuffed Mamba: Oversized States Lead to the Inability to Forget
von: Chen, Yingfa, et al.
Veröffentlicht: (2024)
von: Chen, Yingfa, et al.
Veröffentlicht: (2024)
J1: Exploring Simple Test-Time Scaling for LLM-as-a-Judge
von: Chan, Chi-Min, et al.
Veröffentlicht: (2025)
von: Chan, Chi-Min, et al.
Veröffentlicht: (2025)
It's Not That Simple. An Analysis of Simple Test-Time Scaling
von: Wu, Guojun
Veröffentlicht: (2025)
von: Wu, Guojun
Veröffentlicht: (2025)
On the Role of Temperature Sampling in Test-Time Scaling
von: Wu, Yuheng, et al.
Veröffentlicht: (2025)
von: Wu, Yuheng, et al.
Veröffentlicht: (2025)
Crosslingual Reasoning through Test-Time Scaling
von: Yong, Zheng-Xin, et al.
Veröffentlicht: (2025)
von: Yong, Zheng-Xin, et al.
Veröffentlicht: (2025)
Reinforcement Learning Teachers of Test Time Scaling
von: Cetin, Edoardo, et al.
Veröffentlicht: (2025)
von: Cetin, Edoardo, et al.
Veröffentlicht: (2025)
BlackGoose Rimer: Harnessing RWKV-7 as a Simple yet Superior Replacement for Transformers in Large-Scale Time Series Modeling
von: weile, Li, et al.
Veröffentlicht: (2025)
von: weile, Li, et al.
Veröffentlicht: (2025)
Evaluating Judges as Evaluators: The JETTS Benchmark of LLM-as-Judges as Test-Time Scaling Evaluators
von: Zhou, Yilun, et al.
Veröffentlicht: (2025)
von: Zhou, Yilun, et al.
Veröffentlicht: (2025)
Scaling up Test-Time Compute with Latent Reasoning: A Recurrent Depth Approach
von: Geiping, Jonas, et al.
Veröffentlicht: (2025)
von: Geiping, Jonas, et al.
Veröffentlicht: (2025)
Optimizing Test-Time Compute via Meta Reinforcement Fine-Tuning
von: Qu, Yuxiao, et al.
Veröffentlicht: (2025)
von: Qu, Yuxiao, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Millions of States: Designing a Scalable MoE Architecture with RWKV-7 Meta-learner
von: Xiao, Liu, et al.
Veröffentlicht: (2025) -
Cross-attention for State-based model RWKV-7
von: Xiao, Liu, et al.
Veröffentlicht: (2025) -
WuNeng: Hybrid State with Attention
von: Xiao, Liu, et al.
Veröffentlicht: (2025) -
RWKV-7 "Goose" with Expressive Dynamic State Evolution
von: Peng, Bo, et al.
Veröffentlicht: (2025) -
TrimR: Verifier-based Training-Free Thinking Compression for Efficient Test-Time Scaling
von: Lin, Weizhe, et al.
Veröffentlicht: (2025)