Salvato in:
| Autori principali: | Zhang, Qingyang, Kong, Xinke, Wu, Haitao, Hu, Qinghua, Wu, Minghao, Yang, Baosong, Cheng, Yu, Luo, Yun, Cui, Ganqu, Zhang, Changqing |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2026
|
| Soggetti: | |
| Accesso online: | https://arxiv.org/abs/2604.19295 |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
COME: Test-time adaption by Conservatively Minimizing Entropy
di: Zhang, Qingyang, et al.
Pubblicazione: (2024)
di: Zhang, Qingyang, et al.
Pubblicazione: (2024)
Right Question is Already Half the Answer: Fully Unsupervised LLM Reasoning Incentivization
di: Zhang, Qingyang, et al.
Pubblicazione: (2025)
di: Zhang, Qingyang, et al.
Pubblicazione: (2025)
Computational Reasoning of Large Language Models
di: Wu, Haitao, et al.
Pubblicazione: (2025)
di: Wu, Haitao, et al.
Pubblicazione: (2025)
Scaling Physical Reasoning with the PHYSICS Dataset
di: Zheng, Shenghe, et al.
Pubblicazione: (2025)
di: Zheng, Shenghe, et al.
Pubblicazione: (2025)
The Best of Both Worlds: On the Dilemma of Out-of-distribution Detection
di: Zhang, Qingyang, et al.
Pubblicazione: (2024)
di: Zhang, Qingyang, et al.
Pubblicazione: (2024)
Test-Time Dynamic Image Fusion
di: Cao, Bing, et al.
Pubblicazione: (2024)
di: Cao, Bing, et al.
Pubblicazione: (2024)
Teaching Thinking Models to Reason with Tools: A Full-Pipeline Recipe for Tool-Integrated Reasoning
di: Cheng, Qianjia, et al.
Pubblicazione: (2026)
di: Cheng, Qianjia, et al.
Pubblicazione: (2026)
Meta-Reasoning: Semantics-Symbol Deconstruction for Large Language Models
di: Wang, Yiming, et al.
Pubblicazione: (2023)
di: Wang, Yiming, et al.
Pubblicazione: (2023)
Learning to Reason under Off-Policy Guidance
di: Yan, Jianhao, et al.
Pubblicazione: (2025)
di: Yan, Jianhao, et al.
Pubblicazione: (2025)
Spurious Feature Eraser: Stabilizing Test-Time Adaptation for Vision-Language Foundation Model
di: Ma, Huan, et al.
Pubblicazione: (2024)
di: Ma, Huan, et al.
Pubblicazione: (2024)
Scalable Efficient Training of Large Language Models with Low-dimensional Projected Attention
di: Lv, Xingtai, et al.
Pubblicazione: (2024)
di: Lv, Xingtai, et al.
Pubblicazione: (2024)
Dig2DIG: Dig into Diffusion Information Gains for Image Fusion
di: Cao, Bing, et al.
Pubblicazione: (2025)
di: Cao, Bing, et al.
Pubblicazione: (2025)
Generalized Few-Shot Out-of-Distribution Detection
di: Li, Pinxuan, et al.
Pubblicazione: (2025)
di: Li, Pinxuan, et al.
Pubblicazione: (2025)
Language as a Latent Variable for Reasoning Optimization
di: Wu, Linjuan, et al.
Pubblicazione: (2026)
di: Wu, Linjuan, et al.
Pubblicazione: (2026)
Multimodal Fusion on Low-quality Data: A Comprehensive Survey
di: Zhang, Qingyang, et al.
Pubblicazione: (2024)
di: Zhang, Qingyang, et al.
Pubblicazione: (2024)
Sampling-Efficient Test-Time Scaling: Self-Estimating the Best-of-N Sampling in Early Decoding
di: Wang, Yiming, et al.
Pubblicazione: (2025)
di: Wang, Yiming, et al.
Pubblicazione: (2025)
Think Longer to Explore Deeper: Learn to Explore In-Context via Length-Incentivized Reinforcement Learning
di: Wang, Futing, et al.
Pubblicazione: (2026)
di: Wang, Futing, et al.
Pubblicazione: (2026)
Bridging the Vision-Brain Gap with an Uncertainty-Aware Blur Prior
di: Wu, Haitao, et al.
Pubblicazione: (2025)
di: Wu, Haitao, et al.
Pubblicazione: (2025)
Unveiling Language-Specific Features in Large Language Models via Sparse Autoencoders
di: Deng, Boyi, et al.
Pubblicazione: (2025)
di: Deng, Boyi, et al.
Pubblicazione: (2025)
Helping CLIP See Both the Forest and the Trees: A Decomposition and Description Approach
di: Xue, Leyan, et al.
Pubblicazione: (2025)
di: Xue, Leyan, et al.
Pubblicazione: (2025)
Scaling Test-time Compute for LLM Agents
di: Zhu, King, et al.
Pubblicazione: (2025)
di: Zhu, King, et al.
Pubblicazione: (2025)
New Trends for Modern Machine Translation with Large Reasoning Models
di: Liu, Sinuo, et al.
Pubblicazione: (2025)
di: Liu, Sinuo, et al.
Pubblicazione: (2025)
Draft-OPD: On-Policy Distillation for Speculative Draft Models
di: Lei, Haodi, et al.
Pubblicazione: (2026)
di: Lei, Haodi, et al.
Pubblicazione: (2026)
Understanding and Steering the Cognitive Behaviors of Reasoning Models at Test-Time
di: Zhang, Zhenyu, et al.
Pubblicazione: (2025)
di: Zhang, Zhenyu, et al.
Pubblicazione: (2025)
Relax: An Asynchronous Reinforcement Learning Engine for Omni-Modal Post-Training at Scale
di: Zhang, Liujie, et al.
Pubblicazione: (2026)
di: Zhang, Liujie, et al.
Pubblicazione: (2026)
Teaching Large Reasoning Models Effective Reflection
di: Wang, Hanbin, et al.
Pubblicazione: (2026)
di: Wang, Hanbin, et al.
Pubblicazione: (2026)
Predictive Dynamic Fusion
di: Cao, Bing, et al.
Pubblicazione: (2024)
di: Cao, Bing, et al.
Pubblicazione: (2024)
Selective Learning: Towards Robust Calibration with Dynamic Regularization
di: Han, Zongbo, et al.
Pubblicazione: (2024)
di: Han, Zongbo, et al.
Pubblicazione: (2024)
TSO: Self-Training with Scaled Preference Optimization
di: Chen, Kaihui, et al.
Pubblicazione: (2024)
di: Chen, Kaihui, et al.
Pubblicazione: (2024)
A Contrastive Learning Based Convolutional Neural Network for ERP Brain-Computer Interfaces
di: Cui, Yuntian, et al.
Pubblicazione: (2024)
di: Cui, Yuntian, et al.
Pubblicazione: (2024)
MoE-CT: A Novel Approach For Large Language Models Training With Resistance To Catastrophic Forgetting
di: Li, Tianhao, et al.
Pubblicazione: (2024)
di: Li, Tianhao, et al.
Pubblicazione: (2024)
From Drafts to Answers: Unlocking LLM Potential via Aggregation Fine-Tuning
di: Li, Yafu, et al.
Pubblicazione: (2025)
di: Li, Yafu, et al.
Pubblicazione: (2025)
SimpleVLA-RL: Scaling VLA Training via Reinforcement Learning
di: Li, Haozhan, et al.
Pubblicazione: (2025)
di: Li, Haozhan, et al.
Pubblicazione: (2025)
Embedding Trajectory for Out-of-Distribution Detection in Mathematical Reasoning
di: Wang, Yiming, et al.
Pubblicazione: (2024)
di: Wang, Yiming, et al.
Pubblicazione: (2024)
Policy of Thoughts: Scaling LLM Reasoning via Test-time Policy Evolution
di: Jiao, Zhengbo, et al.
Pubblicazione: (2026)
di: Jiao, Zhengbo, et al.
Pubblicazione: (2026)
Scaling Reasoning without Attention
di: Zhao, Xueliang, et al.
Pubblicazione: (2025)
di: Zhao, Xueliang, et al.
Pubblicazione: (2025)
Z1: Efficient Test-time Scaling with Code
di: Yu, Zhaojian, et al.
Pubblicazione: (2025)
di: Yu, Zhaojian, et al.
Pubblicazione: (2025)
Achieving Gold-Medal-Level Olympiad Reasoning via Simple and Unified Scaling
di: Li, Yafu, et al.
Pubblicazione: (2026)
di: Li, Yafu, et al.
Pubblicazione: (2026)
Large Language Model for Multi-Domain Translation: Benchmarking and Domain CoT Fine-tuning
di: Hu, Tianxiang, et al.
Pubblicazione: (2024)
di: Hu, Tianxiang, et al.
Pubblicazione: (2024)
AnyTrans: Translate AnyText in the Image with Large Scale Models
di: Qian, Zhipeng, et al.
Pubblicazione: (2024)
di: Qian, Zhipeng, et al.
Pubblicazione: (2024)
Documenti analoghi
-
COME: Test-time adaption by Conservatively Minimizing Entropy
di: Zhang, Qingyang, et al.
Pubblicazione: (2024) -
Right Question is Already Half the Answer: Fully Unsupervised LLM Reasoning Incentivization
di: Zhang, Qingyang, et al.
Pubblicazione: (2025) -
Computational Reasoning of Large Language Models
di: Wu, Haitao, et al.
Pubblicazione: (2025) -
Scaling Physical Reasoning with the PHYSICS Dataset
di: Zheng, Shenghe, et al.
Pubblicazione: (2025) -
The Best of Both Worlds: On the Dilemma of Out-of-distribution Detection
di: Zhang, Qingyang, et al.
Pubblicazione: (2024)