Gespeichert in:
| Hauptverfasser: | Luo, Yun, Wang, Futing, Cheng, Qianjia, Yu, Fangchen, Lei, Haodi, Yan, Jianhao, Li, Chenxi, Chen, Jiacheng, Zhao, Yufeng, Wan, Haiyuan, Zhang, Yuchen, Zheng, Shenghe, Yao, Junchi, Zhang, Qingyang, He, Haonan, Zeng, Wenxuan, Sheng, Li, Xie, Chengxing, Zuo, Yuxin, Li, Yizhuo, Wu, Yulun, Huang, Rui, Zhou, Dongzhan, Chen, Kai, Qiao, Yu, Bai, Lei, Cheng, Yu, Ding, Ning, Zhou, Bowen, Ye, Peng, Cui, Ganqu |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | https://arxiv.org/abs/2602.09443 |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
P1: Mastering Physics Olympiads with Reinforcement Learning
von: Chen, Jiacheng, et al.
Veröffentlicht: (2025)
von: Chen, Jiacheng, et al.
Veröffentlicht: (2025)
HiPhO: How Far Are (M)LLMs from Humans in the Latest High School Physics Olympiad Benchmark?
von: Yu, Fangchen, et al.
Veröffentlicht: (2025)
von: Yu, Fangchen, et al.
Veröffentlicht: (2025)
Scaling Physical Reasoning with the PHYSICS Dataset
von: Zheng, Shenghe, et al.
Veröffentlicht: (2025)
von: Zheng, Shenghe, et al.
Veröffentlicht: (2025)
PhysicsMinions: Winning Gold Medals in the Latest Physics Olympiads with a Coevolutionary Multimodal Multi-Agent System
von: Yu, Fangchen, et al.
Veröffentlicht: (2025)
von: Yu, Fangchen, et al.
Veröffentlicht: (2025)
Achieving Gold-Medal-Level Olympiad Reasoning via Simple and Unified Scaling
von: Li, Yafu, et al.
Veröffentlicht: (2026)
von: Li, Yafu, et al.
Veröffentlicht: (2026)
Draft-OPD: On-Policy Distillation for Speculative Draft Models
von: Lei, Haodi, et al.
Veröffentlicht: (2026)
von: Lei, Haodi, et al.
Veröffentlicht: (2026)
Reasoning via Video: The First Evaluation of Video Models' Reasoning Abilities through Maze-Solving Tasks
von: Yang, Cheng, et al.
Veröffentlicht: (2025)
von: Yang, Cheng, et al.
Veröffentlicht: (2025)
SCI-Verifier: Scientific Verifier with Thinking
von: Zheng, Shenghe, et al.
Veröffentlicht: (2025)
von: Zheng, Shenghe, et al.
Veröffentlicht: (2025)
Think Longer to Explore Deeper: Learn to Explore In-Context via Length-Incentivized Reinforcement Learning
von: Wang, Futing, et al.
Veröffentlicht: (2026)
von: Wang, Futing, et al.
Veröffentlicht: (2026)
From What to Why: A Multi-Agent System for Evidence-based Chemical Reaction Condition Reasoning
von: Yang, Cheng, et al.
Veröffentlicht: (2025)
von: Yang, Cheng, et al.
Veröffentlicht: (2025)
Teaching Thinking Models to Reason with Tools: A Full-Pipeline Recipe for Tool-Integrated Reasoning
von: Cheng, Qianjia, et al.
Veröffentlicht: (2026)
von: Cheng, Qianjia, et al.
Veröffentlicht: (2026)
Can Knowledge-Graph-based Retrieval Augmented Generation Really Retrieve What You Need?
von: Yu, Junchi, et al.
Veröffentlicht: (2025)
von: Yu, Junchi, et al.
Veröffentlicht: (2025)
Accessing GPT-4 level Mathematical Olympiad Solutions via Monte Carlo Tree Self-refine with LLaMa-3 8B
von: Zhang, Di, et al.
Veröffentlicht: (2024)
von: Zhang, Di, et al.
Veröffentlicht: (2024)
LLaMA-Berry: Pairwise Optimization for O1-like Olympiad-Level Mathematical Reasoning
von: Zhang, Di, et al.
Veröffentlicht: (2024)
von: Zhang, Di, et al.
Veröffentlicht: (2024)
Potential and Challenges of Model Editing for Social Debiasing
von: Yan, Jianhao, et al.
Veröffentlicht: (2024)
von: Yan, Jianhao, et al.
Veröffentlicht: (2024)
Learning to Reason under Off-Policy Guidance
von: Yan, Jianhao, et al.
Veröffentlicht: (2025)
von: Yan, Jianhao, et al.
Veröffentlicht: (2025)
SophiaVL-R1: Reinforcing MLLMs Reasoning with Thinking Reward
von: Fan, Kaixuan, et al.
Veröffentlicht: (2025)
von: Fan, Kaixuan, et al.
Veröffentlicht: (2025)
LIONs: An Empirically Optimized Approach to Align Language Models
von: Yu, Xiao, et al.
Veröffentlicht: (2024)
von: Yu, Xiao, et al.
Veröffentlicht: (2024)
Improved Laguerre Spectral Methods with Less Round-off Errors and Better Stability
von: Huang, Shenghe, et al.
Veröffentlicht: (2022)
von: Huang, Shenghe, et al.
Veröffentlicht: (2022)
OMIBench: Benchmarking Olympiad-Level Multi-Image Reasoning in Large Vision-Language Model
von: Chen, Qiguang, et al.
Veröffentlicht: (2026)
von: Chen, Qiguang, et al.
Veröffentlicht: (2026)
DeepResearch Arena: The First Exam of LLMs' Research Abilities via Seminar-Grounded Tasks
von: Wan, Haiyuan, et al.
Veröffentlicht: (2025)
von: Wan, Haiyuan, et al.
Veröffentlicht: (2025)
Keys to Robust Edits: from Theoretical Insights to Practical Advances
von: Yan, Jianhao, et al.
Veröffentlicht: (2024)
von: Yan, Jianhao, et al.
Veröffentlicht: (2024)
ELICIT: LLM Augmentation via External In-Context Capability
von: Wang, Futing, et al.
Veröffentlicht: (2024)
von: Wang, Futing, et al.
Veröffentlicht: (2024)
From Drafts to Answers: Unlocking LLM Potential via Aggregation Fine-Tuning
von: Li, Yafu, et al.
Veröffentlicht: (2025)
von: Li, Yafu, et al.
Veröffentlicht: (2025)
Deformation-based In-Context Learning for Point Cloud Understanding
von: Lin, Chengxing, et al.
Veröffentlicht: (2026)
von: Lin, Chengxing, et al.
Veröffentlicht: (2026)
UltraIF: Advancing Instruction Following from the Wild
von: An, Kaikai, et al.
Veröffentlicht: (2025)
von: An, Kaikai, et al.
Veröffentlicht: (2025)
MokA: Multimodal Low-Rank Adaptation for MLLMs
von: Wei, Yake, et al.
Veröffentlicht: (2025)
von: Wei, Yake, et al.
Veröffentlicht: (2025)
Attention Reallocation: Towards Zero-cost and Controllable Hallucination Mitigation of MLLMs
von: Tu, Chongjun, et al.
Veröffentlicht: (2025)
von: Tu, Chongjun, et al.
Veröffentlicht: (2025)
Unveiling Attractor Cycles in Large Language Models: A Dynamical Systems View of Successive Paraphrasing
von: Wang, Zhilin, et al.
Veröffentlicht: (2025)
von: Wang, Zhilin, et al.
Veröffentlicht: (2025)
Submodular flows and extreme flows on measurable spaces
von: Yu, Jing, et al.
Veröffentlicht: (2026)
von: Yu, Jing, et al.
Veröffentlicht: (2026)
Impacts of Food‐Based Flock Size on Foraging Patterns, Activity Time Budget and Foraging Efficiency: Flexible Behavioral Responses of the Wintering Oriental Storks (Ciconia boyciana) to Changes in Aquaculture at Shengjin Lake, China
von: Lei Cheng, et al.
Veröffentlicht: (2025)
von: Lei Cheng, et al.
Veröffentlicht: (2025)
Debiased Offline Representation Learning for Fast Online Adaptation in Non-stationary Dynamics
von: Zhang, Xinyu, et al.
Veröffentlicht: (2024)
von: Zhang, Xinyu, et al.
Veröffentlicht: (2024)
GuardTrace-VL: Detecting Unsafe Multimodel Reasoning via Iterative Safety Supervision
von: Xiang, Yuxiao, et al.
Veröffentlicht: (2025)
von: Xiang, Yuxiao, et al.
Veröffentlicht: (2025)
LabBuilder: Protocol-Grounded 3D Layout Generation for Interactable and Safe Laboratory
von: Cao, Jianbao, et al.
Veröffentlicht: (2026)
von: Cao, Jianbao, et al.
Veröffentlicht: (2026)
Bench2Drive-VL: Benchmarks for Closed-Loop Autonomous Driving with Vision-Language Models
von: Jia, Xiaosong, et al.
Veröffentlicht: (2026)
von: Jia, Xiaosong, et al.
Veröffentlicht: (2026)
TEMPO: Scaling Test-time Training for Large Reasoning Models
von: Zhang, Qingyang, et al.
Veröffentlicht: (2026)
von: Zhang, Qingyang, et al.
Veröffentlicht: (2026)
Adaptive Stopping for Multi-Turn LLM Reasoning
von: Zhou, Xiaofan, et al.
Veröffentlicht: (2026)
von: Zhou, Xiaofan, et al.
Veröffentlicht: (2026)
Q-Adapter: Customizing Pre-trained LLMs to New Preferences with Forgetting Mitigation
von: Li, Yi-Chen, et al.
Veröffentlicht: (2024)
von: Li, Yi-Chen, et al.
Veröffentlicht: (2024)
Invisible Backdoor Attacks on Diffusion Models
von: Li, Sen, et al.
Veröffentlicht: (2024)
von: Li, Sen, et al.
Veröffentlicht: (2024)
Enhancing Table Recognition with Vision LLMs: A Benchmark and Neighbor-Guided Toolchain Reasoner
von: Zhou, Yitong, et al.
Veröffentlicht: (2024)
von: Zhou, Yitong, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
P1: Mastering Physics Olympiads with Reinforcement Learning
von: Chen, Jiacheng, et al.
Veröffentlicht: (2025) -
HiPhO: How Far Are (M)LLMs from Humans in the Latest High School Physics Olympiad Benchmark?
von: Yu, Fangchen, et al.
Veröffentlicht: (2025) -
Scaling Physical Reasoning with the PHYSICS Dataset
von: Zheng, Shenghe, et al.
Veröffentlicht: (2025) -
PhysicsMinions: Winning Gold Medals in the Latest Physics Olympiads with a Coevolutionary Multimodal Multi-Agent System
von: Yu, Fangchen, et al.
Veröffentlicht: (2025) -
Achieving Gold-Medal-Level Olympiad Reasoning via Simple and Unified Scaling
von: Li, Yafu, et al.
Veröffentlicht: (2026)