Saved in:
| Main Authors: | Zhao, Zhengyang, Ma, Lu, Jiang, Yizhen, Ma, Xiaochen, Meng, Zimo, Shen, Chengyu, Tang, Lexiang, Sun, Haoze, Pei, Peng, Zhang, Wentao |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | https://arxiv.org/abs/2601.09233 |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Learning What Reinforcement Learning Can't: Interleaved Online Fine-Tuning for Hardest Questions
by: Ma, Lu, et al.
Published: (2025)
by: Ma, Lu, et al.
Published: (2025)
Training with Harnesses: On-Policy Harness Self-Distillation for Complex Reasoning
by: Zhao, Zhengyang, et al.
Published: (2026)
by: Zhao, Zhengyang, et al.
Published: (2026)
Leash: Adaptive Length Penalty and Reward Shaping for Efficient Large Reasoning Model
by: Li, Yanhao, et al.
Published: (2025)
by: Li, Yanhao, et al.
Published: (2025)
GIFT: Games as Informal Training for Generalizable LLMs
by: Lyu, Nuoyan, et al.
Published: (2026)
by: Lyu, Nuoyan, et al.
Published: (2026)
Let's Verify Math Questions Step by Step
by: Shen, Chengyu, et al.
Published: (2025)
by: Shen, Chengyu, et al.
Published: (2025)
DataFlex: A Unified Framework for Data-Centric Dynamic Training of Large Language Models
by: Liang, Hao, et al.
Published: (2026)
by: Liang, Hao, et al.
Published: (2026)
Towards Next-Generation LLM Training: From the Data-Centric Perspective
by: Liang, Hao, et al.
Published: (2026)
by: Liang, Hao, et al.
Published: (2026)
DARO: Difficulty-Aware Reweighting Policy Optimization
by: Zhou, Jingyu, et al.
Published: (2025)
by: Zhou, Jingyu, et al.
Published: (2025)
Semantic Surgery: Zero-Shot Concept Erasure in Diffusion Models
by: Xiong, Lexiang, et al.
Published: (2025)
by: Xiong, Lexiang, et al.
Published: (2025)
Are Bigger Encoders Always Better in Vision Large Models?
by: Li, Bozhou, et al.
Published: (2024)
by: Li, Bozhou, et al.
Published: (2024)
Uncovering Cross-Objective Interference in Multi-Objective Alignment
by: Lu, Yining, et al.
Published: (2026)
by: Lu, Yining, et al.
Published: (2026)
SpanNorm: Reconciling Training Stability and Performance in Deep Transformers
by: Wang, Chao, et al.
Published: (2026)
by: Wang, Chao, et al.
Published: (2026)
Faster and Better 3D Splatting via Group Training
by: Wang, Chengbo, et al.
Published: (2024)
by: Wang, Chengbo, et al.
Published: (2024)
Q-DiT: Accurate Post-Training Quantization for Diffusion Transformers
by: Chen, Lei, et al.
Published: (2024)
by: Chen, Lei, et al.
Published: (2024)
Thinking by Subtraction: Confidence-Driven Contrastive Decoding for LLM Reasoning
by: Tang, Lexiang, et al.
Published: (2026)
by: Tang, Lexiang, et al.
Published: (2026)
TraceAV-Bench: Benchmarking Multi-Hop Trajectory Reasoning over Long Audio-Visual Videos
by: Feng, Hengyi, et al.
Published: (2026)
by: Feng, Hengyi, et al.
Published: (2026)
DataFlow: An LLM-Driven Framework for Unified Data Preparation and Workflow Automation in the Era of Data-Centric AI
by: Liang, Hao, et al.
Published: (2025)
by: Liang, Hao, et al.
Published: (2025)
GIFT: Guided Importance-Aware Fine-Tuning for Diffusion Language Models
by: Xu, Guowei, et al.
Published: (2025)
by: Xu, Guowei, et al.
Published: (2025)
ANDES: Agent Native Data Evolving Synthesis Tool for Autonomous Instruction Alignment
by: Zhao, Zhengyang, et al.
Published: (2026)
by: Zhao, Zhengyang, et al.
Published: (2026)
Not All Tokens and Heads Are Equally Important: Dual-Level Attention Intervention for Hallucination Mitigation
by: Tang, Lexiang, et al.
Published: (2025)
by: Tang, Lexiang, et al.
Published: (2025)
High-Temperature Gibbs States are Unentangled and Efficiently Preparable
by: Bakshi, Ainesh, et al.
Published: (2024)
by: Bakshi, Ainesh, et al.
Published: (2024)
Objective Metrics for Evaluating Large Language Models Using External Data Sources
by: Du, Haoze, et al.
Published: (2025)
by: Du, Haoze, et al.
Published: (2025)
Dynamic Perturbation-Adaptive Adversarial Training on Medical Image Classification
by: Li, Shuai, et al.
Published: (2024)
by: Li, Shuai, et al.
Published: (2024)
Boosting Few-Shot Segmentation via Instance-Aware Data Augmentation and Local Consensus Guided Cross Attention
by: Guo, Li, et al.
Published: (2024)
by: Guo, Li, et al.
Published: (2024)
MetaMolGen: A Neural Graph Motif Generation Model for De Novo Molecular Design
by: Yan, Zimo, et al.
Published: (2025)
by: Yan, Zimo, et al.
Published: (2025)
K12-KGraph: A Curriculum-Aligned Knowledge Graph for Benchmarking and Training Educational LLMs
by: Liang, Hao, et al.
Published: (2026)
by: Liang, Hao, et al.
Published: (2026)
GIFT: Global Irreplaceability Frame Targeting for Efficient Video Understanding
by: Ma, Junpeng, et al.
Published: (2026)
by: Ma, Junpeng, et al.
Published: (2026)
Spatial-Spectral Binarized Neural Network for Panchromatic and Multi-spectral Images Fusion
by: Jiang, Yizhen, et al.
Published: (2025)
by: Jiang, Yizhen, et al.
Published: (2025)
Zero-to-Hero: Zero-Shot Initialization Empowering Reference-Based Video Appearance Editing
by: Su, Tongtong, et al.
Published: (2025)
by: Su, Tongtong, et al.
Published: (2025)
High-Temperature Fermionic Gibbs States are Mixtures of Gaussian States
by: Ramkumar, Akshar, et al.
Published: (2025)
by: Ramkumar, Akshar, et al.
Published: (2025)
Extrinsic derivatives for SDEs and SPDEs with distribution dependent noise
by: Ma, Xiaochen, et al.
Published: (2026)
by: Ma, Xiaochen, et al.
Published: (2026)
Coupling Methods and Applications on Path Dependent McKean-Vlasov SDEs
by: Huang, Xing, et al.
Published: (2024)
by: Huang, Xing, et al.
Published: (2024)
From Saying to Communicating: The Generic Development of Classroom Academic Presentations by Chinese First‐Year College Students
by: Junming Ma, et al.
Published: (2025)
by: Junming Ma, et al.
Published: (2025)
P‐5.10: A Method of Reducing the Warpage of Medium Size AMOLED Modules in High Temperature and Humidity Environment
by: Jianbing Ou, et al.
Published: (2024)
by: Jianbing Ou, et al.
Published: (2024)
Robust Training for Speaker Verification against Noisy Labels
by: Fang, Zhihua, et al.
Published: (2022)
by: Fang, Zhihua, et al.
Published: (2022)
HetSSNet: Spatial-Spectral Heterogeneous Graph Learning Network for Panchromatic and Multispectral Images Fusion
by: Ma, Mengting, et al.
Published: (2025)
by: Ma, Mengting, et al.
Published: (2025)
GIFT: Unlocking Full Potential of Labels in Distilled Dataset at Near-zero Cost
by: Shang, Xinyi, et al.
Published: (2024)
by: Shang, Xinyi, et al.
Published: (2024)
Dr. Post-Training: A Data Regularization Perspective on LLM Post-Training
by: Hu, Pingbang, et al.
Published: (2026)
by: Hu, Pingbang, et al.
Published: (2026)
HWL-HIN: A Hypergraph-Level Hypergraph Isomorphism Network as Powerful as the Hypergraph Weisfeiler-Lehman Test with Application to Higher-Order Network Robustness
by: Tian, Chengyu, et al.
Published: (2025)
by: Tian, Chengyu, et al.
Published: (2025)
A Factuality and Diversity Reconciled Decoding Method for Knowledge-Grounded Dialogue Generation
by: Yang, Chenxu, et al.
Published: (2024)
by: Yang, Chenxu, et al.
Published: (2024)
Similar Items
-
Learning What Reinforcement Learning Can't: Interleaved Online Fine-Tuning for Hardest Questions
by: Ma, Lu, et al.
Published: (2025) -
Training with Harnesses: On-Policy Harness Self-Distillation for Complex Reasoning
by: Zhao, Zhengyang, et al.
Published: (2026) -
Leash: Adaptive Length Penalty and Reward Shaping for Efficient Large Reasoning Model
by: Li, Yanhao, et al.
Published: (2025) -
GIFT: Games as Informal Training for Generalizable LLMs
by: Lyu, Nuoyan, et al.
Published: (2026) -
Let's Verify Math Questions Step by Step
by: Shen, Chengyu, et al.
Published: (2025)