Guardado en:
| Autores principales: | Guo, Wei, Lu, Siyuan, Ran, Xiangdong, Tong, Yiqi, Ban, Yikun, Xu, Zelong, Fan, Jing, Huang, Zixuan, Zhang, Xiao, Hu, Zhaojun, Zhuang, Fuzhen |
|---|---|
| Formato: | Preprint |
| Publicado: |
2026
|
| Materias: | |
| Acceso en línea: | https://arxiv.org/abs/2602.18749 |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
H2Tune: Federated Foundation Model Fine-Tuning with Hybrid Heterogeneity
por: Guo, Wei, et al.
Publicado: (2025)
por: Guo, Wei, et al.
Publicado: (2025)
UniFAR: A Unified Facet-Aware Retrieval Framework for Scientific Documents
por: Dou, Zheng, et al.
Publicado: (2026)
por: Dou, Zheng, et al.
Publicado: (2026)
A Comprehensive Survey of Federated Transfer Learning: Challenges, Methods and Applications
por: Guo, Wei, et al.
Publicado: (2024)
por: Guo, Wei, et al.
Publicado: (2024)
Proto-EVFL: Enhanced Vertical Federated Learning via Dual Prototype with Extremely Unaligned Data
por: Guo, Wei, et al.
Publicado: (2025)
por: Guo, Wei, et al.
Publicado: (2025)
Counterfactual Credit Policy Optimization for Multi-Agent Collaboration
por: Li, Zhongyi, et al.
Publicado: (2026)
por: Li, Zhongyi, et al.
Publicado: (2026)
Does Your Reasoning Model Implicitly Know When to Stop Thinking?
por: Huang, Zixuan, et al.
Publicado: (2026)
por: Huang, Zixuan, et al.
Publicado: (2026)
Weak-Driven Learning: How Weak Agents make Strong Agents Stronger
por: Chen, Zehao, et al.
Publicado: (2026)
por: Chen, Zehao, et al.
Publicado: (2026)
Real-Time Aligned Reward Model beyond Semantics
por: Huang, Zixuan, et al.
Publicado: (2026)
por: Huang, Zixuan, et al.
Publicado: (2026)
Contextual Rollout Bandits for Reinforcement Learning with Verifiable Rewards
por: Lu, Xiaodong, et al.
Publicado: (2026)
por: Lu, Xiaodong, et al.
Publicado: (2026)
Heterogeneous Agent Collaborative Reinforcement Learning
por: Zhang, Zhixia, et al.
Publicado: (2026)
por: Zhang, Zhixia, et al.
Publicado: (2026)
CoTMR: Chain-of-Thought Multi-Scale Reasoning for Training-Free Zero-Shot Composed Image Retrieval
por: Sun, Zelong, et al.
Publicado: (2025)
por: Sun, Zelong, et al.
Publicado: (2025)
AgriCHN: A Comprehensive Cross-domain Resource for Chinese Agricultural Named Entity Recognition
por: Zeng, Lingxiao, et al.
Publicado: (2025)
por: Zeng, Lingxiao, et al.
Publicado: (2025)
Novel reinforced wood material with a biomimetic hierarchical square honeycomb structure under quasi‐static loading: Simulation and experimental study
por: Zixuan Fan, et al.
Publicado: (2025)
por: Zixuan Fan, et al.
Publicado: (2025)
Mitigating Spurious Correlations Between Question and Answer via Chain-of-Thought Correctness Perception Distillation
por: Xie, Hongyan, et al.
Publicado: (2025)
por: Xie, Hongyan, et al.
Publicado: (2025)
Bridging Social Psychology and LLM Reasoning: Conflict-Aware Meta-Review Generation via Cognitive Alignment
por: Chen, Wei, et al.
Publicado: (2025)
por: Chen, Wei, et al.
Publicado: (2025)
Adaptive Robust Estimator for Multi-Agent Reinforcement Learning
por: Li, Zhongyi, et al.
Publicado: (2026)
por: Li, Zhongyi, et al.
Publicado: (2026)
Electrochemical Regioselective C(sp 2 )–H Selenylation of Pyrrolo[2,3‐ d ]pyrimidine Derivatives With Diselenides
por: Zixuan Liu, et al.
Publicado: (2026)
por: Zixuan Liu, et al.
Publicado: (2026)
Learnable Sampler Distillation for Discrete Diffusion Models
por: Fu, Feiyang, et al.
Publicado: (2025)
por: Fu, Feiyang, et al.
Publicado: (2025)
Your Group-Relative Advantage Is Biased
por: Yang, Fengkai, et al.
Publicado: (2026)
por: Yang, Fengkai, et al.
Publicado: (2026)
GCL-OT: Graph Contrastive Learning with Optimal Transport for Heterophilic Text-Attributed Graphs
por: Ren, Yating, et al.
Publicado: (2025)
por: Ren, Yating, et al.
Publicado: (2025)
A Market-Clearing-based Sensitivity Model for Locational Marginal and Average Carbon Emission
por: Lu, Zelong
Publicado: (2024)
por: Lu, Zelong
Publicado: (2024)
Adaptive Batch-Wise Sample Scheduling for Direct Preference Optimization
por: Huang, Zixuan, et al.
Publicado: (2025)
por: Huang, Zixuan, et al.
Publicado: (2025)
CODA: Difficulty-Aware Compute Allocation for Adaptive Reasoning
por: Wu, Siye, et al.
Publicado: (2026)
por: Wu, Siye, et al.
Publicado: (2026)
HESTIA: A Hessian-Guided Differentiable Quantization-Aware Training Framework for Extremely Low-Bit LLMs
por: Wang, Guoan, et al.
Publicado: (2026)
por: Wang, Guoan, et al.
Publicado: (2026)
BiFedKD: Bidirectional Federated Knowledge Distillation Framework for Non-IID and Long-Tailed ECG Monitoring
por: Shu, Zixuan, et al.
Publicado: (2026)
por: Shu, Zixuan, et al.
Publicado: (2026)
STRIDE: Learnable Stepwise Language Feedback for LLM Reasoning
por: Zhang, Junjie, et al.
Publicado: (2026)
por: Zhang, Junjie, et al.
Publicado: (2026)
Skill-Aware Data Selection and Fine-Tuning for Data-Efficient Reasoning Distillation
por: Zhang, Lechen, et al.
Publicado: (2026)
por: Zhang, Lechen, et al.
Publicado: (2026)
On Multilinear Forms for Mod $p$ Representations of $\mathrm{GL}_2(\mathbb{Q}_p)$
por: Fan, Yikun
Publicado: (2026)
por: Fan, Yikun
Publicado: (2026)
Why Distillation can Outperform Zero-RL: The Role of Flexible Reasoning
por: Hu, Xiao, et al.
Publicado: (2025)
por: Hu, Xiao, et al.
Publicado: (2025)
A General Deep Learning Framework for Wireless Resource Allocation under Discrete Constraints
por: Wang, Yikun, et al.
Publicado: (2026)
por: Wang, Yikun, et al.
Publicado: (2026)
FLeW: Facet-Level and Adaptive Weighted Representation Learning of Scientific Documents
por: Dou, Zheng, et al.
Publicado: (2025)
por: Dou, Zheng, et al.
Publicado: (2025)
FedPOB: Sample-Efficient Federated Prompt Optimization via Bandits
por: Lu, Pingchen, et al.
Publicado: (2025)
por: Lu, Pingchen, et al.
Publicado: (2025)
Mobility-Aware Asynchronous Federated Learning with Dynamic Sparsification
por: Yan, Jintao, et al.
Publicado: (2025)
por: Yan, Jintao, et al.
Publicado: (2025)
A Federated Online Restless Bandit Framework for Cooperative Resource Allocation
por: Tong, Jingwen, et al.
Publicado: (2024)
por: Tong, Jingwen, et al.
Publicado: (2024)
Learnability-Guided Diffusion for Dataset Distillation
por: Chan-Santiago, Jeffrey A., et al.
Publicado: (2026)
por: Chan-Santiago, Jeffrey A., et al.
Publicado: (2026)
SynGR: Unleashing the Potential of Cross-Modal Synergy for Generative Recommendation
por: Chen, Wei, et al.
Publicado: (2026)
por: Chen, Wei, et al.
Publicado: (2026)
Small-Scale-Fading-Aware Resource Allocation in Wireless Federated Learning
por: Wang, Jiacheng, et al.
Publicado: (2025)
por: Wang, Jiacheng, et al.
Publicado: (2025)
Policy Improvement Reinforcement Learning
por: Wang, Huaiyang, et al.
Publicado: (2026)
por: Wang, Huaiyang, et al.
Publicado: (2026)
LLMBoost: Make Large Language Models Stronger with Boosting
por: Chen, Zehao, et al.
Publicado: (2025)
por: Chen, Zehao, et al.
Publicado: (2025)
Neural Exploitation and Exploration of Contextual Bandits
por: Ban, Yikun, et al.
Publicado: (2023)
por: Ban, Yikun, et al.
Publicado: (2023)
Ejemplares similares
-
H2Tune: Federated Foundation Model Fine-Tuning with Hybrid Heterogeneity
por: Guo, Wei, et al.
Publicado: (2025) -
UniFAR: A Unified Facet-Aware Retrieval Framework for Scientific Documents
por: Dou, Zheng, et al.
Publicado: (2026) -
A Comprehensive Survey of Federated Transfer Learning: Challenges, Methods and Applications
por: Guo, Wei, et al.
Publicado: (2024) -
Proto-EVFL: Enhanced Vertical Federated Learning via Dual Prototype with Extremely Unaligned Data
por: Guo, Wei, et al.
Publicado: (2025) -
Counterfactual Credit Policy Optimization for Multi-Agent Collaboration
por: Li, Zhongyi, et al.
Publicado: (2026)