From Atoms to Chains: Divergence-Guided Reasoning Curriculum for Unlabeled LLM Domain Adaptation
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Wang, Yongqi, Ji, Xiaofeng, Wang, Jie, Li, Qingbin, Xiong, Xiao, Yang, Zheming, Xu, Jian, Qiu, Minghui, Wu, Xinxiao |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
CURE: Critical-Token-Guided Re-Concatenation for Entropy-Collapse Prevention
von: Li, Qingbin, et al.
Veröffentlicht: (2025)
von: Li, Qingbin, et al.
Veröffentlicht: (2025)
METOR: A Unified Framework for Mutual Enhancement of Objects and Relationships in Open-vocabulary Video Visual Relationship Detection
von: Wang, Yongqi, et al.
Veröffentlicht: (2025)
von: Wang, Yongqi, et al.
Veröffentlicht: (2025)
Reinforced MLLM: A Survey on RL-Based Reasoning in Multimodal Large Language Models
von: Zhou, Guanghao, et al.
Veröffentlicht: (2025)
von: Zhou, Guanghao, et al.
Veröffentlicht: (2025)
End-to-end Open-vocabulary Video Visual Relationship Detection using Multi-modal Prompting
von: Wang, Yongqi, et al.
Veröffentlicht: (2024)
von: Wang, Yongqi, et al.
Veröffentlicht: (2024)
Not All Negative Samples Are Equal: LLMs Learn Better from Plausible Reasoning
von: Di, Zixiang, et al.
Veröffentlicht: (2026)
von: Di, Zixiang, et al.
Veröffentlicht: (2026)
ElimPCL: Eliminating Noise Accumulation with Progressive Curriculum Labeling for Source-Free Domain Adaptation
von: Cheng, Jie, et al.
Veröffentlicht: (2025)
von: Cheng, Jie, et al.
Veröffentlicht: (2025)
From Answers to Arguments: Toward Trustworthy Clinical Diagnostic Reasoning with Toulmin-Guided Curriculum Goal-Conditioned Learning
von: Zhan, Chen, et al.
Veröffentlicht: (2026)
von: Zhan, Chen, et al.
Veröffentlicht: (2026)
Data-free Multi-label Image Recognition via LLM-powered Prompt Tuning
von: Yang, Shuo, et al.
Veröffentlicht: (2024)
von: Yang, Shuo, et al.
Veröffentlicht: (2024)
From Reasoning Chains to Verifiable Subproblems: Curriculum Reinforcement Learning Enables Credit Assignment for LLM Reasoning
von: Jiang, Xitai, et al.
Veröffentlicht: (2026)
von: Jiang, Xitai, et al.
Veröffentlicht: (2026)
Heterogeneous Domain Adaptation with Positive and Unlabeled Data
von: Mori, Junki, et al.
Veröffentlicht: (2023)
von: Mori, Junki, et al.
Veröffentlicht: (2023)
Towards Real-world Lens Active Alignment with Unlabeled Data via Domain Adaptation
von: Li, Wenyong, et al.
Veröffentlicht: (2026)
von: Li, Wenyong, et al.
Veröffentlicht: (2026)
Domain Adaptation with Cauchy-Schwarz Divergence
von: Yin, Wenzhe, et al.
Veröffentlicht: (2024)
von: Yin, Wenzhe, et al.
Veröffentlicht: (2024)
Logo-VGR: Visual Grounded Reasoning for Open-world Logo Recognition
von: Liang, Zichen, et al.
Veröffentlicht: (2025)
von: Liang, Zichen, et al.
Veröffentlicht: (2025)
Protecting Model Adaptation from Trojans in the Unlabeled Data
von: Sheng, Lijun, et al.
Veröffentlicht: (2024)
von: Sheng, Lijun, et al.
Veröffentlicht: (2024)
Mimic Human Cognition, Master Multi-Image Reasoning: A Meta-Action Framework for Enhanced Visual Understanding
von: Yin, Jianghao, et al.
Veröffentlicht: (2026)
von: Yin, Jianghao, et al.
Veröffentlicht: (2026)
On $f$-Divergence Principled Domain Adaptation: An Improved Framework
von: Wang, Ziqiao, et al.
Veröffentlicht: (2024)
von: Wang, Ziqiao, et al.
Veröffentlicht: (2024)
Domain-Hierarchy Adaptation via Chain of Iterative Reasoning for Few-shot Hierarchical Text Classification
von: Ji, Ke, et al.
Veröffentlicht: (2024)
von: Ji, Ke, et al.
Veröffentlicht: (2024)
CLIP the Divergence: Language-guided Unsupervised Domain Adaptation
von: Zhu, Jinjing, et al.
Veröffentlicht: (2024)
von: Zhu, Jinjing, et al.
Veröffentlicht: (2024)
GThinker: Towards General Multimodal Reasoning via Cue-Guided Rethinking
von: Zhan, Yufei, et al.
Veröffentlicht: (2025)
von: Zhan, Yufei, et al.
Veröffentlicht: (2025)
LLM-powered Query Expansion for Enhancing Boundary Prediction in Language-driven Action Localization
von: Shang, Zirui, et al.
Veröffentlicht: (2025)
von: Shang, Zirui, et al.
Veröffentlicht: (2025)
LLM-enhanced Action-aware Multi-modal Prompt Tuning for Image-Text Matching
von: Tian, Mengxiao, et al.
Veröffentlicht: (2025)
von: Tian, Mengxiao, et al.
Veröffentlicht: (2025)
ThinkDrive: Chain-of-Thought Guided Progressive Reinforcement Learning Fine-Tuning for Autonomous Driving
von: Zhao, Chang, et al.
Veröffentlicht: (2026)
von: Zhao, Chang, et al.
Veröffentlicht: (2026)
Probing the Crossover between Dynamical Phases with Local Correlations in a Rydberg Atom Array
von: Wu, Xiaofeng, et al.
Veröffentlicht: (2025)
von: Wu, Xiaofeng, et al.
Veröffentlicht: (2025)
Reasoning Curriculum: Bootstrapping Broad LLM Reasoning from Math
von: Pang, Bo, et al.
Veröffentlicht: (2025)
von: Pang, Bo, et al.
Veröffentlicht: (2025)
Simulate, Refocus and Ensemble: An Attention-Refocusing Scheme for Domain Generalization
von: Wang, Ziyi, et al.
Veröffentlicht: (2025)
von: Wang, Ziyi, et al.
Veröffentlicht: (2025)
Uncertainty-quantified Rollout Policy Adaptation for Unlabelled Cross-domain Temporal Grounding
von: Hu, Jian, et al.
Veröffentlicht: (2025)
von: Hu, Jian, et al.
Veröffentlicht: (2025)
Self-Evolving Curriculum for LLM Reasoning
von: Chen, Xiaoyin, et al.
Veröffentlicht: (2025)
von: Chen, Xiaoyin, et al.
Veröffentlicht: (2025)
Enhancing Continuous Domain Adaptation with Multi-Path Transfer Curriculum
von: Liu, Hanbing, et al.
Veröffentlicht: (2024)
von: Liu, Hanbing, et al.
Veröffentlicht: (2024)
Automatic Curriculum Expert Iteration for Reliable LLM Reasoning
von: Zhao, Zirui, et al.
Veröffentlicht: (2024)
von: Zhao, Zirui, et al.
Veröffentlicht: (2024)
LLM Reasoning Is Latent, Not the Chain of Thought
von: Wang, Wenshuo
Veröffentlicht: (2026)
von: Wang, Wenshuo
Veröffentlicht: (2026)
Learning Unlabeled Clients Divergence for Federated Semi-Supervised Learning via Anchor Model Aggregation
von: Elbatel, Marawan, et al.
Veröffentlicht: (2024)
von: Elbatel, Marawan, et al.
Veröffentlicht: (2024)
Federated Source-free Domain Adaptation for Classification: Weighted Cluster Aggregation for Unlabeled Data
von: Mori, Junki, et al.
Veröffentlicht: (2024)
von: Mori, Junki, et al.
Veröffentlicht: (2024)
DiSCTT: Consensus-Guided Self-Curriculum for Efficient Test-Time Adaptation in Reasoning
von: Moradi, Mohammad Mahdi, et al.
Veröffentlicht: (2026)
von: Moradi, Mohammad Mahdi, et al.
Veröffentlicht: (2026)
SORM‐Enhanced Inverse Reliability Analysis for Geotechnical Multiobjective Reliability‐Based Design Optimization
von: Tao Wang, et al.
Veröffentlicht: (2024)
von: Tao Wang, et al.
Veröffentlicht: (2024)
Empowering Source-Free Domain Adaptation via MLLM-Guided Reliability-Based Curriculum Learning
von: Chen, Dongjie, et al.
Veröffentlicht: (2024)
von: Chen, Dongjie, et al.
Veröffentlicht: (2024)
An Area-Efficient 20-100-GHz Phase-Invariant Switch-Type Attenuator Achieving 0.1-dB Tuning Step in 65-nm CMOS
von: Li, Qingbin, et al.
Veröffentlicht: (2025)
von: Li, Qingbin, et al.
Veröffentlicht: (2025)
From Literature to Lab: Closed-Loop Advancement of Perovskite Solar Cells via Domain Knowledge Guided LLM
von: Sun, Penglei, et al.
Veröffentlicht: (2026)
von: Sun, Penglei, et al.
Veröffentlicht: (2026)
From OSS to Open Source AI: an Exploratory Study of Collaborative Development Paradigm Divergence
von: Ye, Hengzhi, et al.
Veröffentlicht: (2026)
von: Ye, Hengzhi, et al.
Veröffentlicht: (2026)
DisSR: Disentangling Speech Representation for Degradation-Prior Guided Cross-Domain Speech Restoration
von: Liang, Ziqi, et al.
Veröffentlicht: (2026)
von: Liang, Ziqi, et al.
Veröffentlicht: (2026)
Where and What: Reasoning Dynamic and Implicit Preferences in Situated Conversational Recommendation
von: Lin, Dongding, et al.
Veröffentlicht: (2026)
von: Lin, Dongding, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
CURE: Critical-Token-Guided Re-Concatenation for Entropy-Collapse Prevention
von: Li, Qingbin, et al.
Veröffentlicht: (2025) -
METOR: A Unified Framework for Mutual Enhancement of Objects and Relationships in Open-vocabulary Video Visual Relationship Detection
von: Wang, Yongqi, et al.
Veröffentlicht: (2025) -
Reinforced MLLM: A Survey on RL-Based Reasoning in Multimodal Large Language Models
von: Zhou, Guanghao, et al.
Veröffentlicht: (2025) -
End-to-end Open-vocabulary Video Visual Relationship Detection using Multi-modal Prompting
von: Wang, Yongqi, et al.
Veröffentlicht: (2024) -
Not All Negative Samples Are Equal: LLMs Learn Better from Plausible Reasoning
von: Di, Zixiang, et al.
Veröffentlicht: (2026)