CAdam: Confidence-Based Optimization for Online Learning
Fuente:
arXiv
Saved in:
| Main Authors: | Wang, Shaowen, Liu, Anan, Xiao, Jian, Liu, Huan, Yang, Yuekui, Xu, Cong, Pu, Qianqian, Zheng, Suncong, Zhang, Wei, Wang, Di, Jiang, Jie, Li, Jian |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Efficient Cross-Architecture Knowledge Transfer for Large-Scale Online User Response Prediction
by: Wu, Yucheng, et al.
Published: (2026)
by: Wu, Yucheng, et al.
Published: (2026)
ICPO: Intrinsic Confidence-Driven Group Relative Preference Optimization for Efficient Reinforcement Learning
by: Wang, Jinpeng, et al.
Published: (2025)
by: Wang, Jinpeng, et al.
Published: (2025)
ROAD: Adaptive Data Mixing for Offline-to-Online Reinforcement Learning via Bi-Level Optimization
by: Yang, Letian, et al.
Published: (2026)
by: Yang, Letian, et al.
Published: (2026)
FlexHB: a More Efficient and Flexible Framework for Hyperparameter Optimization
by: Zhang, Yang, et al.
Published: (2024)
by: Zhang, Yang, et al.
Published: (2024)
Revealing the Learning Dynamics of Long-Context Continual Pre-training
by: Liang, Yupu, et al.
Published: (2026)
by: Liang, Yupu, et al.
Published: (2026)
MFmamba: A Multi-function Network for Panchromatic Image Resolution Restoration Based on State-Space Model
by: Jiang, Qian, et al.
Published: (2025)
by: Jiang, Qian, et al.
Published: (2025)
Dynamic Correlation Learning and Regularization for Multi-Label Confidence Calibration
by: Chen, Tianshui, et al.
Published: (2024)
by: Chen, Tianshui, et al.
Published: (2024)
Understanding LLM Behaviors via Compression: Data Generation, Knowledge Acquisition and Scaling Laws
by: Pan, Zhixuan, et al.
Published: (2025)
by: Pan, Zhixuan, et al.
Published: (2025)
LoRA-GA: Low-Rank Adaptation with Gradient Approximation
by: Wang, Shaowen, et al.
Published: (2024)
by: Wang, Shaowen, et al.
Published: (2024)
CAdam: Context-Adaptive Moment Estimation for 3D Gaussian Densification in Generative Distillation
by: Chung, SeungJeh, et al.
Published: (2026)
by: Chung, SeungJeh, et al.
Published: (2026)
Inference Time Optimization with Confidence Dynamics
by: Wang, Yu, et al.
Published: (2026)
by: Wang, Yu, et al.
Published: (2026)
Designing Pin-pression Gripper and Learning its Dexterous Grasping with Online In-hand Adjustment
by: Xiao, Hewen, et al.
Published: (2025)
by: Xiao, Hewen, et al.
Published: (2025)
Online Shopping and Financial Literacy Among the Elderly
by: Xu Cui, et al.
Published: (2025)
by: Xu Cui, et al.
Published: (2025)
Pushing the Limits of Low-Bit Optimizers: A Focus on EMA Dynamics
by: Xu, Cong, et al.
Published: (2025)
by: Xu, Cong, et al.
Published: (2025)
Adaptive Cache Pollution Control for Large Language Model Inference Workloads Using Temporal CNN-Based Prediction and Priority-Aware Replacement
by: Liu, Songze, et al.
Published: (2025)
by: Liu, Songze, et al.
Published: (2025)
Reasoning-Enhanced Object-Centric Learning for Videos
by: Li, Jian, et al.
Published: (2024)
by: Li, Jian, et al.
Published: (2024)
End-to-End Optimization of LLM-Driven Multi-Agent Search Systems via Heterogeneous-Group-Based Reinforcement Learning
by: Chen, Guanzhong, et al.
Published: (2025)
by: Chen, Guanzhong, et al.
Published: (2025)
Generalization in Online Reinforcement Learning for Mobile Agents
by: Gu, Li, et al.
Published: (2026)
by: Gu, Li, et al.
Published: (2026)
The Smoluchowski-Kramers approximation for a system with arbitrary friction depending on both state and distribution
by: Liu, Xueru, et al.
Published: (2024)
by: Liu, Xueru, et al.
Published: (2024)
AdaS&S: a One-Shot Supernet Approach for Automatic Embedding Size Search in Deep Recommender System
by: Wei, He, et al.
Published: (2024)
by: Wei, He, et al.
Published: (2024)
G$^2$IFWI: Gradient-Coupled Diffusion Priors for Regularizing Implicit Full Waveform Inversion
by: Wang, Shaowen
Published: (2026)
by: Wang, Shaowen
Published: (2026)
Asynchronous Parallel Reinforcement Learning for Optimizing Propulsive Performance in Fin Ray Control
by: Liu, Xin-Yang, et al.
Published: (2024)
by: Liu, Xin-Yang, et al.
Published: (2024)
MPA: Multimodal Prototype Augmentation for Few-Shot Learning
by: Wu, Liwen, et al.
Published: (2026)
by: Wu, Liwen, et al.
Published: (2026)
A total-shear-stress-conserved wall model for large-eddy simulation of high-Reynolds number wall turbulence
by: Liu, Huan-Cong, et al.
Published: (2024)
by: Liu, Huan-Cong, et al.
Published: (2024)
SDOG: Scalable Scheduling of Flows Based on Dynamic Online Grouping in Industrial Time‐Sensitive Networks
by: Chang Liu, et al.
Published: (2025)
by: Chang Liu, et al.
Published: (2025)
Designed Synthesis of an Aluminum Molecular Ring Based Rotaxane and Polyrotaxane
by: Ya‐Jie Liu, et al.
Published: (2024)
by: Ya‐Jie Liu, et al.
Published: (2024)
Designed Synthesis of an Aluminum Molecular Ring Based Rotaxane and Polyrotaxane
by: Ya‐Jie Liu, et al.
Published: (2024)
by: Ya‐Jie Liu, et al.
Published: (2024)
Revealing Vulnerabilities in Stable Diffusion via Targeted Attacks
by: Zhang, Chenyu, et al.
Published: (2024)
by: Zhang, Chenyu, et al.
Published: (2024)
Occlusion-Aware 3D Hand-Object Pose Estimation with Masked AutoEncoders
by: Yang, Hui, et al.
Published: (2025)
by: Yang, Hui, et al.
Published: (2025)
PPO‐Based Reinforcement Learning for the Semi‐Active Vibration Control of MDOF Platform
by: Wei Huang, et al.
Published: (2026)
by: Wei Huang, et al.
Published: (2026)
CRPO: Confidence-Reward Driven Preference Optimization for Machine Translation
by: Cui, Guofeng, et al.
Published: (2025)
by: Cui, Guofeng, et al.
Published: (2025)
Online Meta Adaptation for Variable-Rate Learned Image Compression
by: Jiang, Wei, et al.
Published: (2021)
by: Jiang, Wei, et al.
Published: (2021)
SSD: Spatial-Semantic Head Decoupling for Efficient Autoregressive Image Generation
by: Jian, Siyong, et al.
Published: (2025)
by: Jian, Siyong, et al.
Published: (2025)
Minimum Area Confidence Set Optimality for Simultaneous Confidence Bands for Percentiles With Applications to Drug Shelf‐Life Estimation
by: Lingjiao Wang, et al.
Published: (2025)
by: Lingjiao Wang, et al.
Published: (2025)
Minimizing Upper Confidence Bounds: A Data-Driven Framework for Stochastic Programming
by: Liu, Shixin, et al.
Published: (2024)
by: Liu, Shixin, et al.
Published: (2024)
Online Optimization for Learning to Communicate over Time-Correlated Channels
by: Wu, Zheshun, et al.
Published: (2024)
by: Wu, Zheshun, et al.
Published: (2024)
Microwave‐to‐Optics Conversion Using Magnetostatic Modes and a Tunable Optical Cavity
by: Wei‐Jiang Wu, et al.
Published: (2024)
by: Wei‐Jiang Wu, et al.
Published: (2024)
Learning an Opponent-aware Anti-jamming Strategy via Online Convex Optimization
by: Liu, Liangqi, et al.
Published: (2026)
by: Liu, Liangqi, et al.
Published: (2026)
Deep Incomplete Multi-view Learning via Cyclic Permutation of VAEs
by: Gao, Xin, et al.
Published: (2025)
by: Gao, Xin, et al.
Published: (2025)
Optimized In‐Solution and Gas‐Phase Chemistry Enables High‐Efficiency Interactome Mapping by DSBSO‐Based Cross‐Linking Mass Spectrometry
by: Pin‐Lian Jiang, et al.
Published: (2025)
by: Pin‐Lian Jiang, et al.
Published: (2025)
Similar Items
-
Efficient Cross-Architecture Knowledge Transfer for Large-Scale Online User Response Prediction
by: Wu, Yucheng, et al.
Published: (2026) -
ICPO: Intrinsic Confidence-Driven Group Relative Preference Optimization for Efficient Reinforcement Learning
by: Wang, Jinpeng, et al.
Published: (2025) -
ROAD: Adaptive Data Mixing for Offline-to-Online Reinforcement Learning via Bi-Level Optimization
by: Yang, Letian, et al.
Published: (2026) -
FlexHB: a More Efficient and Flexible Framework for Hyperparameter Optimization
by: Zhang, Yang, et al.
Published: (2024) -
Revealing the Learning Dynamics of Long-Context Continual Pre-training
by: Liang, Yupu, et al.
Published: (2026)