Gespeichert in:
| Hauptverfasser: | Tran, Quan M., Huang, Zhuo, Zhang, Wenbin, Han, Bo, Yatani, Koji, Sugiyama, Masashi, Liu, Tongliang |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | https://arxiv.org/abs/2602.05810 |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
BrokenBind: Universal Modality Exploration beyond Dataset Boundaries
von: Huang, Zhuo, et al.
Veröffentlicht: (2026)
von: Huang, Zhuo, et al.
Veröffentlicht: (2026)
Towards Effective Evaluations and Comparisons for LLM Unlearning Methods
von: Wang, Qizhou, et al.
Veröffentlicht: (2024)
von: Wang, Qizhou, et al.
Veröffentlicht: (2024)
BadLabel: A Robust Perspective on Evaluating and Enhancing Label-noise Learning
von: Zhang, Jingfeng, et al.
Veröffentlicht: (2023)
von: Zhang, Jingfeng, et al.
Veröffentlicht: (2023)
Towards Scalable Oversight with Collaborative Multi-Agent Debate in Error Detection
von: Chen, Yongqiang, et al.
Veröffentlicht: (2025)
von: Chen, Yongqiang, et al.
Veröffentlicht: (2025)
Reinforcement Learning with Verifiable yet Noisy Rewards under Imperfect Verifiers
von: Cai, Xin-Qiang, et al.
Veröffentlicht: (2025)
von: Cai, Xin-Qiang, et al.
Veröffentlicht: (2025)
On the Thinking-Language Modeling Gap in Large Language Models
von: Liu, Chenxi, et al.
Veröffentlicht: (2025)
von: Liu, Chenxi, et al.
Veröffentlicht: (2025)
Exploring Criteria of Loss Reweighting to Enhance LLM Unlearning
von: Yang, Puning, et al.
Veröffentlicht: (2025)
von: Yang, Puning, et al.
Veröffentlicht: (2025)
Mind the Gap Between Prototypes and Images in Cross-domain Finetuning
von: Tian, Hongduan, et al.
Veröffentlicht: (2024)
von: Tian, Hongduan, et al.
Veröffentlicht: (2024)
Non-stationary Online Learning for Curved Losses: Improved Dynamic Regret via Mixability
von: Zhang, Yu-Jie, et al.
Veröffentlicht: (2025)
von: Zhang, Yu-Jie, et al.
Veröffentlicht: (2025)
Is Gradient Ascent Really Necessary? Memorize to Forget for Machine Unlearning
von: Huang, Zhuo, et al.
Veröffentlicht: (2026)
von: Huang, Zhuo, et al.
Veröffentlicht: (2026)
MeGU: Machine-Guided Unlearning with Target Feature Disentanglement
von: Wang, Haoyu, et al.
Veröffentlicht: (2026)
von: Wang, Haoyu, et al.
Veröffentlicht: (2026)
Enriching Disentanglement: From Logical Definitions to Quantitative Metrics
von: Zhang, Yivan, et al.
Veröffentlicht: (2023)
von: Zhang, Yivan, et al.
Veröffentlicht: (2023)
A Category-theoretical Meta-analysis of Definitions of Disentanglement
von: Zhang, Yivan, et al.
Veröffentlicht: (2023)
von: Zhang, Yivan, et al.
Veröffentlicht: (2023)
A Fast Algorithm for the Real-Valued Combinatorial Pure Exploration of Multi-Armed Bandit
von: Nakamura, Shintaro, et al.
Veröffentlicht: (2023)
von: Nakamura, Shintaro, et al.
Veröffentlicht: (2023)
Research as Resistance: Recognizing and Reconsidering HCI's Role in Technology Hype Cycles
von: Sramek, Zefan, et al.
Veröffentlicht: (2025)
von: Sramek, Zefan, et al.
Veröffentlicht: (2025)
What Is Preference Optimization Doing, and Why?
von: Wang, Yue, et al.
Veröffentlicht: (2025)
von: Wang, Yue, et al.
Veröffentlicht: (2025)
VEC-SBM: Optimal Community Detection with Vectorial Edges Covariates
von: Braun, Guillaume, et al.
Veröffentlicht: (2024)
von: Braun, Guillaume, et al.
Veröffentlicht: (2024)
Riemannian Langevin Dynamics: Strong Convergence of Geometric Euler-Maruyama Scheme
von: Zhan, Zhiyuan, et al.
Veröffentlicht: (2026)
von: Zhan, Zhiyuan, et al.
Veröffentlicht: (2026)
Multi-Player Approaches for Dueling Bandits
von: Raveh, Or, et al.
Veröffentlicht: (2024)
von: Raveh, Or, et al.
Veröffentlicht: (2024)
Label Distribution Learning with Biased Annotations by Learning Multi-Label Representation
von: Kou, Zhiqiang, et al.
Veröffentlicht: (2025)
von: Kou, Zhiqiang, et al.
Veröffentlicht: (2025)
Decoupling the Class Label and the Target Concept in Machine Unlearning
von: Zhu, Jianing, et al.
Veröffentlicht: (2024)
von: Zhu, Jianing, et al.
Veröffentlicht: (2024)
VI-CuRL: Stabilizing Verifier-Independent RL Reasoning via Confidence-Guided Variance Reduction
von: Cai, Xin-Qiang, et al.
Veröffentlicht: (2026)
von: Cai, Xin-Qiang, et al.
Veröffentlicht: (2026)
Parallel Simulation for Log-concave Sampling and Score-based Diffusion Models
von: Zhou, Huanjian, et al.
Veröffentlicht: (2024)
von: Zhou, Huanjian, et al.
Veröffentlicht: (2024)
Towards Understanding Valuable Preference Data for Large Language Model Alignment
von: Zhang, Zizhuo, et al.
Veröffentlicht: (2025)
von: Zhang, Zizhuo, et al.
Veröffentlicht: (2025)
Enhancing Sample Selection Against Label Noise by Cutting Mislabeled Easy Examples
von: Yuan, Suqin, et al.
Veröffentlicht: (2025)
von: Yuan, Suqin, et al.
Veröffentlicht: (2025)
On the Over-Memorization During Natural, Robust and Catastrophic Overfitting
von: Lin, Runqi, et al.
Veröffentlicht: (2023)
von: Lin, Runqi, et al.
Veröffentlicht: (2023)
Practical estimation of the optimal classification error with soft labels and calibration
von: Ushio, Ryota, et al.
Veröffentlicht: (2025)
von: Ushio, Ryota, et al.
Veröffentlicht: (2025)
The Survival Bandit Problem
von: Riou, Charles, et al.
Veröffentlicht: (2022)
von: Riou, Charles, et al.
Veröffentlicht: (2022)
Thompson Exploration with Best Challenger Rule in Best Arm Identification
von: Lee, Jongyeong, et al.
Veröffentlicht: (2023)
von: Lee, Jongyeong, et al.
Veröffentlicht: (2023)
From Debate to Equilibrium: Belief-Driven Multi-Agent LLM Reasoning via Bayesian Nash Equilibrium
von: Yi, Xie, et al.
Veröffentlicht: (2025)
von: Yi, Xie, et al.
Veröffentlicht: (2025)
Reinforcement Learning with Options and State Representation
von: Ghriss, Ayoub, et al.
Veröffentlicht: (2024)
von: Ghriss, Ayoub, et al.
Veröffentlicht: (2024)
FedImpro: Measuring and Improving Client Update in Federated Learning
von: Tang, Zhenheng, et al.
Veröffentlicht: (2024)
von: Tang, Zhenheng, et al.
Veröffentlicht: (2024)
What If the Input is Expanded in OOD Detection?
von: Zhang, Boxuan, et al.
Veröffentlicht: (2024)
von: Zhang, Boxuan, et al.
Veröffentlicht: (2024)
Offline Reinforcement Learning from Datasets with Structured Non-Stationarity
von: Ackermann, Johannes, et al.
Veröffentlicht: (2024)
von: Ackermann, Johannes, et al.
Veröffentlicht: (2024)
Learnability Gaps of Strategic Classification
von: Cohen, Lee, et al.
Veröffentlicht: (2024)
von: Cohen, Lee, et al.
Veröffentlicht: (2024)
COBRA: Contextual Bandit Algorithm for Ensuring Truthful Strategic Agents
von: Verma, Arun, et al.
Veröffentlicht: (2025)
von: Verma, Arun, et al.
Veröffentlicht: (2025)
Bridging the Gap between Chemical Reaction Pretraining and Conditional Molecule Generation with a Unified Model
von: Qiang, Bo, et al.
Veröffentlicht: (2023)
von: Qiang, Bo, et al.
Veröffentlicht: (2023)
Weak-to-Strong Diffusion with Reflection
von: Bai, Lichen, et al.
Veröffentlicht: (2025)
von: Bai, Lichen, et al.
Veröffentlicht: (2025)
Off-Policy Corrected Reward Modeling for Reinforcement Learning from Human Feedback
von: Ackermann, Johannes, et al.
Veröffentlicht: (2025)
von: Ackermann, Johannes, et al.
Veröffentlicht: (2025)
Towards Scalable Oversight via Partitioned Human Supervision
von: Yin, Ren, et al.
Veröffentlicht: (2025)
von: Yin, Ren, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
BrokenBind: Universal Modality Exploration beyond Dataset Boundaries
von: Huang, Zhuo, et al.
Veröffentlicht: (2026) -
Towards Effective Evaluations and Comparisons for LLM Unlearning Methods
von: Wang, Qizhou, et al.
Veröffentlicht: (2024) -
BadLabel: A Robust Perspective on Evaluating and Enhancing Label-noise Learning
von: Zhang, Jingfeng, et al.
Veröffentlicht: (2023) -
Towards Scalable Oversight with Collaborative Multi-Agent Debate in Error Detection
von: Chen, Yongqiang, et al.
Veröffentlicht: (2025) -
Reinforcement Learning with Verifiable yet Noisy Rewards under Imperfect Verifiers
von: Cai, Xin-Qiang, et al.
Veröffentlicht: (2025)