Salvato in:
| Autori principali: | Chen, Yongqiang, Niu, Gang, Cheng, James, Han, Bo, Sugiyama, Masashi |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | https://arxiv.org/abs/2510.20963 |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Towards Scalable Oversight via Partitioned Human Supervision
di: Yin, Ren, et al.
Pubblicazione: (2025)
di: Yin, Ren, et al.
Pubblicazione: (2025)
BrokenBind: Universal Modality Exploration beyond Dataset Boundaries
di: Huang, Zhuo, et al.
Pubblicazione: (2026)
di: Huang, Zhuo, et al.
Pubblicazione: (2026)
Knowledge Divergence and the Value of Debate for Scalable Oversight
di: Young, Robin
Pubblicazione: (2026)
di: Young, Robin
Pubblicazione: (2026)
What Is Preference Optimization Doing, and Why?
di: Wang, Yue, et al.
Pubblicazione: (2025)
di: Wang, Yue, et al.
Pubblicazione: (2025)
Decoupling the Class Label and the Target Concept in Machine Unlearning
di: Zhu, Jianing, et al.
Pubblicazione: (2024)
di: Zhu, Jianing, et al.
Pubblicazione: (2024)
Rethinking Consistent Multi-Label Classification Under Inexact Supervision
di: Wang, Wei, et al.
Pubblicazione: (2025)
di: Wang, Wei, et al.
Pubblicazione: (2025)
Dual-Decoupling Learning and Metric-Adaptive Thresholding for Semi-Supervised Multi-Label Learning
di: Xiao, Jia-Hao, et al.
Pubblicazione: (2024)
di: Xiao, Jia-Hao, et al.
Pubblicazione: (2024)
Robust Multi-View Learning via Representation Fusion of Sample-Level Attention and Alignment of Simulated Perturbation
di: Xu, Jie, et al.
Pubblicazione: (2025)
di: Xu, Jie, et al.
Pubblicazione: (2025)
Exploring Health Misinformation Detection with Multi-Agent Debate
di: Chen, Chih-Han, et al.
Pubblicazione: (2025)
di: Chen, Chih-Han, et al.
Pubblicazione: (2025)
Generating Chain-of-Thoughts with a Pairwise-Comparison Approach to Searching for the Most Promising Intermediate Thought
di: Zhang, Zhen-Yu, et al.
Pubblicazione: (2024)
di: Zhang, Zhen-Yu, et al.
Pubblicazione: (2024)
A Fast Algorithm for the Real-Valued Combinatorial Pure Exploration of Multi-Armed Bandit
di: Nakamura, Shintaro, et al.
Pubblicazione: (2023)
di: Nakamura, Shintaro, et al.
Pubblicazione: (2023)
Towards Effective Evaluations and Comparisons for LLM Unlearning Methods
di: Wang, Qizhou, et al.
Pubblicazione: (2024)
di: Wang, Qizhou, et al.
Pubblicazione: (2024)
Learning with Complementary Labels Revisited: The Selected-Completely-at-Random Setting Is More Practical
di: Wang, Wei, et al.
Pubblicazione: (2023)
di: Wang, Wei, et al.
Pubblicazione: (2023)
Multi-Player Approaches for Dueling Bandits
di: Raveh, Or, et al.
Pubblicazione: (2024)
di: Raveh, Or, et al.
Pubblicazione: (2024)
VEC-SBM: Optimal Community Detection with Vectorial Edges Covariates
di: Braun, Guillaume, et al.
Pubblicazione: (2024)
di: Braun, Guillaume, et al.
Pubblicazione: (2024)
In-context Demonstration Matters: On Prompt Optimization for Pseudo-Supervision Refinement
di: Zhang, Zhen-Yu, et al.
Pubblicazione: (2024)
di: Zhang, Zhen-Yu, et al.
Pubblicazione: (2024)
Balancing Similarity and Complementarity for Federated Learning
di: Yan, Kunda, et al.
Pubblicazione: (2024)
di: Yan, Kunda, et al.
Pubblicazione: (2024)
Multi-Agent Debate with Memory Masking
di: Tian, Hongduan, et al.
Pubblicazione: (2026)
di: Tian, Hongduan, et al.
Pubblicazione: (2026)
How Interpretable Are Interpretable Graph Neural Networks?
di: Chen, Yongqiang, et al.
Pubblicazione: (2024)
di: Chen, Yongqiang, et al.
Pubblicazione: (2024)
Towards Understanding Valuable Preference Data for Large Language Model Alignment
di: Zhang, Zizhuo, et al.
Pubblicazione: (2025)
di: Zhang, Zizhuo, et al.
Pubblicazione: (2025)
Accessible, Realistic, and Fair Evaluation of Positive-Unlabeled Learning Algorithms
di: Wang, Wei, et al.
Pubblicazione: (2025)
di: Wang, Wei, et al.
Pubblicazione: (2025)
From Debate to Equilibrium: Belief-Driven Multi-Agent LLM Reasoning via Bayesian Nash Equilibrium
di: Yi, Xie, et al.
Pubblicazione: (2025)
di: Yi, Xie, et al.
Pubblicazione: (2025)
Accurate Forgetting for Heterogeneous Federated Continual Learning
di: Wuerkaixi, Abudukelimu, et al.
Pubblicazione: (2025)
di: Wuerkaixi, Abudukelimu, et al.
Pubblicazione: (2025)
Multi-Label Knowledge Distillation
di: Yang, Penghui, et al.
Pubblicazione: (2023)
di: Yang, Penghui, et al.
Pubblicazione: (2023)
Bifrost: Steering Strategic Trajectories to Bridge Contextual Gaps for Self-Improving Agents
di: Tran, Quan M., et al.
Pubblicazione: (2026)
di: Tran, Quan M., et al.
Pubblicazione: (2026)
Counterfactual Reasoning for Multi-Label Image Classification via Patching-Based Training
di: Xie, Ming-Kun, et al.
Pubblicazione: (2024)
di: Xie, Ming-Kun, et al.
Pubblicazione: (2024)
Realistic Evaluation of Deep Partial-Label Learning Algorithms
di: Wang, Wei, et al.
Pubblicazione: (2025)
di: Wang, Wei, et al.
Pubblicazione: (2025)
Reinforcement Learning with Verifiable yet Noisy Rewards under Imperfect Verifiers
di: Cai, Xin-Qiang, et al.
Pubblicazione: (2025)
di: Cai, Xin-Qiang, et al.
Pubblicazione: (2025)
VerificAgent: Domain-Specific Memory Verification for Scalable Oversight of Aligned Computer-Use Agents
di: Nguyen, Thong Q., et al.
Pubblicazione: (2025)
di: Nguyen, Thong Q., et al.
Pubblicazione: (2025)
Scaling Laws For Scalable Oversight
di: Engels, Joshua, et al.
Pubblicazione: (2025)
di: Engels, Joshua, et al.
Pubblicazione: (2025)
A Complete Decomposition of KL Error using Refined Information and Mode Interaction Selection
di: Enouen, James, et al.
Pubblicazione: (2024)
di: Enouen, James, et al.
Pubblicazione: (2024)
Riemannian Langevin Dynamics: Strong Convergence of Geometric Euler-Maruyama Scheme
di: Zhan, Zhiyuan, et al.
Pubblicazione: (2026)
di: Zhan, Zhiyuan, et al.
Pubblicazione: (2026)
Enriching Disentanglement: From Logical Definitions to Quantitative Metrics
di: Zhang, Yivan, et al.
Pubblicazione: (2023)
di: Zhang, Yivan, et al.
Pubblicazione: (2023)
From Coefficients to Directions: Rethinking Model Merging with Directional Alignment
di: Chen, Zhikang, et al.
Pubblicazione: (2025)
di: Chen, Zhikang, et al.
Pubblicazione: (2025)
Steering LLMs via Scalable Interactive Oversight
di: Zhou, Enyu, et al.
Pubblicazione: (2026)
di: Zhou, Enyu, et al.
Pubblicazione: (2026)
Direct Distillation between Different Domains
di: Tang, Jialiang, et al.
Pubblicazione: (2024)
di: Tang, Jialiang, et al.
Pubblicazione: (2024)
VI-CuRL: Stabilizing Verifier-Independent RL Reasoning via Confidence-Guided Variance Reduction
di: Cai, Xin-Qiang, et al.
Pubblicazione: (2026)
di: Cai, Xin-Qiang, et al.
Pubblicazione: (2026)
Locally Estimated Global Perturbations are Better than Local Perturbations for Federated Sharpness-aware Minimization
di: Fan, Ziqing, et al.
Pubblicazione: (2024)
di: Fan, Ziqing, et al.
Pubblicazione: (2024)
Practical estimation of the optimal classification error with soft labels and calibration
di: Ushio, Ryota, et al.
Pubblicazione: (2025)
di: Ushio, Ryota, et al.
Pubblicazione: (2025)
The Survival Bandit Problem
di: Riou, Charles, et al.
Pubblicazione: (2022)
di: Riou, Charles, et al.
Pubblicazione: (2022)
Documenti analoghi
-
Towards Scalable Oversight via Partitioned Human Supervision
di: Yin, Ren, et al.
Pubblicazione: (2025) -
BrokenBind: Universal Modality Exploration beyond Dataset Boundaries
di: Huang, Zhuo, et al.
Pubblicazione: (2026) -
Knowledge Divergence and the Value of Debate for Scalable Oversight
di: Young, Robin
Pubblicazione: (2026) -
What Is Preference Optimization Doing, and Why?
di: Wang, Yue, et al.
Pubblicazione: (2025) -
Decoupling the Class Label and the Target Concept in Machine Unlearning
di: Zhu, Jianing, et al.
Pubblicazione: (2024)