Learning Robust Diffusion Models from Imprecise Supervision
Fuente:
arXiv
Salvato in:
| Autori principali: | Wu, Dong-Dong, Cui, Jiacheng, Wang, Wei, Shen, Zhiqiang, Sugiyama, Masashi |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Imprecise Label Learning: A Unified Framework for Learning with Various Imprecise Label Configurations
di: Chen, Hao, et al.
Pubblicazione: (2023)
di: Chen, Hao, et al.
Pubblicazione: (2023)
A Frustratingly Simple Yet Highly Effective Attack Baseline: Over 90% Success Rate Against the Strong Black-box Models of GPT-4.5/4o/o1
di: Li, Zhaoyi, et al.
Pubblicazione: (2025)
di: Li, Zhaoyi, et al.
Pubblicazione: (2025)
Towards Scalable Oversight via Partitioned Human Supervision
di: Yin, Ren, et al.
Pubblicazione: (2025)
di: Yin, Ren, et al.
Pubblicazione: (2025)
Offline Reinforcement Learning from Datasets with Structured Non-Stationarity
di: Ackermann, Johannes, et al.
Pubblicazione: (2024)
di: Ackermann, Johannes, et al.
Pubblicazione: (2024)
A General Framework for Learning from Weak Supervision
di: Chen, Hao, et al.
Pubblicazione: (2024)
di: Chen, Hao, et al.
Pubblicazione: (2024)
Off-Policy Corrected Reward Modeling for Reinforcement Learning from Human Feedback
di: Ackermann, Johannes, et al.
Pubblicazione: (2025)
di: Ackermann, Johannes, et al.
Pubblicazione: (2025)
On the Overlooked Pitfalls of Weight Decay and How to Mitigate Them: A Gradient-Norm Perspective
di: Xie, Zeke, et al.
Pubblicazione: (2020)
di: Xie, Zeke, et al.
Pubblicazione: (2020)
On Symmetric Losses for Robust Policy Optimization with Noisy Preferences
di: Nishimori, Soichiro, et al.
Pubblicazione: (2025)
di: Nishimori, Soichiro, et al.
Pubblicazione: (2025)
VI-CuRL: Stabilizing Verifier-Independent RL Reasoning via Confidence-Guided Variance Reduction
di: Cai, Xin-Qiang, et al.
Pubblicazione: (2026)
di: Cai, Xin-Qiang, et al.
Pubblicazione: (2026)
Distributionally Robust Statistical Verification with Imprecise Neural Networks
di: Dutta, Souradeep, et al.
Pubblicazione: (2023)
di: Dutta, Souradeep, et al.
Pubblicazione: (2023)
Impact of Noisy Supervision in Foundation Model Learning
di: Chen, Hao, et al.
Pubblicazione: (2024)
di: Chen, Hao, et al.
Pubblicazione: (2024)
Reinforcement Learning with Verifiable yet Noisy Rewards under Imperfect Verifiers
di: Cai, Xin-Qiang, et al.
Pubblicazione: (2025)
di: Cai, Xin-Qiang, et al.
Pubblicazione: (2025)
VLBM: Variational Latent Basis Modeling for OOD Robust Multivariate Time Series Forecasting
di: Zhang, Xudong, et al.
Pubblicazione: (2026)
di: Zhang, Xudong, et al.
Pubblicazione: (2026)
Next-Gen CAPTCHAs: Leveraging the Cognitive Gap for Scalable and Diverse GUI-Agent Defense
di: Liu, Jiacheng, et al.
Pubblicazione: (2026)
di: Liu, Jiacheng, et al.
Pubblicazione: (2026)
Gradient Regularization Prevents Reward Hacking in Reinforcement Learning from Human Feedback and Verifiable Rewards
di: Ackermann, Johannes, et al.
Pubblicazione: (2026)
di: Ackermann, Johannes, et al.
Pubblicazione: (2026)
Mitigating Reward Hacking in RLHF via Advantage Sign Robustness
di: Ono, Shinnosuke, et al.
Pubblicazione: (2026)
di: Ono, Shinnosuke, et al.
Pubblicazione: (2026)
LLMSurgeon: Diagnosing Data Mixture of Large Language Models
di: Luo, Yaxin, et al.
Pubblicazione: (2026)
di: Luo, Yaxin, et al.
Pubblicazione: (2026)
From Coefficients to Directions: Rethinking Model Merging with Directional Alignment
di: Chen, Zhikang, et al.
Pubblicazione: (2025)
di: Chen, Zhikang, et al.
Pubblicazione: (2025)
Robust Multi-View Learning via Representation Fusion of Sample-Level Attention and Alignment of Simulated Perturbation
di: Xu, Jie, et al.
Pubblicazione: (2025)
di: Xu, Jie, et al.
Pubblicazione: (2025)
Accurate Forgetting for Heterogeneous Federated Continual Learning
di: Wuerkaixi, Abudukelimu, et al.
Pubblicazione: (2025)
di: Wuerkaixi, Abudukelimu, et al.
Pubblicazione: (2025)
Mahjax: A GPU-Accelerated Mahjong Simulator for Reinforcement Learning in JAX
di: Nishimori, Soichiro, et al.
Pubblicazione: (2026)
di: Nishimori, Soichiro, et al.
Pubblicazione: (2026)
AllMatch: Exploiting All Unlabeled Data for Semi-Supervised Learning
di: Wu, Zhiyu, et al.
Pubblicazione: (2024)
di: Wu, Zhiyu, et al.
Pubblicazione: (2024)
Action-Agnostic Point-Level Supervision for Temporal Action Detection
di: Yoshida, Shuhei M., et al.
Pubblicazione: (2024)
di: Yoshida, Shuhei M., et al.
Pubblicazione: (2024)
Learning Multimodal Energy-Based Model with Multimodal Variational Auto-Encoder via MCMC Revision
di: Cui, Jiali, et al.
Pubblicazione: (2026)
di: Cui, Jiali, et al.
Pubblicazione: (2026)
Stabilizing Reinforcement Learning for Diffusion Language Models
di: Zhong, Jianyuan, et al.
Pubblicazione: (2026)
di: Zhong, Jianyuan, et al.
Pubblicazione: (2026)
Causal Graph Learning via Distributional Invariance of Cause-Effect Relationship
di: Nguyen, Nang Hung, et al.
Pubblicazione: (2026)
di: Nguyen, Nang Hung, et al.
Pubblicazione: (2026)
Fast and Scalable Analytical Diffusion
di: Shang, Xinyi, et al.
Pubblicazione: (2026)
di: Shang, Xinyi, et al.
Pubblicazione: (2026)
Sharpness-Aware Black-Box Optimization
di: Ye, Feiyang, et al.
Pubblicazione: (2024)
di: Ye, Feiyang, et al.
Pubblicazione: (2024)
Goal-Conditioned Supervised Learning for LLM Fine-Tuning
di: Li, Shijun, et al.
Pubblicazione: (2026)
di: Li, Shijun, et al.
Pubblicazione: (2026)
A Survey on Diffusion Language Models
di: Li, Tianyi, et al.
Pubblicazione: (2025)
di: Li, Tianyi, et al.
Pubblicazione: (2025)
WISTERIA: Learning Clinical Representations from Noisy Supervision via Multi-View Consistency in Electronic Health Records
di: Dong, Ruan, et al.
Pubblicazione: (2026)
di: Dong, Ruan, et al.
Pubblicazione: (2026)
Diffusion Model-Based Data Synthesis Aided Federated Semi-Supervised Learning
di: Wang, Zhongwei, et al.
Pubblicazione: (2025)
di: Wang, Zhongwei, et al.
Pubblicazione: (2025)
Task-Distributionally Robust Data-Free Meta-Learning
di: Hu, Zixuan, et al.
Pubblicazione: (2023)
di: Hu, Zixuan, et al.
Pubblicazione: (2023)
Can LLMs Learn to Reason Robustly under Noisy Supervision?
di: Yang, Shenzhi, et al.
Pubblicazione: (2026)
di: Yang, Shenzhi, et al.
Pubblicazione: (2026)
Predictive Learning in Energy-based Models with Attractor Structures
di: Dong, Xingsi, et al.
Pubblicazione: (2025)
di: Dong, Xingsi, et al.
Pubblicazione: (2025)
How Do Diffusion Models Improve Adversarial Robustness?
di: Yuezhang, Liu, et al.
Pubblicazione: (2025)
di: Yuezhang, Liu, et al.
Pubblicazione: (2025)
PhysioME: A Robust Multimodal Self-Supervised Framework for Physiological Signals with Missing Modalities
di: Lee, Cheol-Hui, et al.
Pubblicazione: (2025)
di: Lee, Cheol-Hui, et al.
Pubblicazione: (2025)
USE: Uncertainty Structure Estimation for Robust Semi-Supervised Learning
di: Chen, Tsao-Lun, et al.
Pubblicazione: (2026)
di: Chen, Tsao-Lun, et al.
Pubblicazione: (2026)
Robust Semi-Supervised Learning in Open Environments
di: Guo, Lan-Zhe, et al.
Pubblicazione: (2024)
di: Guo, Lan-Zhe, et al.
Pubblicazione: (2024)
Maximum Entropy Reinforcement Learning with Diffusion Policy
di: Dong, Xiaoyi, et al.
Pubblicazione: (2025)
di: Dong, Xiaoyi, et al.
Pubblicazione: (2025)
Documenti analoghi
-
Imprecise Label Learning: A Unified Framework for Learning with Various Imprecise Label Configurations
di: Chen, Hao, et al.
Pubblicazione: (2023) -
A Frustratingly Simple Yet Highly Effective Attack Baseline: Over 90% Success Rate Against the Strong Black-box Models of GPT-4.5/4o/o1
di: Li, Zhaoyi, et al.
Pubblicazione: (2025) -
Towards Scalable Oversight via Partitioned Human Supervision
di: Yin, Ren, et al.
Pubblicazione: (2025) -
Offline Reinforcement Learning from Datasets with Structured Non-Stationarity
di: Ackermann, Johannes, et al.
Pubblicazione: (2024) -
A General Framework for Learning from Weak Supervision
di: Chen, Hao, et al.
Pubblicazione: (2024)