Out-of-Distribution Learning with Human Feedback
Fuente:
arXiv
Guardado en:
| Autores principales: | Bai, Haoyue, Du, Xuefeng, Rainey, Katie, Parameswaran, Shibin, Li, Yixuan |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Feed Two Birds with One Scone: Exploiting Wild Data for Both Out-of-Distribution Generalization and Detection
por: Bai, Haoyue, et al.
Publicado: (2023)
por: Bai, Haoyue, et al.
Publicado: (2023)
HYPO: Hyperspherical Out-of-Distribution Generalization
por: Bai, Haoyue, et al.
Publicado: (2024)
por: Bai, Haoyue, et al.
Publicado: (2024)
When and How Does In-Distribution Label Help Out-of-Distribution Detection?
por: Du, Xuefeng, et al.
Publicado: (2024)
por: Du, Xuefeng, et al.
Publicado: (2024)
AHA: Human-Assisted Out-of-Distribution Generalization and Detection
por: Bai, Haoyue, et al.
Publicado: (2024)
por: Bai, Haoyue, et al.
Publicado: (2024)
How Does Unlabeled Data Provably Help Out-of-Distribution Detection?
por: Du, Xuefeng, et al.
Publicado: (2024)
por: Du, Xuefeng, et al.
Publicado: (2024)
Towards Robust Out-of-Distribution Generalization: Data Augmentation and Neural Architecture Search Approaches
por: Bai, Haoyue
Publicado: (2024)
por: Bai, Haoyue
Publicado: (2024)
Understanding the Learning Dynamics of Alignment with Human Feedback
por: Im, Shawn, et al.
Publicado: (2024)
por: Im, Shawn, et al.
Publicado: (2024)
Corruption Robust Offline Reinforcement Learning with Human Feedback
por: Mandal, Debmalya, et al.
Publicado: (2024)
por: Mandal, Debmalya, et al.
Publicado: (2024)
Corruption-robust Offline Multi-agent Reinforcement Learning From Human Feedback
por: Nika, Andi, et al.
Publicado: (2026)
por: Nika, Andi, et al.
Publicado: (2026)
Taming False Positives in Out-of-Distribution Detection with Human Feedback
por: Vishwakarma, Harit, et al.
Publicado: (2024)
por: Vishwakarma, Harit, et al.
Publicado: (2024)
HaloScope: Harnessing Unlabeled LLM Generations for Hallucination Detection
por: Du, Xuefeng, et al.
Publicado: (2024)
por: Du, Xuefeng, et al.
Publicado: (2024)
Adaptive Scoring and Thresholding with Human Feedback for Robust Out-of-Distribution Detection
por: Yamada, Daisuke, et al.
Publicado: (2025)
por: Yamada, Daisuke, et al.
Publicado: (2025)
Foundations of Unknown-aware Machine Learning
por: Du, Xuefeng
Publicado: (2025)
por: Du, Xuefeng
Publicado: (2025)
Learning from Imperfect Human Feedback: a Tale from Corruption-Robust Dueling
por: Cheng, Yuwei, et al.
Publicado: (2024)
por: Cheng, Yuwei, et al.
Publicado: (2024)
Understanding Multimodal LLMs Under Distribution Shifts: An Information-Theoretic Approach
por: Oh, Changdae, et al.
Publicado: (2025)
por: Oh, Changdae, et al.
Publicado: (2025)
OpenOOD v1.5: Enhanced Benchmark for Out-of-Distribution Detection
por: Zhang, Jingyang, et al.
Publicado: (2023)
por: Zhang, Jingyang, et al.
Publicado: (2023)
Latent space analysis and generalization to out-of-distribution data
por: Rainey, Katie, et al.
Publicado: (2025)
por: Rainey, Katie, et al.
Publicado: (2025)
ConjNorm: Tractable Density Estimation for Out-of-Distribution Detection
por: Peng, Bo, et al.
Publicado: (2024)
por: Peng, Bo, et al.
Publicado: (2024)
Generalized Out-of-Distribution Detection: A Survey
por: Yang, Jingkang, et al.
Publicado: (2021)
por: Yang, Jingkang, et al.
Publicado: (2021)
How Does Fine-Tuning Impact Out-of-Distribution Detection for Vision-Language Models?
por: Ming, Yifei, et al.
Publicado: (2023)
por: Ming, Yifei, et al.
Publicado: (2023)
Bridging the Domain Gap in Equation Distillation with Reinforcement Feedback
por: Ying, Wangyang, et al.
Publicado: (2025)
por: Ying, Wangyang, et al.
Publicado: (2025)
Representation Learning on Out of Distribution in Tabular Data
por: Ginanjar, Achmad, et al.
Publicado: (2025)
por: Ginanjar, Achmad, et al.
Publicado: (2025)
Unknown Aware AI-Generated Content Attribution
por: Thieu, Ellie, et al.
Publicado: (2026)
por: Thieu, Ellie, et al.
Publicado: (2026)
Distributionally Robust Reinforcement Learning with Human Feedback
por: Mandal, Debmalya, et al.
Publicado: (2025)
por: Mandal, Debmalya, et al.
Publicado: (2025)
On the Out-of-Distribution Generalization of Self-Supervised Learning
por: Qiang, Wenwen, et al.
Publicado: (2025)
por: Qiang, Wenwen, et al.
Publicado: (2025)
Distributional Equivalence in Linear Non-Gaussian Latent-Variable Cyclic Causal Models: Characterization and Learning
por: Dai, Haoyue, et al.
Publicado: (2026)
por: Dai, Haoyue, et al.
Publicado: (2026)
Continual Contrastive Learning on Tabular Data with Out of Distribution
por: Ginanjar, Achmad, et al.
Publicado: (2025)
por: Ginanjar, Achmad, et al.
Publicado: (2025)
The Risk of Federated Learning to Skew Fine-Tuning Features and Underperform Out-of-Distribution Robustness
por: Du, Mengyao, et al.
Publicado: (2024)
por: Du, Mengyao, et al.
Publicado: (2024)
Policy Teaching via Data Poisoning in Learning from Human Preferences
por: Nika, Andi, et al.
Publicado: (2025)
por: Nika, Andi, et al.
Publicado: (2025)
Extragradient Preference Optimization (EGPO): Beyond Last-Iterate Convergence for Nash Learning from Human Feedback
por: Zhou, Runlong, et al.
Publicado: (2025)
por: Zhou, Runlong, et al.
Publicado: (2025)
Reward Model Learning vs. Direct Policy Optimization: A Comparative Analysis of Learning from Human Preferences
por: Nika, Andi, et al.
Publicado: (2024)
por: Nika, Andi, et al.
Publicado: (2024)
Reinforcement Learning from Human Feedback
por: Lambert, Nathan
Publicado: (2025)
por: Lambert, Nathan
Publicado: (2025)
Informativeness of Reward Functions in Reinforcement Learning
por: Devidze, Rati, et al.
Publicado: (2024)
por: Devidze, Rati, et al.
Publicado: (2024)
Optimal Ridge Regularization for Out-of-Distribution Prediction
por: Patil, Pratik, et al.
Publicado: (2024)
por: Patil, Pratik, et al.
Publicado: (2024)
GOLD: Graph Out-of-Distribution Detection via Implicit Adversarial Latent Generation
por: Wang, Danny, et al.
Publicado: (2025)
por: Wang, Danny, et al.
Publicado: (2025)
CGRL: Causal-Guided Representation Learning for Graph Out-of-Distribution Generalization
por: Lu, Bowen, et al.
Publicado: (2026)
por: Lu, Bowen, et al.
Publicado: (2026)
Incremental Causal Graph Learning for Online Cyberattack Detection in Cyber-Physical Infrastructures
por: Malarkkan, Arun Vignesh, et al.
Publicado: (2025)
por: Malarkkan, Arun Vignesh, et al.
Publicado: (2025)
Robust Reinforcement Learning from Corrupted Human Feedback
por: Bukharin, Alexander, et al.
Publicado: (2024)
por: Bukharin, Alexander, et al.
Publicado: (2024)
Where's the liability in the Generative Era? Recovery-based Black-Box Detection of AI-Generated Content
por: Bai, Haoyue, et al.
Publicado: (2025)
por: Bai, Haoyue, et al.
Publicado: (2025)
Generalizing Graph Neural Networks on Out-Of-Distribution Graphs
por: Fan, Shaohua, et al.
Publicado: (2021)
por: Fan, Shaohua, et al.
Publicado: (2021)
Ejemplares similares
-
Feed Two Birds with One Scone: Exploiting Wild Data for Both Out-of-Distribution Generalization and Detection
por: Bai, Haoyue, et al.
Publicado: (2023) -
HYPO: Hyperspherical Out-of-Distribution Generalization
por: Bai, Haoyue, et al.
Publicado: (2024) -
When and How Does In-Distribution Label Help Out-of-Distribution Detection?
por: Du, Xuefeng, et al.
Publicado: (2024) -
AHA: Human-Assisted Out-of-Distribution Generalization and Detection
por: Bai, Haoyue, et al.
Publicado: (2024) -
How Does Unlabeled Data Provably Help Out-of-Distribution Detection?
por: Du, Xuefeng, et al.
Publicado: (2024)