Towards Reliable, Uncertainty-Aware Alignment
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Banerjee, Debangshu, Saha, Kintan, Gopalan, Aditya |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Towards Reliable Alignment: Uncertainty-aware RLHF
von: Banerjee, Debangshu, et al.
Veröffentlicht: (2024)
von: Banerjee, Debangshu, et al.
Veröffentlicht: (2024)
Bad Values but Good Behavior: Learning Highly Misspecified Bandits and MDPs
von: Banerjee, Debangshu, et al.
Veröffentlicht: (2023)
von: Banerjee, Debangshu, et al.
Veröffentlicht: (2023)
Reliable Policy Iteration: Performance Robustness Across Architecture and Environment Perturbations
von: Eshwar, S. R., et al.
Veröffentlicht: (2025)
von: Eshwar, S. R., et al.
Veröffentlicht: (2025)
Data Shifts Hurt CoT: A Theoretical Study
von: Yin, Lang, et al.
Veröffentlicht: (2025)
von: Yin, Lang, et al.
Veröffentlicht: (2025)
Why DPO is a Misspecified Estimator and How to Fix It
von: Gopalan, Aditya, et al.
Veröffentlicht: (2025)
von: Gopalan, Aditya, et al.
Veröffentlicht: (2025)
Shifting Uncertainty to Critical Moments: Towards Reliable Uncertainty Quantification for VLA Model
von: Tang, Yanchuan, et al.
Veröffentlicht: (2026)
von: Tang, Yanchuan, et al.
Veröffentlicht: (2026)
EARL: Entropy-Aware RL Alignment of LLMs for Reliable RTL Code Generation
von: Shi, Jiahe, et al.
Veröffentlicht: (2025)
von: Shi, Jiahe, et al.
Veröffentlicht: (2025)
FedUAF: Uncertainty-Aware Fusion with Reliability-Guided Aggregation for Multimodal Federated Sentiment Analysis
von: Zhu, Xianxun, et al.
Veröffentlicht: (2026)
von: Zhu, Xianxun, et al.
Veröffentlicht: (2026)
Towards Reliable Time Series Forecasting under Future Uncertainty: Ambiguity and Novelty Rejection Mechanisms
von: Feng, Ninghui, et al.
Veröffentlicht: (2025)
von: Feng, Ninghui, et al.
Veröffentlicht: (2025)
Conformal Feedback Alignment: Quantifying Answer-Level Reliability for Robust LLM Alignment
von: Chen, Tiejin, et al.
Veröffentlicht: (2026)
von: Chen, Tiejin, et al.
Veröffentlicht: (2026)
Towards Label-Free Biological Reasoning Synthetic Dataset Creation via Uncertainty Filtering
von: Stoisser, Josefa Lia, et al.
Veröffentlicht: (2025)
von: Stoisser, Josefa Lia, et al.
Veröffentlicht: (2025)
Why Uncertainty Calibration Matters for Reliable Perturbation-based Explanations
von: Decker, Thomas, et al.
Veröffentlicht: (2025)
von: Decker, Thomas, et al.
Veröffentlicht: (2025)
Reliable Trajectory Prediction and Uncertainty Quantification with Conditioned Diffusion Models
von: Neumeier, Marion, et al.
Veröffentlicht: (2024)
von: Neumeier, Marion, et al.
Veröffentlicht: (2024)
Towards Reliable Testing of Machine Unlearning
von: Mazhar, Anna, et al.
Veröffentlicht: (2026)
von: Mazhar, Anna, et al.
Veröffentlicht: (2026)
Calibrating LLM Judges: Linear Probes for Fast and Reliable Uncertainty Estimation
von: Radharapu, Bhaktipriya, et al.
Veröffentlicht: (2025)
von: Radharapu, Bhaktipriya, et al.
Veröffentlicht: (2025)
Tractable Uncertainty-Aware Meta-Learning
von: Park, Young-Jin, et al.
Veröffentlicht: (2022)
von: Park, Young-Jin, et al.
Veröffentlicht: (2022)
UCPO: Uncertainty-Aware Policy Optimization
von: Zeng, Xianzhou, et al.
Veröffentlicht: (2026)
von: Zeng, Xianzhou, et al.
Veröffentlicht: (2026)
CARE: Confounder-Aware Aggregation for Reliable LLM Evaluation
von: Zhao, Jitian, et al.
Veröffentlicht: (2026)
von: Zhao, Jitian, et al.
Veröffentlicht: (2026)
Extending Epistemic Uncertainty Beyond Parameters Would Assist in Designing Reliable LLMs
von: Nguyen-Hien, T. Duy, et al.
Veröffentlicht: (2025)
von: Nguyen-Hien, T. Duy, et al.
Veröffentlicht: (2025)
Beyond Fluency: Toward Reliable Trajectories in Agentic IR
von: Sinha, Anushree, et al.
Veröffentlicht: (2026)
von: Sinha, Anushree, et al.
Veröffentlicht: (2026)
An Uncertainty-Aware ED-LSTM for Probabilistic Suffix Prediction
von: Mustroph, Henryk, et al.
Veröffentlicht: (2025)
von: Mustroph, Henryk, et al.
Veröffentlicht: (2025)
Uncertainty Gating for Cost-Aware Explainable Artificial Intelligence
von: Mikriukov, Georgii, et al.
Veröffentlicht: (2026)
von: Mikriukov, Georgii, et al.
Veröffentlicht: (2026)
Uncertainty-Aware Decision Transformer for Stochastic Driving Environments
von: Li, Zenan, et al.
Veröffentlicht: (2023)
von: Li, Zenan, et al.
Veröffentlicht: (2023)
Bayesian-Symbolic Integration for Uncertainty-Aware Parking Prediction
von: Nezhadettehad, Alireza, et al.
Veröffentlicht: (2026)
von: Nezhadettehad, Alireza, et al.
Veröffentlicht: (2026)
The Promise of Analog Deep Learning: Recent Advances, Challenges and Opportunities
von: Datar, Aditya, et al.
Veröffentlicht: (2024)
von: Datar, Aditya, et al.
Veröffentlicht: (2024)
Towards Uncertainty Quantification in Generative Model Learning
von: Morales, Giorgio, et al.
Veröffentlicht: (2025)
von: Morales, Giorgio, et al.
Veröffentlicht: (2025)
Learning Graph Structures and Uncertainty for Accurate and Calibrated Time-series Forecasting
von: Kamarthi, Harshavardhan, et al.
Veröffentlicht: (2024)
von: Kamarthi, Harshavardhan, et al.
Veröffentlicht: (2024)
Towards a Learning Theory of Representation Alignment
von: Insulla, Francesco, et al.
Veröffentlicht: (2025)
von: Insulla, Francesco, et al.
Veröffentlicht: (2025)
Is Your Explanation Reliable: Confidence-Aware Explanation on Graph Neural Networks
von: Zhang, Jiaxing, et al.
Veröffentlicht: (2025)
von: Zhang, Jiaxing, et al.
Veröffentlicht: (2025)
Towards One Model for Classical Dimensionality Reduction: A Probabilistic Perspective on UMAP and t-SNE
von: Ravuri, Aditya, et al.
Veröffentlicht: (2024)
von: Ravuri, Aditya, et al.
Veröffentlicht: (2024)
Towards Reliable Evaluation of Adversarial Robustness for Spiking Neural Networks
von: Wang, Jihang, et al.
Veröffentlicht: (2025)
von: Wang, Jihang, et al.
Veröffentlicht: (2025)
Introducing Interval Neural Networks for Uncertainty-Aware System Identification
von: Ferah, Mehmet Ali, et al.
Veröffentlicht: (2025)
von: Ferah, Mehmet Ali, et al.
Veröffentlicht: (2025)
Informative Perturbation Selection for Uncertainty-Aware Post-hoc Explanations
von: Chugh, Sumedha, et al.
Veröffentlicht: (2026)
von: Chugh, Sumedha, et al.
Veröffentlicht: (2026)
Uncertainty-Aware Reward-Free Exploration with General Function Approximation
von: Zhang, Junkai, et al.
Veröffentlicht: (2024)
von: Zhang, Junkai, et al.
Veröffentlicht: (2024)
DARC: Disagreement-Aware Alignment via Risk-Constrained Decoding
von: Zou, Mingxi, et al.
Veröffentlicht: (2026)
von: Zou, Mingxi, et al.
Veröffentlicht: (2026)
Uncertainty-Aware Deep Learning Framework for Remaining Useful Life Prediction in Turbofan Engines with Learned Aleatoric Uncertainty
von: Sharma, Krishang
Veröffentlicht: (2025)
von: Sharma, Krishang
Veröffentlicht: (2025)
Physics-Guided Tiny-Mamba Transformer for Reliability-Aware Early Fault Warning
von: Li, Changyu, et al.
Veröffentlicht: (2026)
von: Li, Changyu, et al.
Veröffentlicht: (2026)
Towards Effective and Efficient Graph Alignment without Supervision
von: Chen, Songyang, et al.
Veröffentlicht: (2026)
von: Chen, Songyang, et al.
Veröffentlicht: (2026)
Deep Reinforcement Learning for Inventory Networks: Toward Reliable Policy Optimization
von: Alvo, Matias, et al.
Veröffentlicht: (2023)
von: Alvo, Matias, et al.
Veröffentlicht: (2023)
Entropy-Guided Loop: Achieving Reasoning through Uncertainty-Aware Generation
von: Correa, Andrew G. A., et al.
Veröffentlicht: (2025)
von: Correa, Andrew G. A., et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Towards Reliable Alignment: Uncertainty-aware RLHF
von: Banerjee, Debangshu, et al.
Veröffentlicht: (2024) -
Bad Values but Good Behavior: Learning Highly Misspecified Bandits and MDPs
von: Banerjee, Debangshu, et al.
Veröffentlicht: (2023) -
Reliable Policy Iteration: Performance Robustness Across Architecture and Environment Perturbations
von: Eshwar, S. R., et al.
Veröffentlicht: (2025) -
Data Shifts Hurt CoT: A Theoretical Study
von: Yin, Lang, et al.
Veröffentlicht: (2025) -
Why DPO is a Misspecified Estimator and How to Fix It
von: Gopalan, Aditya, et al.
Veröffentlicht: (2025)