Truth or Deceit? A Bayesian Decoding Game Enhances Consistency and Reliability
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Zhang, Weitong, Zang, Chengqi, Kainz, Bernhard |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Stability and Generalizability in SDE Diffusion Models with Measure-Preserving Dynamics
von: Zhang, Weitong, et al.
Veröffentlicht: (2024)
von: Zhang, Weitong, et al.
Veröffentlicht: (2024)
Wasserstein Equilibrium Decoding for Reliable Medical Visual Question Answering
von: Hagen, Luca, et al.
Veröffentlicht: (2026)
von: Hagen, Luca, et al.
Veröffentlicht: (2026)
Multi-Agent Reasoning for Cardiovascular Imaging Phenotype Analysis
von: Zhang, Weitong, et al.
Veröffentlicht: (2025)
von: Zhang, Weitong, et al.
Veröffentlicht: (2025)
Towards Effective MLLM Jailbreaking Through Balanced On-Topicness and OOD-Intensity
von: Li, Zuoou, et al.
Veröffentlicht: (2025)
von: Li, Zuoou, et al.
Veröffentlicht: (2025)
TruthEval: A Dataset to Evaluate LLM Truthfulness and Reliability
von: Khatun, Aisha, et al.
Veröffentlicht: (2024)
von: Khatun, Aisha, et al.
Veröffentlicht: (2024)
Resource-efficient Medical Image Analysis with Self-adapting Forward-Forward Networks
von: Müller, Johanna P., et al.
Veröffentlicht: (2024)
von: Müller, Johanna P., et al.
Veröffentlicht: (2024)
Graph Conditioned Diffusion for Controllable Histopathology Image Generation
von: Cechnicka, Sarah, et al.
Veröffentlicht: (2025)
von: Cechnicka, Sarah, et al.
Veröffentlicht: (2025)
Uncovering Hidden Subspaces in Video Diffusion Models Using Re-Identification
von: Dombrowski, Mischa, et al.
Veröffentlicht: (2024)
von: Dombrowski, Mischa, et al.
Veröffentlicht: (2024)
Real-Time Diffusion Policies for Games: Enhancing Consistency Policies with Q-Ensembles
von: Zhang, Ruoqi, et al.
Veröffentlicht: (2025)
von: Zhang, Ruoqi, et al.
Veröffentlicht: (2025)
A Dynamic Equilibrium Model for Automated Market Makers
von: Zang, Chengqi, et al.
Veröffentlicht: (2026)
von: Zang, Chengqi, et al.
Veröffentlicht: (2026)
Incentivizing Truthfulness and Collaborative Fairness in Bayesian Learning
von: Sim, Rachael Hwee Ling, et al.
Veröffentlicht: (2026)
von: Sim, Rachael Hwee Ling, et al.
Veröffentlicht: (2026)
Geometry-Aware Uncertainty Coresets for Robust Visual In-Context Learning in Histopathology
von: Erick, Franciskus Xaverius, et al.
Veröffentlicht: (2026)
von: Erick, Franciskus Xaverius, et al.
Veröffentlicht: (2026)
SH2: Self-Highlighted Hesitation Helps You Decode More Truthfully
von: Kai, Jushi, et al.
Veröffentlicht: (2024)
von: Kai, Jushi, et al.
Veröffentlicht: (2024)
Fact or Fiction? Can LLMs be Reliable Annotators for Political Truths?
von: Chatrath, Veronica, et al.
Veröffentlicht: (2024)
von: Chatrath, Veronica, et al.
Veröffentlicht: (2024)
Return-to-Go Is More Than a Number: Q-Guided Alignment for Return-Conditioned Supervised Learning
von: Yang, Yuxiao, et al.
Veröffentlicht: (2026)
von: Yang, Yuxiao, et al.
Veröffentlicht: (2026)
Consistent Paths Lead to Truth: Self-Rewarding Reinforcement Learning for LLM Reasoning
von: Zhang, Kongcheng, et al.
Veröffentlicht: (2025)
von: Zhang, Kongcheng, et al.
Veröffentlicht: (2025)
Towards Reliable Truth-Aligned Uncertainty Estimation in Large Language Models
von: Srey, Ponhvoan, et al.
Veröffentlicht: (2026)
von: Srey, Ponhvoan, et al.
Veröffentlicht: (2026)
CTFlow: Video-Inspired Latent Flow Matching for 3D CT Synthesis
von: Wang, Jiayi, et al.
Veröffentlicht: (2025)
von: Wang, Jiayi, et al.
Veröffentlicht: (2025)
TriAlign: Towards Universal Truth Consistency in Personalized LLM Alignment
von: Nguyen, Thi-Nhung, et al.
Veröffentlicht: (2026)
von: Nguyen, Thi-Nhung, et al.
Veröffentlicht: (2026)
Toward Reliable Scientific Hypothesis Generation: Evaluating Truthfulness and Hallucination in Large Language Models
von: Xiong, Guangzhi, et al.
Veröffentlicht: (2025)
von: Xiong, Guangzhi, et al.
Veröffentlicht: (2025)
Measuring and Aligning Abstraction in Vision-Language Models with Medical Taxonomies
von: Schaper, Ben, et al.
Veröffentlicht: (2026)
von: Schaper, Ben, et al.
Veröffentlicht: (2026)
Synthetic Concept Evolution Under the Axiom of Truth
von: Rubendall, C. Mike "WildFacts"
Veröffentlicht: (2025)
von: Rubendall, C. Mike "WildFacts"
Veröffentlicht: (2025)
OGLS-SD: On-Policy Self-Distillation with Outcome-Guided Logit Steering for LLM Reasoning
von: Yang, Yuxiao, et al.
Veröffentlicht: (2026)
von: Yang, Yuxiao, et al.
Veröffentlicht: (2026)
Exploring Criteria of Loss Reweighting to Enhance LLM Unlearning
von: Yang, Puning, et al.
Veröffentlicht: (2025)
von: Yang, Puning, et al.
Veröffentlicht: (2025)
STED and Consistency Scoring: A Framework for Evaluating LLM Structured Output Reliability
von: Wang, Guanghui, et al.
Veröffentlicht: (2025)
von: Wang, Guanghui, et al.
Veröffentlicht: (2025)
Enhancing Vision-Language Model Reliability with Uncertainty-Guided Dropout Decoding
von: Fang, Yixiong, et al.
Veröffentlicht: (2024)
von: Fang, Yixiong, et al.
Veröffentlicht: (2024)
Consistency as a Testable Property: Statistical Methods to Evaluate AI Agent Reliability
von: Raj, Harsh, et al.
Veröffentlicht: (2026)
von: Raj, Harsh, et al.
Veröffentlicht: (2026)
Incentivizing Truthful Language Models via Peer Elicitation Games
von: Chen, Baiting, et al.
Veröffentlicht: (2025)
von: Chen, Baiting, et al.
Veröffentlicht: (2025)
Uncertainty-Aware Reward-Free Exploration with General Function Approximation
von: Zhang, Junkai, et al.
Veröffentlicht: (2024)
von: Zhang, Junkai, et al.
Veröffentlicht: (2024)
Illusions of Confidence? Diagnosing LLM Truthfulness via Neighborhood Consistency
von: Xu, Haoming, et al.
Veröffentlicht: (2026)
von: Xu, Haoming, et al.
Veröffentlicht: (2026)
Scientific Knowledge-driven Decoding Constraints Improving the Reliability of LLMs
von: Ma, Maotian, et al.
Veröffentlicht: (2026)
von: Ma, Maotian, et al.
Veröffentlicht: (2026)
TruthStance: An Annotated Dataset of Conversations on Truth Social
von: Ameen, Fathima, et al.
Veröffentlicht: (2026)
von: Ameen, Fathima, et al.
Veröffentlicht: (2026)
A Collective Variational Principle Unifying Bayesian Inference, Game Theory, and Thermodynamics
von: Bouchaffra, Djamel, et al.
Veröffentlicht: (2026)
von: Bouchaffra, Djamel, et al.
Veröffentlicht: (2026)
Decoding Memories: An Efficient Pipeline for Self-Consistency Hallucination Detection
von: Gao, Weizhi, et al.
Veröffentlicht: (2025)
von: Gao, Weizhi, et al.
Veröffentlicht: (2025)
System Design for Maintaining Internal State Consistency in Long-Horizon Robotic Tabletop Games
von: Zhao, Guangyu, et al.
Veröffentlicht: (2026)
von: Zhao, Guangyu, et al.
Veröffentlicht: (2026)
How Reliable are LLMs as Knowledge Bases? Re-thinking Facutality and Consistency
von: Zheng, Danna, et al.
Veröffentlicht: (2024)
von: Zheng, Danna, et al.
Veröffentlicht: (2024)
Bayesian Optimization-based Search for Agent Control in Automated Game Testing
von: Celemin, Carlos
Veröffentlicht: (2025)
von: Celemin, Carlos
Veröffentlicht: (2025)
Compositional Consistency-Guided Decoding for Three-Way Logical Question Answering
von: Huang, Tianyi, et al.
Veröffentlicht: (2026)
von: Huang, Tianyi, et al.
Veröffentlicht: (2026)
Bayesian Neural Networks: A Min-Max Game Framework
von: Hong, Junping, et al.
Veröffentlicht: (2023)
von: Hong, Junping, et al.
Veröffentlicht: (2023)
TruthX: Alleviating Hallucinations by Editing Large Language Models in Truthful Space
von: Zhang, Shaolei, et al.
Veröffentlicht: (2024)
von: Zhang, Shaolei, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Stability and Generalizability in SDE Diffusion Models with Measure-Preserving Dynamics
von: Zhang, Weitong, et al.
Veröffentlicht: (2024) -
Wasserstein Equilibrium Decoding for Reliable Medical Visual Question Answering
von: Hagen, Luca, et al.
Veröffentlicht: (2026) -
Multi-Agent Reasoning for Cardiovascular Imaging Phenotype Analysis
von: Zhang, Weitong, et al.
Veröffentlicht: (2025) -
Towards Effective MLLM Jailbreaking Through Balanced On-Topicness and OOD-Intensity
von: Li, Zuoou, et al.
Veröffentlicht: (2025) -
TruthEval: A Dataset to Evaluate LLM Truthfulness and Reliability
von: Khatun, Aisha, et al.
Veröffentlicht: (2024)