Increasing LLM response trustworthiness using voting ensembles
Fuente:
arXiv
Salvato in:
| Autori principali: | Nair-Kanneganti, Aparna, Chan, Trevor J., Goldfinger, Shir, Mackay, Emily, Anthony, Brian, Pouch, Alison |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Faithfulness metric fusion: Improving the evaluation of LLM trustworthiness across domains
di: Malin, Ben, et al.
Pubblicazione: (2025)
di: Malin, Ben, et al.
Pubblicazione: (2025)
Learning the Domain Specific Inverse NUFFT for Accelerated Spiral MRI using Diffusion Models
di: Chan, Trevor J., et al.
Pubblicazione: (2024)
di: Chan, Trevor J., et al.
Pubblicazione: (2024)
Generative AI voting: fair collective choice is resilient to LLM biases and inconsistencies
di: Majumdar, Srijoni, et al.
Pubblicazione: (2024)
di: Majumdar, Srijoni, et al.
Pubblicazione: (2024)
Developing trustworthy AI applications with foundation models
di: Mock, Michael, et al.
Pubblicazione: (2024)
di: Mock, Michael, et al.
Pubblicazione: (2024)
GENEOnet: Statistical analysis supporting explainability and trustworthiness
di: Bocchi, Giovanni, et al.
Pubblicazione: (2025)
di: Bocchi, Giovanni, et al.
Pubblicazione: (2025)
Follow Your Heart: Landmark-Guided Transducer Pose Scoring for Point-of-Care Echocardiography
di: Guo, Zaiyang, et al.
Pubblicazione: (2026)
di: Guo, Zaiyang, et al.
Pubblicazione: (2026)
Adaptive Composition of Machine Learning as a Service (MLaaS) for IoT Environments
di: Kanneganti, Deepak, et al.
Pubblicazione: (2025)
di: Kanneganti, Deepak, et al.
Pubblicazione: (2025)
Translating Dietary Standards into Healthy Meals with Minimal Substitutions
di: Chan, Trevor, et al.
Pubblicazione: (2026)
di: Chan, Trevor, et al.
Pubblicazione: (2026)
Standards for trustworthy AI in the European Union: technical rationale, structural challenges, and an implementation path
di: Bisconti, Piercosma, et al.
Pubblicazione: (2026)
di: Bisconti, Piercosma, et al.
Pubblicazione: (2026)
The METRIC-framework for assessing data quality for trustworthy AI in medicine: a systematic review
di: Schwabe, Daniel, et al.
Pubblicazione: (2024)
di: Schwabe, Daniel, et al.
Pubblicazione: (2024)
Humans learn to prefer trustworthy AI over human partners
di: Jiang, Yaomin, et al.
Pubblicazione: (2025)
di: Jiang, Yaomin, et al.
Pubblicazione: (2025)
Increasing AI Explainability by LLM Driven Standard Processes
di: Jansen, Marc, et al.
Pubblicazione: (2025)
di: Jansen, Marc, et al.
Pubblicazione: (2025)
Tell me the truth: A system to measure the trustworthiness of Large Language Models
di: Lipizzi, Carlo
Pubblicazione: (2024)
di: Lipizzi, Carlo
Pubblicazione: (2024)
Planning in the Dark: LLM-Symbolic Planning Pipeline without Experts
di: Huang, Sukai, et al.
Pubblicazione: (2024)
di: Huang, Sukai, et al.
Pubblicazione: (2024)
A feature-stable and explainable machine learning framework for trustworthy decision-making under incomplete clinical data
di: Andrys-Olek, Justyna, et al.
Pubblicazione: (2026)
di: Andrys-Olek, Justyna, et al.
Pubblicazione: (2026)
Presentations are not always linear! GNN meets LLM for Document-to-Presentation Transformation with Attribution
di: Maheshwari, Himanshu, et al.
Pubblicazione: (2024)
di: Maheshwari, Himanshu, et al.
Pubblicazione: (2024)
LLM-based Agent Simulation for Maternal Health Interventions: Uncertainty Estimation and Decision-focused Evaluation
di: Martinson, Sarah, et al.
Pubblicazione: (2025)
di: Martinson, Sarah, et al.
Pubblicazione: (2025)
ALIGN: Prompt-based Attribute Alignment for Reliable, Responsible, and Personalized LLM-based Decision-Making
di: Ravichandran, Bharadwaj, et al.
Pubblicazione: (2025)
di: Ravichandran, Bharadwaj, et al.
Pubblicazione: (2025)
Correlated Mutations for Integer Programming
di: Shir, Ofer M., et al.
Pubblicazione: (2025)
di: Shir, Ofer M., et al.
Pubblicazione: (2025)
Addressing Unboundedness in Quadratically-Constrained Mixed-Integer Problems
di: Zepko, Guy, et al.
Pubblicazione: (2024)
di: Zepko, Guy, et al.
Pubblicazione: (2024)
Evaluating the Effectiveness of Index-Based Treatment Allocation
di: Boehmer, Niclas, et al.
Pubblicazione: (2024)
di: Boehmer, Niclas, et al.
Pubblicazione: (2024)
Optimal bounds for dissatisfaction in perpetual voting
di: Kozachinskiy, Alexander, et al.
Pubblicazione: (2024)
di: Kozachinskiy, Alexander, et al.
Pubblicazione: (2024)
Ensuring trustworthy and ethical behaviour in intelligent logical agents
di: Costantini, Stefania
Pubblicazione: (2024)
di: Costantini, Stefania
Pubblicazione: (2024)
Automated Prediction of Breast Cancer Response to Neoadjuvant Chemotherapy from DWI Data
di: Nitzan, Shir, et al.
Pubblicazione: (2024)
di: Nitzan, Shir, et al.
Pubblicazione: (2024)
Bullying the Machine: How Personas Increase LLM Vulnerability
di: Xu, Ziwei, et al.
Pubblicazione: (2025)
di: Xu, Ziwei, et al.
Pubblicazione: (2025)
Beyond Listenership: AI-Predicted Interventions Drive Improvements in Maternal Health Behaviours
di: Dasgupta, Arpan, et al.
Pubblicazione: (2025)
di: Dasgupta, Arpan, et al.
Pubblicazione: (2025)
EXGnet: a single-lead explainable-AI guided multiresolution network with train-only quantitative features for trustworthy ECG arrhythmia classification
di: Showrav, Tushar Talukder, et al.
Pubblicazione: (2025)
di: Showrav, Tushar Talukder, et al.
Pubblicazione: (2025)
Chasing Progress, Not Perfection: Revisiting Strategies for End-to-End LLM Plan Generation
di: Huang, Sukai, et al.
Pubblicazione: (2024)
di: Huang, Sukai, et al.
Pubblicazione: (2024)
C-voting: Confidence-Based Test-Time Voting without Explicit Energy Functions
di: Kubo, Kenji, et al.
Pubblicazione: (2026)
di: Kubo, Kenji, et al.
Pubblicazione: (2026)
Region-wise stacking ensembles for estimating brain-age using MRI
di: Antonopoulos, Georgios, et al.
Pubblicazione: (2025)
di: Antonopoulos, Georgios, et al.
Pubblicazione: (2025)
CDQuant: Greedy Coordinate Descent for Accurate LLM Quantization
di: Nair, Pranav Ajit, et al.
Pubblicazione: (2024)
di: Nair, Pranav Ajit, et al.
Pubblicazione: (2024)
Response-Aware User Memory Selection for LLM Personalization
di: Fisher, Jillian, et al.
Pubblicazione: (2026)
di: Fisher, Jillian, et al.
Pubblicazione: (2026)
A Generative AI Technique for Synthesizing a Digital Twin for U.S. Residential Solar Adoption and Generation
di: Kishore, Aparna, et al.
Pubblicazione: (2024)
di: Kishore, Aparna, et al.
Pubblicazione: (2024)
The Price of Agreement: Measuring LLM Sycophancy in Agentic Financial Applications
di: Zhao, Zhenyu, et al.
Pubblicazione: (2026)
di: Zhao, Zhenyu, et al.
Pubblicazione: (2026)
Encrypted Prompt: Securing LLM Applications Against Unauthorized Actions
di: Chan, Shih-Han
Pubblicazione: (2025)
di: Chan, Shih-Han
Pubblicazione: (2025)
STARFISH: faST Accuracy Recovery in pruned networks From Internal State Healing
di: Maon, Shir, et al.
Pubblicazione: (2026)
di: Maon, Shir, et al.
Pubblicazione: (2026)
Increasing LLM Coding Capabilities through Diverse Synthetic Coding Tasks
di: Abed, Amal, et al.
Pubblicazione: (2025)
di: Abed, Amal, et al.
Pubblicazione: (2025)
Physical foundations for trustworthy medical imaging: a review for artificial intelligence researchers
di: Cobo, Miriam, et al.
Pubblicazione: (2025)
di: Cobo, Miriam, et al.
Pubblicazione: (2025)
Flow-of-Options: Diversified and Improved LLM Reasoning by Thinking Through Options
di: Nair, Lakshmi, et al.
Pubblicazione: (2025)
di: Nair, Lakshmi, et al.
Pubblicazione: (2025)
Real-Time Gait Adaptation for Quadrupeds using Model Predictive Control and Reinforcement Learning
di: Kotecha, Prakrut, et al.
Pubblicazione: (2025)
di: Kotecha, Prakrut, et al.
Pubblicazione: (2025)
Documenti analoghi
-
Faithfulness metric fusion: Improving the evaluation of LLM trustworthiness across domains
di: Malin, Ben, et al.
Pubblicazione: (2025) -
Learning the Domain Specific Inverse NUFFT for Accelerated Spiral MRI using Diffusion Models
di: Chan, Trevor J., et al.
Pubblicazione: (2024) -
Generative AI voting: fair collective choice is resilient to LLM biases and inconsistencies
di: Majumdar, Srijoni, et al.
Pubblicazione: (2024) -
Developing trustworthy AI applications with foundation models
di: Mock, Michael, et al.
Pubblicazione: (2024) -
GENEOnet: Statistical analysis supporting explainability and trustworthiness
di: Bocchi, Giovanni, et al.
Pubblicazione: (2025)