Towards Transparent AI Grading: Semantic Entropy as a Signal for Human-AI Disagreement
Fuente:
arXiv
Guardado en:
| Autores principales: | Iyer, Karrtik, Ravikiran, Manikandan, Pendse, Prasanna, Mohanty, Shayan |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
The Tokenization Bottleneck: How Vocabulary Extension Improves Chemistry Representation Learning in Pretrained Language Models
por: Kalamkar, Prathamesh, et al.
Publicado: (2025)
por: Kalamkar, Prathamesh, et al.
Publicado: (2025)
A Value-Based Approach to AI Ethics: Accountability, Transparency, Explainability, and Usability
por: Vish Iyer
Publicado: (2025)
por: Vish Iyer
Publicado: (2025)
TrueGradeAI: Retrieval-Augmented and Bias-Resistant AI for Transparent and Explainable Digital Assessments
por: Thakur, Rakesh, et al.
Publicado: (2025)
por: Thakur, Rakesh, et al.
Publicado: (2025)
AI-Powered Annotation Pipelines for Stabilizing Large Language Models: A Human-AI Synergy Approach
por: Pathak, Gangesh, et al.
Publicado: (2025)
por: Pathak, Gangesh, et al.
Publicado: (2025)
Humanizing AI Grading: Student-Centered Insights on Fairness, Trust, Consistency and Transparency
por: Riahi, Bahare, et al.
Publicado: (2026)
por: Riahi, Bahare, et al.
Publicado: (2026)
Learning from Disagreement: Clinician Overrides as Implicit Preference Signals for Clinical AI in Value-Based Care
por: Singh, Prabhjot, et al.
Publicado: (2026)
por: Singh, Prabhjot, et al.
Publicado: (2026)
STABLEVAL: Disagreement-Aware and Stable Evaluation of AI Systems
por: Bonagiri, Akash, et al.
Publicado: (2026)
por: Bonagiri, Akash, et al.
Publicado: (2026)
AI Safeguards, Generative AI and the Pandora Box: AI Safety Measures to Protect Businesses and Personal Reputation
por: Kumar, Prasanna
Publicado: (2026)
por: Kumar, Prasanna
Publicado: (2026)
Human-AI Collaborative Uncertainty Quantification
por: Noorani, Sima, et al.
Publicado: (2025)
por: Noorani, Sima, et al.
Publicado: (2025)
Cross-Model Disagreement as a Label-Free Correctness Signal
por: Gorbett, Matt, et al.
Publicado: (2026)
por: Gorbett, Matt, et al.
Publicado: (2026)
"Let's Agree to Disagree": Investigating the Disagreement Problem in Explainable AI for Text Summarization
por: Aswani, Seema, et al.
Publicado: (2024)
por: Aswani, Seema, et al.
Publicado: (2024)
Towards Meaningful Transparency in Civic AI Systems
por: Murray-Rust, Dave, et al.
Publicado: (2025)
por: Murray-Rust, Dave, et al.
Publicado: (2025)
RFBES at SemEval-2024 Task 8: Investigating Syntactic and Semantic Features for Distinguishing AI-Generated and Human-Written Texts
por: Rad, Mohammad Heydari, et al.
Publicado: (2024)
por: Rad, Mohammad Heydari, et al.
Publicado: (2024)
Towards Trustworthy and Explainable AI for Perception Models: From Concept to Prototype Vehicle Deployment
por: Beemelmanns, Till, et al.
Publicado: (2026)
por: Beemelmanns, Till, et al.
Publicado: (2026)
Optimizing Generative AI's Accuracy and Transparency in Inductive Thematic Analysis: A Human-AI Comparison
por: Nyaaba, Matthew, et al.
Publicado: (2025)
por: Nyaaba, Matthew, et al.
Publicado: (2025)
Towards AI Transparency and Accountability: A Global Framework for Exchanging Information on AI Systems
por: Buckley, Warren, et al.
Publicado: (2023)
por: Buckley, Warren, et al.
Publicado: (2023)
AI, Climate, and Transparency: Operationalizing and Improving the AI Act
por: Alder, Nicolas, et al.
Publicado: (2024)
por: Alder, Nicolas, et al.
Publicado: (2024)
Benchmarking Edge AI Platforms for High-Performance ML Inference
por: Jayanth, Rakshith, et al.
Publicado: (2024)
por: Jayanth, Rakshith, et al.
Publicado: (2024)
AI Model Passport: Data and System Traceability Framework for Transparent AI in Health
por: Kalokyri, Varvara, et al.
Publicado: (2025)
por: Kalokyri, Varvara, et al.
Publicado: (2025)
Overview of AI Grading of Physics Olympiad Exams
por: McGinness, Lachlan
Publicado: (2025)
por: McGinness, Lachlan
Publicado: (2025)
Argumentative Human-AI Decision-Making: Toward AI Agents That Reason With Us, Not For Us
por: Vasileiou, Stylianos Loukas, et al.
Publicado: (2026)
por: Vasileiou, Stylianos Loukas, et al.
Publicado: (2026)
We Are All Creators: Generative AI, Collective Knowledge, and the Path Towards Human-AI Synergy
por: Linares-Pellicer, Jordi, et al.
Publicado: (2025)
por: Linares-Pellicer, Jordi, et al.
Publicado: (2025)
The Dark Side of AI Transformers: Sentiment Polarization & the Loss of Business Neutrality by NLP Transformers
por: Kumar, Prasanna
Publicado: (2026)
por: Kumar, Prasanna
Publicado: (2026)
Generative AI-Driven High-Fidelity Human Motion Simulation
por: Iyer, Hari, et al.
Publicado: (2025)
por: Iyer, Hari, et al.
Publicado: (2025)
Towards Dialogues for Joint Human-AI Reasoning and Value Alignment
por: Bezou-Vrakatseli, Elfia, et al.
Publicado: (2024)
por: Bezou-Vrakatseli, Elfia, et al.
Publicado: (2024)
Measuring What Matters: Benchmarking Generative, Multimodal, and Agentic AI in Healthcare
por: Desikan, Prasanna, et al.
Publicado: (2026)
por: Desikan, Prasanna, et al.
Publicado: (2026)
Discrete-Space Generative AI Pipeline for Semantic Transmission of Signals
por: Kokalj-Filipovic, Silvija, et al.
Publicado: (2026)
por: Kokalj-Filipovic, Silvija, et al.
Publicado: (2026)
Future of Code with Generative AI: Transparency and Safety in the Era of AI Generated Software
por: Hanson, David
Publicado: (2025)
por: Hanson, David
Publicado: (2025)
Holistic Explainable AI (H-XAI): Extending Transparency Beyond Developers in AI-Driven Decision Making
por: Lakkaraju, Kausik, et al.
Publicado: (2025)
por: Lakkaraju, Kausik, et al.
Publicado: (2025)
Toward Safe and Responsible AI Agents: A Three-Pillar Model for Transparency, Accountability, and Trustworthiness
por: Cheng, Edward C., et al.
Publicado: (2026)
por: Cheng, Edward C., et al.
Publicado: (2026)
Beyond Autocomplete: Designing CopilotLens Towards Transparent and Explainable AI Coding Agents
por: Ye, Runlong, et al.
Publicado: (2025)
por: Ye, Runlong, et al.
Publicado: (2025)
Trends in AI and Human-AI Interaction in Clinical Trials -- A Hybrid Human-AI Exploration
por: Woolley, Sandra, et al.
Publicado: (2026)
por: Woolley, Sandra, et al.
Publicado: (2026)
Hey AI Can You Grade My Essay?: Automatic Essay Grading
por: Maliha, Maisha, et al.
Publicado: (2024)
por: Maliha, Maisha, et al.
Publicado: (2024)
Language Models in Dialogue: Conversational Maxims for Human-AI Interactions
por: Miehling, Erik, et al.
Publicado: (2024)
por: Miehling, Erik, et al.
Publicado: (2024)
Transparent AI: The Case for Interpretability and Explainability
por: Ramachandram, Dhanesh, et al.
Publicado: (2025)
por: Ramachandram, Dhanesh, et al.
Publicado: (2025)
HumanStudy-Bench: Towards AI Agent Design for Participant Simulation
por: Liu, Xuan, et al.
Publicado: (2026)
por: Liu, Xuan, et al.
Publicado: (2026)
Toward Agentic Environments: GenAI and the Convergence of AI, Sustainability, and Human-Centric Spaces
por: Pospieszny, Przemek, et al.
Publicado: (2025)
por: Pospieszny, Przemek, et al.
Publicado: (2025)
AI Application in Anti-Money Laundering for Sustainable and Transparent Financial Systems
por: Nie, Chuanhao, et al.
Publicado: (2025)
por: Nie, Chuanhao, et al.
Publicado: (2025)
Human-AI Coevolution
por: Pedreschi, Dino, et al.
Publicado: (2023)
por: Pedreschi, Dino, et al.
Publicado: (2023)
BotzoneBench: Scalable LLM Evaluation via Graded AI Anchors
por: Li, Lingfeng, et al.
Publicado: (2026)
por: Li, Lingfeng, et al.
Publicado: (2026)
Ejemplares similares
-
The Tokenization Bottleneck: How Vocabulary Extension Improves Chemistry Representation Learning in Pretrained Language Models
por: Kalamkar, Prathamesh, et al.
Publicado: (2025) -
A Value-Based Approach to AI Ethics: Accountability, Transparency, Explainability, and Usability
por: Vish Iyer
Publicado: (2025) -
TrueGradeAI: Retrieval-Augmented and Bias-Resistant AI for Transparent and Explainable Digital Assessments
por: Thakur, Rakesh, et al.
Publicado: (2025) -
AI-Powered Annotation Pipelines for Stabilizing Large Language Models: A Human-AI Synergy Approach
por: Pathak, Gangesh, et al.
Publicado: (2025) -
Humanizing AI Grading: Student-Centered Insights on Fairness, Trust, Consistency and Transparency
por: Riahi, Bahare, et al.
Publicado: (2026)