RELIC: Investigating Large Language Model Responses using Self-Consistency
Fuente:
arXiv
Saved in:
| Main Authors: | Cheng, Furui, Zouhar, Vilém, Arora, Simran, Sachan, Mrinmaya, Strobelt, Hendrik, El-Assady, Mennatallah |
|---|---|
| Format: | Preprint |
| Published: |
2023
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
CafGa: Customizing Feature Attributions to Explain Language Models
by: Boyle, Alan, et al.
Published: (2025)
by: Boyle, Alan, et al.
Published: (2025)
Understanding Large Language Model Behaviors through Interactive Counterfactual Generation and Analysis
by: Cheng, Furui, et al.
Published: (2024)
by: Cheng, Furui, et al.
Published: (2024)
AutoTutor meets Large Language Models: A Language Model Tutor with Rich Pedagogy and Guardrails
by: Chowdhury, Sankalan Pal, et al.
Published: (2024)
by: Chowdhury, Sankalan Pal, et al.
Published: (2024)
Pearmut: Human Evaluation of Translation Made Trivial
by: Zouhar, Vilém, et al.
Published: (2026)
by: Zouhar, Vilém, et al.
Published: (2026)
PleaSQLarify: Visual Pragmatic Repair for Natural Language Database Querying
by: Chan, Robin Shing Moon, et al.
Published: (2026)
by: Chan, Robin Shing Moon, et al.
Published: (2026)
DxHF: Providing High-Quality Human Feedback for LLM Alignment via Interactive Decomposition
by: Shi, Danqing, et al.
Published: (2025)
by: Shi, Danqing, et al.
Published: (2025)
Deconstructing Human-AI Collaboration: Agency, Interaction, and Adaptation
by: Holter, Steffen, et al.
Published: (2024)
by: Holter, Steffen, et al.
Published: (2024)
iToT: An Interactive System for Customized Tree-of-Thought Generation
by: Boyle, Alan, et al.
Published: (2024)
by: Boyle, Alan, et al.
Published: (2024)
Unsupervised Word-level Quality Estimation for Machine Translation Through the Lens of Annotators (Dis)agreement
by: Sarti, Gabriele, et al.
Published: (2025)
by: Sarti, Gabriele, et al.
Published: (2025)
QE4PE: Word-level Quality Estimation for Human Post-Editing
by: Sarti, Gabriele, et al.
Published: (2025)
by: Sarti, Gabriele, et al.
Published: (2025)
AI-Assisted Human Evaluation of Machine Translation
by: Zouhar, Vilém, et al.
Published: (2024)
by: Zouhar, Vilém, et al.
Published: (2024)
How to Select Datapoints for Efficient Human Evaluation of NLG Models?
by: Zouhar, Vilém, et al.
Published: (2025)
by: Zouhar, Vilém, et al.
Published: (2025)
ProvenanceWidgets: A Library of UI Control Elements to Track and Dynamically Overlay Analytic Provenance
by: Narechania, Arpit, et al.
Published: (2024)
by: Narechania, Arpit, et al.
Published: (2024)
iNNspector: Visual, Interactive Deep Model Debugging
by: Spinner, Thilo, et al.
Published: (2024)
by: Spinner, Thilo, et al.
Published: (2024)
Generating Pedagogically Meaningful Visuals for Math Word Problems: A New Benchmark and Analysis of Text-to-Image Models
by: Wang, Junling, et al.
Published: (2025)
by: Wang, Junling, et al.
Published: (2025)
Dia-Lingle: A Gamified Interface for Dialectal Data Collection
by: Sun, Jiugeng, et al.
Published: (2025)
by: Sun, Jiugeng, et al.
Published: (2025)
GEMS -- Guided Evolutionary Molecule Design for Sustainable Chemicals
by: Robinson, Coelina, et al.
Published: (2026)
by: Robinson, Coelina, et al.
Published: (2026)
Mapping out the Space of Human Feedback for Reinforcement Learning: A Conceptual Framework
by: Metz, Yannick, et al.
Published: (2024)
by: Metz, Yannick, et al.
Published: (2024)
Abstraction Alignment: Comparing Model-Learned and Human-Encoded Conceptual Relationships
by: Boggust, Angie, et al.
Published: (2024)
by: Boggust, Angie, et al.
Published: (2024)
Can LLMs Model Incorrect Student Reasoning? A Case Study on Distractor Generation
by: Zengaffinen, Yanick, et al.
Published: (2026)
by: Zengaffinen, Yanick, et al.
Published: (2026)
LLM-based Cognitive Models of Students with Misconceptions
by: Sonkar, Shashank, et al.
Published: (2024)
by: Sonkar, Shashank, et al.
Published: (2024)
VACP: Visual Analytics Context Protocol
by: Stähle, Tobias, et al.
Published: (2026)
by: Stähle, Tobias, et al.
Published: (2026)
TopoAlign: Topology-Aware Visual Representation Alignment
by: Yan, Xinyuan, et al.
Published: (2026)
by: Yan, Xinyuan, et al.
Published: (2026)
Challenges and Opportunities for Visual Analytics in Jurisprudence
by: Fürst, Daniel, et al.
Published: (2024)
by: Fürst, Daniel, et al.
Published: (2024)
Consistency of Responses and Continuations Generated by Large Language Models on Social Media
by: Xu, Wentao, et al.
Published: (2025)
by: Xu, Wentao, et al.
Published: (2025)
How to Engage Your Readers? Generating Guiding Questions to Promote Active Reading
by: Cui, Peng, et al.
Published: (2024)
by: Cui, Peng, et al.
Published: (2024)
A Design Space for Intelligent Agents in Mixed-Initiative Visual Analytics
by: Stähle, Tobias, et al.
Published: (2025)
by: Stähle, Tobias, et al.
Published: (2025)
Multilingual Performance Biases of Large Language Models in Education
by: Gupta, Vansh, et al.
Published: (2025)
by: Gupta, Vansh, et al.
Published: (2025)
Implicit Personalization in Language Models: A Systematic Study
by: Jin, Zhijing, et al.
Published: (2024)
by: Jin, Zhijing, et al.
Published: (2024)
Characterizing Grounded Theory Approaches in Visualization
by: Diehl, Alexandra, et al.
Published: (2022)
by: Diehl, Alexandra, et al.
Published: (2022)
Word Synchronization Challenge: A Benchmark for Word Association Responses for Large Language Models
by: Cazalets, Tanguy, et al.
Published: (2025)
by: Cazalets, Tanguy, et al.
Published: (2025)
SemanticTours: A Conceptual Framework for Non-Linear, Knowledge Graph-Driven Data Tours
by: Fürst, Daniel, et al.
Published: (2025)
by: Fürst, Daniel, et al.
Published: (2025)
Early-Exit and Instant Confidence Translation Quality Estimation
by: Zouhar, Vilém, et al.
Published: (2025)
by: Zouhar, Vilém, et al.
Published: (2025)
Bridging Instead of Replacing Online Coding Communities with AI through Community-Enriched Chatbot Designs
by: Wang, Junling, et al.
Published: (2026)
by: Wang, Junling, et al.
Published: (2026)
Exploring Large Language Models for Specialist-level Oncology Care
by: Palepu, Anil, et al.
Published: (2024)
by: Palepu, Anil, et al.
Published: (2024)
Exploring the Impact of Personality Traits on Conversational Recommender Systems: A Simulation with Large Language Models
by: Zhao, Xiaoyan, et al.
Published: (2025)
by: Zhao, Xiaoyan, et al.
Published: (2025)
Are Humans as Brittle as Large Language Models?
by: Li, Jiahui, et al.
Published: (2025)
by: Li, Jiahui, et al.
Published: (2025)
"Harmless to You, Hurtful to Me!": Investigating the Detection of Toxic Languages Grounded in the Perspective of Youth
by: Li, Yaqiong, et al.
Published: (2025)
by: Li, Yaqiong, et al.
Published: (2025)
Reducing Privacy Risks in Online Self-Disclosures with Language Models
by: Dou, Yao, et al.
Published: (2023)
by: Dou, Yao, et al.
Published: (2023)
PersonaLLM: Investigating the Ability of Large Language Models to Express Personality Traits
by: Jiang, Hang, et al.
Published: (2023)
by: Jiang, Hang, et al.
Published: (2023)
Similar Items
-
CafGa: Customizing Feature Attributions to Explain Language Models
by: Boyle, Alan, et al.
Published: (2025) -
Understanding Large Language Model Behaviors through Interactive Counterfactual Generation and Analysis
by: Cheng, Furui, et al.
Published: (2024) -
AutoTutor meets Large Language Models: A Language Model Tutor with Rich Pedagogy and Guardrails
by: Chowdhury, Sankalan Pal, et al.
Published: (2024) -
Pearmut: Human Evaluation of Translation Made Trivial
by: Zouhar, Vilém, et al.
Published: (2026) -
PleaSQLarify: Visual Pragmatic Repair for Natural Language Database Querying
by: Chan, Robin Shing Moon, et al.
Published: (2026)