iRULER: Intelligible Rubric-Based User-Defined LLM Evaluation for Revision
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Bai, Jingwen, Cheong, Wei Soon, Muller, Philippe, Lim, Brian Y |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Editable XAI: Toward Bidirectional Human-AI Alignment with Co-Editable Explanations of Interpretable Attributes
von: Chen, Haoyang, et al.
Veröffentlicht: (2026)
von: Chen, Haoyang, et al.
Veröffentlicht: (2026)
Who Defines "Best"? Towards Interactive, User-Defined Evaluation of LLM Leaderboards
von: Jung, Minji, et al.
Veröffentlicht: (2026)
von: Jung, Minji, et al.
Veröffentlicht: (2026)
Beyond Scores: Explainable Intelligent Assessment Strengthens Pre-service Teachers' Assessment Literacy
von: Wei, Yuang, et al.
Veröffentlicht: (2026)
von: Wei, Yuang, et al.
Veröffentlicht: (2026)
User Experience with LLM-powered Conversational Recommendation Systems: A Case of Music Recommendation
von: Yun, Sojeong, et al.
Veröffentlicht: (2025)
von: Yun, Sojeong, et al.
Veröffentlicht: (2025)
PREFINE: Personalized Story Generation via Simulated User Critics and User-Specific Rubric Generation
von: Ueda, Kentaro, et al.
Veröffentlicht: (2025)
von: Ueda, Kentaro, et al.
Veröffentlicht: (2025)
Varif.ai to Vary and Verify User-Driven Diversity in Scalable Image Generation
von: Michelessa, M., et al.
Veröffentlicht: (2025)
von: Michelessa, M., et al.
Veröffentlicht: (2025)
U-Define: Designing User Workflows for Hard and Soft Constraints in LLM-Based Planning
von: Lee, Christine P, et al.
Veröffentlicht: (2026)
von: Lee, Christine P, et al.
Veröffentlicht: (2026)
InFerActive: Interactive Tree-Based Exploration of LLM Sampling for Safety Evaluation
von: Hwangbo, Junhyeong, et al.
Veröffentlicht: (2025)
von: Hwangbo, Junhyeong, et al.
Veröffentlicht: (2025)
LLM-Based Educational Simulation: Evaluating Temporal Student Persona Stability Across ADHD Profiles
von: Gonnermann-Müller, Jana, et al.
Veröffentlicht: (2026)
von: Gonnermann-Müller, Jana, et al.
Veröffentlicht: (2026)
Designing an LLM-Based Behavioral Activation Chatbot for Young People with Depression: Insights from an Evaluation with Artificial Users and Clinical Experts
von: Kuhlmeier, Florian Onur, et al.
Veröffentlicht: (2025)
von: Kuhlmeier, Florian Onur, et al.
Veröffentlicht: (2025)
CoPrompter: User-Centric Evaluation of LLM Instruction Alignment for Improved Prompt Engineering
von: Joshi, Ishika, et al.
Veröffentlicht: (2024)
von: Joshi, Ishika, et al.
Veröffentlicht: (2024)
Generative AI as a Tool or Leader? Exploring AI-Augmented Thinking in Student Programming Tasks
von: Zhong, Tianlong, et al.
Veröffentlicht: (2024)
von: Zhong, Tianlong, et al.
Veröffentlicht: (2024)
AI-Generated Rubric Interfaces: K-12 Teachers' Perceptions and Practices
von: Riahi, Bahare, et al.
Veröffentlicht: (2026)
von: Riahi, Bahare, et al.
Veröffentlicht: (2026)
Enhancing XAI Interpretation through a Reverse Mapping from Insights to Visualizations
von: Nuthalapati, Aniket, et al.
Veröffentlicht: (2025)
von: Nuthalapati, Aniket, et al.
Veröffentlicht: (2025)
Assessment of Sign Language-Based versus Touch-Based Input for Deaf Users Interacting with Intelligent Personal Assistants
von: Tran, Nina, et al.
Veröffentlicht: (2024)
von: Tran, Nina, et al.
Veröffentlicht: (2024)
User Perceptions of an LLM-Based Chatbot for Cognitive Reappraisal of Stress: Feasibility Study
von: Bhattacharjee, Ananya, et al.
Veröffentlicht: (2026)
von: Bhattacharjee, Ananya, et al.
Veröffentlicht: (2026)
Evaluation of Task Specific Productivity Improvements Using a Generative Artificial Intelligence Personal Assistant Tool
von: Freeman, Brian S., et al.
Veröffentlicht: (2024)
von: Freeman, Brian S., et al.
Veröffentlicht: (2024)
Building Intelligent User Interfaces for Human-AI Alignment
von: Shi, Danqing
Veröffentlicht: (2026)
von: Shi, Danqing
Veröffentlicht: (2026)
User Agency and System Automation in Interactive Intelligent Systems
von: Langerak, Thomas
Veröffentlicht: (2025)
von: Langerak, Thomas
Veröffentlicht: (2025)
CrowdGenUI: Aligning LLM-Based UI Generation with Crowdsourced User Preferences
von: Liu, Yimeng, et al.
Veröffentlicht: (2024)
von: Liu, Yimeng, et al.
Veröffentlicht: (2024)
EvalLM: Interactive Evaluation of Large Language Model Prompts on User-Defined Criteria
von: Kim, Tae Soo, et al.
Veröffentlicht: (2023)
von: Kim, Tae Soo, et al.
Veröffentlicht: (2023)
Human and LLM-Based Voice Assistant Interaction: An Analytical Framework for User Verbal and Nonverbal Behaviors
von: Chan, Szeyi, et al.
Veröffentlicht: (2024)
von: Chan, Szeyi, et al.
Veröffentlicht: (2024)
Users and Wizards in Conversations: How WoZ Interface Choices Define Human-Robot Interactions
von: Torubarova, Ekaterina, et al.
Veröffentlicht: (2026)
von: Torubarova, Ekaterina, et al.
Veröffentlicht: (2026)
Incremental XAI: Memorable Understanding of AI with Incremental Explanations
von: Bo, Jessica Y., et al.
Veröffentlicht: (2024)
von: Bo, Jessica Y., et al.
Veröffentlicht: (2024)
Thinking Assistants: LLM-Based Conversational Assistants that Help Users Think By Asking rather than Answering
von: Park, Soya, et al.
Veröffentlicht: (2023)
von: Park, Soya, et al.
Veröffentlicht: (2023)
Assessing Policy Updates: Toward Trust-Preserving Intelligent User Interfaces
von: Solomon, Matan, et al.
Veröffentlicht: (2025)
von: Solomon, Matan, et al.
Veröffentlicht: (2025)
Canvil: Designerly Adaptation for LLM-Powered User Experiences
von: Feng, K. J. Kevin, et al.
Veröffentlicht: (2024)
von: Feng, K. J. Kevin, et al.
Veröffentlicht: (2024)
From Control to Foresight: Simulation as a New Paradigm for Human-Agent Collaboration
von: He, Gaole, et al.
Veröffentlicht: (2026)
von: He, Gaole, et al.
Veröffentlicht: (2026)
Guiding, not Driving: Design and Evaluation of a Command-Based User Interface for Teleoperation of Autonomous Vehicles
von: Tener, Felix, et al.
Veröffentlicht: (2025)
von: Tener, Felix, et al.
Veröffentlicht: (2025)
Evaluating User Experience and Data Quality in Gamified Data Collection for Appearance-Based Gaze Estimation
von: Yue, Mingtao, et al.
Veröffentlicht: (2024)
von: Yue, Mingtao, et al.
Veröffentlicht: (2024)
Temporal Drift in Privacy Recall: Users Misremember From Verbatim Loss to Gist-Based Overexposure
von: Guo, Haoze, et al.
Veröffentlicht: (2025)
von: Guo, Haoze, et al.
Veröffentlicht: (2025)
Digital Twins for Extended Reality Tourism: User Experience Evaluation Across User Groups
von: Warsinke, Maximilian, et al.
Veröffentlicht: (2025)
von: Warsinke, Maximilian, et al.
Veröffentlicht: (2025)
Interface Framework for Human-AI Collaboration within Intelligent User Interface Ecosystems
von: Andru, Shruthi, et al.
Veröffentlicht: (2026)
von: Andru, Shruthi, et al.
Veröffentlicht: (2026)
User Misconceptions of LLM-Based Conversational Programming Assistants
von: O'Brien, Gabrielle, et al.
Veröffentlicht: (2025)
von: O'Brien, Gabrielle, et al.
Veröffentlicht: (2025)
Final Happiness: What Intelligent User Interfaces Can Do for The Lonely Dying
von: Meng, Yibo, et al.
Veröffentlicht: (2025)
von: Meng, Yibo, et al.
Veröffentlicht: (2025)
MYCloth: Towards Intelligent and Interactive Online T-Shirt Customization based on User's Preference
von: Liu, Yexin, et al.
Veröffentlicht: (2024)
von: Liu, Yexin, et al.
Veröffentlicht: (2024)
Seeing the Reasoning: How LLM Rationales Influence User Trust and Decision-Making in Factual Verification Tasks
von: Sun, Xin, et al.
Veröffentlicht: (2026)
von: Sun, Xin, et al.
Veröffentlicht: (2026)
Be Friendly, Not Friends: How LLM Sycophancy Shapes User Trust
von: Sun, Yuan, et al.
Veröffentlicht: (2025)
von: Sun, Yuan, et al.
Veröffentlicht: (2025)
Transferable XAI: Relating Understanding Across Domains with Explanation Transfer
von: Wang, Fei, et al.
Veröffentlicht: (2026)
von: Wang, Fei, et al.
Veröffentlicht: (2026)
Stable Personas: Dual-Assessment of Temporal Stability in LLM-Based Human Simulation
von: Gonnermann-Müller, Jana, et al.
Veröffentlicht: (2026)
von: Gonnermann-Müller, Jana, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
Editable XAI: Toward Bidirectional Human-AI Alignment with Co-Editable Explanations of Interpretable Attributes
von: Chen, Haoyang, et al.
Veröffentlicht: (2026) -
Who Defines "Best"? Towards Interactive, User-Defined Evaluation of LLM Leaderboards
von: Jung, Minji, et al.
Veröffentlicht: (2026) -
Beyond Scores: Explainable Intelligent Assessment Strengthens Pre-service Teachers' Assessment Literacy
von: Wei, Yuang, et al.
Veröffentlicht: (2026) -
User Experience with LLM-powered Conversational Recommendation Systems: A Case of Music Recommendation
von: Yun, Sojeong, et al.
Veröffentlicht: (2025) -
PREFINE: Personalized Story Generation via Simulated User Critics and User-Specific Rubric Generation
von: Ueda, Kentaro, et al.
Veröffentlicht: (2025)