Deployment-Relevant Alignment Cannot Be Inferred from Model-Level Evaluation Alone
Fuente:
arXiv
Enregistré dans:
| Auteurs principaux: | Vishwarupe, Varad, Shadbolt, Nigel, Jirotka, Marina, Flechais, Ivan |
|---|---|
| Format: | Preprint |
| Publié: |
2026
|
| Sujets: | |
| Accès en ligne: | |
| Tags: |
Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
|
Documents similaires
The Evaluation Differential: When Frontier AI Models Recognise They Are Being Tested
par: Vishwarupe, Varad, et autres
Publié: (2026)
par: Vishwarupe, Varad, et autres
Publié: (2026)
NeurIPS Should Require Reproducibility Standards for Frontier AI Safety Claims
par: Vishwarupe, Varad, et autres
Publié: (2026)
par: Vishwarupe, Varad, et autres
Publié: (2026)
The Collaboration Gap in Human-AI Work
par: Vishwarupe, Varad, et autres
Publié: (2026)
par: Vishwarupe, Varad, et autres
Publié: (2026)
To LLM, or Not to LLM: How Designers and Developers Navigate LLMs as Tools or Teammates
par: Vishwarupe, Varad, et autres
Publié: (2026)
par: Vishwarupe, Varad, et autres
Publié: (2026)
From Rights to Rites: Expectations Management in Smart-Home AI
par: Vishwarupe, Varad, et autres
Publié: (2026)
par: Vishwarupe, Varad, et autres
Publié: (2026)
From Sycophantic Consensus to Pluralistic Repair: Why AI Alignment Must Surface Disagreement
par: Vishwarupe, Varad, et autres
Publié: (2026)
par: Vishwarupe, Varad, et autres
Publié: (2026)
Follow-up Attention: An Empirical Study of Developer and Neural Model Code Exploration
par: Paltenghi, Matteo, et autres
Publié: (2022)
par: Paltenghi, Matteo, et autres
Publié: (2022)
Testing software for non-discrimination: an updated and extended audit in the Italian car insurance domain
par: Rondina, Marco, et autres
Publié: (2025)
par: Rondina, Marco, et autres
Publié: (2025)
MentalGame: Predicting Personality-Job Fitness for Software Developers Using Multi-Genre Games and Machine Learning Approaches
par: Elyasi, Soroush, et autres
Publié: (2026)
par: Elyasi, Soroush, et autres
Publié: (2026)
Human-In-the-Loop Software Development Agents
par: Takerngsaksiri, Wannita, et autres
Publié: (2024)
par: Takerngsaksiri, Wannita, et autres
Publié: (2024)
A Transformer-Based Approach for Smart Invocation of Automatic Code Completion
par: de Moor, Aral, et autres
Publié: (2024)
par: de Moor, Aral, et autres
Publié: (2024)
Requirements Engineering for Older Adult Digital Health Software: A Systematic Literature Review
par: Xiao, Yuqing, et autres
Publié: (2024)
par: Xiao, Yuqing, et autres
Publié: (2024)
Deriva-ML: A Continuous FAIRness Approach to Reproducible Machine Learning Models
par: Li, Zhiwei, et autres
Publié: (2024)
par: Li, Zhiwei, et autres
Publié: (2024)
Collaborative AI in Sentiment Analysis: System Architecture, Data Prediction and Deployment Strategies
par: Zhang, Chaofeng, et autres
Publié: (2024)
par: Zhang, Chaofeng, et autres
Publié: (2024)
AICoFe: Implementation and Deployment of an AI-Based Collaborative Feedback System for Higher Education
par: Becerra, Alvaro, et autres
Publié: (2026)
par: Becerra, Alvaro, et autres
Publié: (2026)
AISSA: Implementation and Deployment of an AI-based Student Slides Analysis tool for Academic Presentations
par: Becerra, Alvaro, et autres
Publié: (2026)
par: Becerra, Alvaro, et autres
Publié: (2026)
Enhanced Web User Interface Design Via Cross-Device Responsiveness Assessment Using An Improved HCI-INTEGRATED DL Schemes
par: Balasubramanian, Shrinivass Arunachalam
Publié: (2025)
par: Balasubramanian, Shrinivass Arunachalam
Publié: (2025)
Expecting the Unexpected: Developing Autonomous-System Design Principles for Reacting to Unpredicted Events and Conditions
par: Marron, Assaf, et autres
Publié: (2020)
par: Marron, Assaf, et autres
Publié: (2020)
Investigating Multimodal Large Language Models to Support Usability Evaluation
par: Lubos, Sebastian, et autres
Publié: (2025)
par: Lubos, Sebastian, et autres
Publié: (2025)
Evaluating the Quality of Code Comments Generated by Large Language Models for Novice Programmers
par: Fan, Aysa Xuemo, et autres
Publié: (2024)
par: Fan, Aysa Xuemo, et autres
Publié: (2024)
The RealHumanEval: Evaluating Large Language Models' Abilities to Support Programmers
par: Mozannar, Hussein, et autres
Publié: (2024)
par: Mozannar, Hussein, et autres
Publié: (2024)
Human-Centered Evaluation of an LLM-Based Process Modeling Copilot: A Mixed-Methods Study with Domain Experts
par: Lauer, Chantale, et autres
Publié: (2026)
par: Lauer, Chantale, et autres
Publié: (2026)
"I'm Not Reading All of That": Understanding Software Engineers' Level of Cognitive Engagement with Agentic Coding Assistants
par: Catalan, Carlos Rafael, et autres
Publié: (2026)
par: Catalan, Carlos Rafael, et autres
Publié: (2026)
Towards an Appropriate Level of Reliance on AI: A Preliminary Reliance-Control Framework for AI in Software Engineering
par: Ferino, Samuel, et autres
Publié: (2026)
par: Ferino, Samuel, et autres
Publié: (2026)
Natural Language Outlines for Code: Literate Programming in the LLM Era
par: Shi, Kensen, et autres
Publié: (2024)
par: Shi, Kensen, et autres
Publié: (2024)
Operationalizing AI: Empirical Evidence on MLOps Practices, User Satisfaction, and Organizational Context
par: Pasch, Stefan
Publié: (2025)
par: Pasch, Stefan
Publié: (2025)
Beyond Code Generation: An Observational Study of ChatGPT Usage in Software Engineering Practice
par: Khojah, Ranim, et autres
Publié: (2024)
par: Khojah, Ranim, et autres
Publié: (2024)
Feynman: Knowledge-Infused Diagramming Agent for Scalable Visual Designs
par: Wen, Zixin, et autres
Publié: (2026)
par: Wen, Zixin, et autres
Publié: (2026)
Harnessing IoT and Generative AI for Weather-Adaptive Learning in Climate Resilience Education
par: Khan, Imran S. A., et autres
Publié: (2025)
par: Khan, Imran S. A., et autres
Publié: (2025)
SWE-agent: Agent-Computer Interfaces Enable Automated Software Engineering
par: Yang, John, et autres
Publié: (2024)
par: Yang, John, et autres
Publié: (2024)
AutoGen Studio: A No-Code Developer Tool for Building and Debugging Multi-Agent Systems
par: Dibia, Victor, et autres
Publié: (2024)
par: Dibia, Victor, et autres
Publié: (2024)
Qualitative Evaluation of LLM-Designed GUI
par: Sawicki, Bartosz, et autres
Publié: (2026)
par: Sawicki, Bartosz, et autres
Publié: (2026)
Results-Actionability Gap: Understanding How Practitioners Evaluate LLM Products in the Wild
par: van der Maden, Willem, et autres
Publié: (2026)
par: van der Maden, Willem, et autres
Publié: (2026)
How Do Hackathons Foster Creativity? Towards AI Collaborative Evaluation of Creativity at Scale
par: Falk, Jeanette, et autres
Publié: (2025)
par: Falk, Jeanette, et autres
Publié: (2025)
Assistance or Disruption? Exploring and Evaluating the Design and Trade-offs of Proactive AI Programming Support
par: Pu, Kevin, et autres
Publié: (2025)
par: Pu, Kevin, et autres
Publié: (2025)
A Case Study Investigating the Role of Generative AI in Quality Evaluations of Epics in Agile Software Development
par: Geyer, Werner, et autres
Publié: (2025)
par: Geyer, Werner, et autres
Publié: (2025)
From Correctness to Collaboration: Toward a Human-Centered Framework for Evaluating AI Agent Behavior in Software Engineering
par: Dong, Tao, et autres
Publié: (2025)
par: Dong, Tao, et autres
Publié: (2025)
On the Utility of Domain Modeling Assistance with Large Language Models
par: Chaaben, Meriem Ben, et autres
Publié: (2024)
par: Chaaben, Meriem Ben, et autres
Publié: (2024)
Exploring the Efficacy of Robotic Assistants with ChatGPT and Claude in Enhancing ADHD Therapy: Innovating Treatment Paradigms
par: Berrezueta-Guzman, Santiago, et autres
Publié: (2024)
par: Berrezueta-Guzman, Santiago, et autres
Publié: (2024)
Generative AI for CAD Automation: Leveraging Large Language Models for 3D Modelling
par: Kumar, Sumit, et autres
Publié: (2025)
par: Kumar, Sumit, et autres
Publié: (2025)
Documents similaires
-
The Evaluation Differential: When Frontier AI Models Recognise They Are Being Tested
par: Vishwarupe, Varad, et autres
Publié: (2026) -
NeurIPS Should Require Reproducibility Standards for Frontier AI Safety Claims
par: Vishwarupe, Varad, et autres
Publié: (2026) -
The Collaboration Gap in Human-AI Work
par: Vishwarupe, Varad, et autres
Publié: (2026) -
To LLM, or Not to LLM: How Designers and Developers Navigate LLMs as Tools or Teammates
par: Vishwarupe, Varad, et autres
Publié: (2026) -
From Rights to Rites: Expectations Management in Smart-Home AI
par: Vishwarupe, Varad, et autres
Publié: (2026)