To LLM, or Not to LLM: How Designers and Developers Navigate LLMs as Tools or Teammates
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Vishwarupe, Varad, Flechais, Ivan, Shadbolt, Nigel, Jirotka, Marina |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
The Collaboration Gap in Human-AI Work
von: Vishwarupe, Varad, et al.
Veröffentlicht: (2026)
von: Vishwarupe, Varad, et al.
Veröffentlicht: (2026)
From Rights to Rites: Expectations Management in Smart-Home AI
von: Vishwarupe, Varad, et al.
Veröffentlicht: (2026)
von: Vishwarupe, Varad, et al.
Veröffentlicht: (2026)
The Evaluation Differential: When Frontier AI Models Recognise They Are Being Tested
von: Vishwarupe, Varad, et al.
Veröffentlicht: (2026)
von: Vishwarupe, Varad, et al.
Veröffentlicht: (2026)
Deployment-Relevant Alignment Cannot Be Inferred from Model-Level Evaluation Alone
von: Vishwarupe, Varad, et al.
Veröffentlicht: (2026)
von: Vishwarupe, Varad, et al.
Veröffentlicht: (2026)
From Sycophantic Consensus to Pluralistic Repair: Why AI Alignment Must Surface Disagreement
von: Vishwarupe, Varad, et al.
Veröffentlicht: (2026)
von: Vishwarupe, Varad, et al.
Veröffentlicht: (2026)
NeurIPS Should Require Reproducibility Standards for Frontier AI Safety Claims
von: Vishwarupe, Varad, et al.
Veröffentlicht: (2026)
von: Vishwarupe, Varad, et al.
Veröffentlicht: (2026)
Almanac Copilot: Towards Autonomous Electronic Health Record Navigation
von: Zakka, Cyril, et al.
Veröffentlicht: (2024)
von: Zakka, Cyril, et al.
Veröffentlicht: (2024)
SCOPE: A Lightweight-training LLM Framework for Air Traffic Control Readback Monitoring
von: Deng, Qihan, et al.
Veröffentlicht: (2026)
von: Deng, Qihan, et al.
Veröffentlicht: (2026)
Making Language Models Better Tool Learners with Execution Feedback
von: Qiao, Shuofei, et al.
Veröffentlicht: (2023)
von: Qiao, Shuofei, et al.
Veröffentlicht: (2023)
Summaries, Highlights, and Action items: Design, implementation and evaluation of an LLM-powered meeting recap system
von: Asthana, Sumit, et al.
Veröffentlicht: (2023)
von: Asthana, Sumit, et al.
Veröffentlicht: (2023)
Personalized Benchmarking: Evaluating LLMs by Individual Preferences
von: Garbacea, Cristina, et al.
Veröffentlicht: (2026)
von: Garbacea, Cristina, et al.
Veröffentlicht: (2026)
Ink and Individuality: Crafting a Personalised Narrative in the Age of LLMs
von: Wasi, Azmine Toushik, et al.
Veröffentlicht: (2024)
von: Wasi, Azmine Toushik, et al.
Veröffentlicht: (2024)
Retentive Relevance: Capturing Long-Term User Value in Recommendation Systems
von: Bakhshi, Saeideh, et al.
Veröffentlicht: (2025)
von: Bakhshi, Saeideh, et al.
Veröffentlicht: (2025)
MeMemo: On-device Retrieval Augmentation for Private and Personalized Text Generation
von: Wang, Zijie J., et al.
Veröffentlicht: (2024)
von: Wang, Zijie J., et al.
Veröffentlicht: (2024)
User Simulation in the Era of Generative AI: User Modeling, Synthetic Data Generation, and System Evaluation
von: Balog, Krisztian, et al.
Veröffentlicht: (2025)
von: Balog, Krisztian, et al.
Veröffentlicht: (2025)
TreeHop: Generate and Filter Next Query Embeddings Efficiently for Multi-hop Question Answering
von: Li, Zhonghao, et al.
Veröffentlicht: (2025)
von: Li, Zhonghao, et al.
Veröffentlicht: (2025)
FairEval: Evaluating Fairness in LLM-Based Recommendations with Personality Awareness
von: Sah, Chandan Kumar, et al.
Veröffentlicht: (2025)
von: Sah, Chandan Kumar, et al.
Veröffentlicht: (2025)
Anna Karenina Strikes Again: Pre-Trained LLM Embeddings May Favor High-Performing Learners
von: Schleifer, Abigail Gurin, et al.
Veröffentlicht: (2024)
von: Schleifer, Abigail Gurin, et al.
Veröffentlicht: (2024)
SoulSeek: Exploring the Use of Social Cues in LLM-based Information Seeking
von: Shu, Yubo, et al.
Veröffentlicht: (2026)
von: Shu, Yubo, et al.
Veröffentlicht: (2026)
Three Modalities, Two Design Probes, One Prototype, and No Vision: Experience-Based Co-Design of a Multi-modal 3D Data Visualization Tool
von: Kamath, Sanchita S., et al.
Veröffentlicht: (2026)
von: Kamath, Sanchita S., et al.
Veröffentlicht: (2026)
Trustworthy AI Psychotherapy: Multi-Agent LLM Workflow for Counseling and Explainable Mental Disorder Diagnosis
von: Ozgun, Mithat Can, et al.
Veröffentlicht: (2025)
von: Ozgun, Mithat Can, et al.
Veröffentlicht: (2025)
Talk2X -- An Open-Source Toolkit Facilitating Deployment of LLM-Powered Chatbots on the Web
von: Krupp, Lars, et al.
Veröffentlicht: (2025)
von: Krupp, Lars, et al.
Veröffentlicht: (2025)
Balancing Domestic and Global Perspectives: Evaluating Dual-Calibration and LLM-Generated Nudges for Diverse News Recommendation
von: Sun, Ruixuan, et al.
Veröffentlicht: (2026)
von: Sun, Ruixuan, et al.
Veröffentlicht: (2026)
Emotion-Driven Personalized Recommendation for AI-Generated Content Using Multi-Modal Sentiment and Intent Analysis
von: Hu, Zheqi, et al.
Veröffentlicht: (2025)
von: Hu, Zheqi, et al.
Veröffentlicht: (2025)
LLMs as Policy-Agnostic Teammates: A Case Study in Human Proxy Design for Heterogeneous Agent Teams
von: Justus, Aju Ani, et al.
Veröffentlicht: (2025)
von: Justus, Aju Ani, et al.
Veröffentlicht: (2025)
Observations on LLMs for Telecom Domain: Capabilities and Limitations
von: Soman, Sumit, et al.
Veröffentlicht: (2023)
von: Soman, Sumit, et al.
Veröffentlicht: (2023)
Understanding Usage and Engagement in AI-Powered Scientific Research Tools: The Asta Interaction Dataset
von: Haddad, Dany, et al.
Veröffentlicht: (2026)
von: Haddad, Dany, et al.
Veröffentlicht: (2026)
User Preference Modeling for Conversational LLM Agents: Weak Rewards from Retrieval-Augmented Interaction
von: Hao, Yuren, et al.
Veröffentlicht: (2026)
von: Hao, Yuren, et al.
Veröffentlicht: (2026)
From Bytes to Biases: Investigating the Cultural Self-Perception of Large Language Models
von: Messner, Wolfgang, et al.
Veröffentlicht: (2023)
von: Messner, Wolfgang, et al.
Veröffentlicht: (2023)
Clinical Reasoning AI for Oncology Treatment Planning: A Multi-Specialty Case-Based Evaluation
von: Spiess, Philippe E., et al.
Veröffentlicht: (2026)
von: Spiess, Philippe E., et al.
Veröffentlicht: (2026)
RuleAlign: Making Large Language Models Better Physicians with Diagnostic Rule Alignment
von: Wang, Xiaohan, et al.
Veröffentlicht: (2024)
von: Wang, Xiaohan, et al.
Veröffentlicht: (2024)
EasyInstruct: An Easy-to-use Instruction Processing Framework for Large Language Models
von: Ou, Yixin, et al.
Veröffentlicht: (2024)
von: Ou, Yixin, et al.
Veröffentlicht: (2024)
OmniThink: Expanding Knowledge Boundaries in Machine Writing through Thinking
von: Xi, Zekun, et al.
Veröffentlicht: (2025)
von: Xi, Zekun, et al.
Veröffentlicht: (2025)
User Perception of Attention Visualizations: Effects on Interpretability Across Evidence-Based Medical Documents
von: Carvallo, Andrés, et al.
Veröffentlicht: (2025)
von: Carvallo, Andrés, et al.
Veröffentlicht: (2025)
Cascading Adaptors to Leverage English Data to Improve Performance of Question Answering for Low-Resource Languages
von: Pandya, Hariom A., et al.
Veröffentlicht: (2021)
von: Pandya, Hariom A., et al.
Veröffentlicht: (2021)
Bridging the Skills Gap: Evaluating an AI-Assisted Provider Platform to Support Care Providers with Empathetic Delivery of Protocolized Therapy
von: Kearns, William R., et al.
Veröffentlicht: (2024)
von: Kearns, William R., et al.
Veröffentlicht: (2024)
WildVis: Open Source Visualizer for Million-Scale Chat Logs in the Wild
von: Deng, Yuntian, et al.
Veröffentlicht: (2024)
von: Deng, Yuntian, et al.
Veröffentlicht: (2024)
RoTRAG: Rule of Thumb Reasoning for Conversation Harm Detection with Retrieval-Augmented Generation
von: Lee, Juhyeon, et al.
Veröffentlicht: (2026)
von: Lee, Juhyeon, et al.
Veröffentlicht: (2026)
Generative AI in the Construction Industry: A State-of-the-art Analysis
von: Taiwo, Ridwan, et al.
Veröffentlicht: (2024)
von: Taiwo, Ridwan, et al.
Veröffentlicht: (2024)
Towards End-to-End Open Conversational Machine Reading
von: Zhou, Sizhe, et al.
Veröffentlicht: (2022)
von: Zhou, Sizhe, et al.
Veröffentlicht: (2022)
Ähnliche Einträge
-
The Collaboration Gap in Human-AI Work
von: Vishwarupe, Varad, et al.
Veröffentlicht: (2026) -
From Rights to Rites: Expectations Management in Smart-Home AI
von: Vishwarupe, Varad, et al.
Veröffentlicht: (2026) -
The Evaluation Differential: When Frontier AI Models Recognise They Are Being Tested
von: Vishwarupe, Varad, et al.
Veröffentlicht: (2026) -
Deployment-Relevant Alignment Cannot Be Inferred from Model-Level Evaluation Alone
von: Vishwarupe, Varad, et al.
Veröffentlicht: (2026) -
From Sycophantic Consensus to Pluralistic Repair: Why AI Alignment Must Surface Disagreement
von: Vishwarupe, Varad, et al.
Veröffentlicht: (2026)