LLMs Corrupt Your Documents When You Delegate
Fuente:
arXiv
Saved in:
| Main Authors: | Laban, Philippe, Schnabel, Tobias, Neville, Jennifer |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
LLMs Get Lost In Multi-Turn Conversation
by: Laban, Philippe, et al.
Published: (2025)
by: Laban, Philippe, et al.
Published: (2025)
Can AI writing be salvaged? Mitigating Idiosyncrasies and Improving Human-AI Alignment in the Writing Process through Edits
by: Chakrabarty, Tuhin, et al.
Published: (2024)
by: Chakrabarty, Tuhin, et al.
Published: (2024)
Your Students Don't Use LLMs Like You Wish They Did
by: Kobler, Sebastian, et al.
Published: (2026)
by: Kobler, Sebastian, et al.
Published: (2026)
Art or Artifice? Large Language Models and the False Promise of Creativity
by: Chakrabarty, Tuhin, et al.
Published: (2023)
by: Chakrabarty, Tuhin, et al.
Published: (2023)
Do We Talk to Robots Like Therapists, and Do They Respond Accordingly? Language Alignment in AI Emotional Support
by: Chiang, Sophie, et al.
Published: (2025)
by: Chiang, Sophie, et al.
Published: (2025)
LEXI: Large Language Models Experimentation Interface
by: Laban, Guy, et al.
Published: (2024)
by: Laban, Guy, et al.
Published: (2024)
AI, Take the Wheel: What Drives Delegation and Trust in Human-Computer Cooperative Question Answering?
by: Gor, Maharshi, et al.
Published: (2026)
by: Gor, Maharshi, et al.
Published: (2026)
What People Share With a Robot When Feeling Lonely and Stressed and How It Helps Over Time
by: Laban, Guy, et al.
Published: (2025)
by: Laban, Guy, et al.
Published: (2025)
When Can We Trust LLMs in Mental Health? Large-Scale Benchmarks for Reliable LLM Evaluation
by: Badawi, Abeer, et al.
Published: (2025)
by: Badawi, Abeer, et al.
Published: (2025)
The Richest Paradigm You're Not Using: Commercial Videogames at the Intersection of Human-Computer Interaction and Cognitive Science
by: Munneke, Jaap, et al.
Published: (2026)
by: Munneke, Jaap, et al.
Published: (2026)
LLMs as Workers in Human-Computational Algorithms? Replicating Crowdsourcing Pipelines with LLMs
by: Wu, Tongshuang, et al.
Published: (2023)
by: Wu, Tongshuang, et al.
Published: (2023)
When Large Language Models are Reliable for Judging Empathic Communication
by: Kumar, Aakriti, et al.
Published: (2025)
by: Kumar, Aakriti, et al.
Published: (2025)
DiscussLLM: Teaching Large Language Models When to Speak
by: Patel, Deep Anil, et al.
Published: (2025)
by: Patel, Deep Anil, et al.
Published: (2025)
"Harmless to You, Hurtful to Me!": Investigating the Detection of Toxic Languages Grounded in the Perspective of Youth
by: Li, Yaqiong, et al.
Published: (2025)
by: Li, Yaqiong, et al.
Published: (2025)
Sniff AI: Is My 'Spicy' Your 'Spicy'? Exploring LLM's Perceptual Alignment with Human Smell Experiences
by: Zhong, Shu, et al.
Published: (2024)
by: Zhong, Shu, et al.
Published: (2024)
Search Engines in an AI Era: The False Promise of Factual and Verifiable Source-Cited Responses
by: Venkit, Pranav Narayanan, et al.
Published: (2024)
by: Venkit, Pranav Narayanan, et al.
Published: (2024)
"You tell me": A Dataset of GPT-4-Based Behaviour Change Support Conversations
by: Meyer, Selina, et al.
Published: (2024)
by: Meyer, Selina, et al.
Published: (2024)
When Avatars Have Personality: Effects on Engagement and Communication in Immersive Medical Training
by: Dollis, Julia S., et al.
Published: (2025)
by: Dollis, Julia S., et al.
Published: (2025)
CALYPSO: LLMs as Dungeon Masters' Assistants
by: Zhu, Andrew, et al.
Published: (2023)
by: Zhu, Andrew, et al.
Published: (2023)
ConsistencyAI: A Benchmark to Assess LLMs' Factual Consistency When Responding to Different Demographic Groups
by: Banyas, Peter, et al.
Published: (2025)
by: Banyas, Peter, et al.
Published: (2025)
Large Language Models Help Humans Verify Truthfulness -- Except When They Are Convincingly Wrong
by: Si, Chenglei, et al.
Published: (2023)
by: Si, Chenglei, et al.
Published: (2023)
STL: Still Tricky Logic (for System Validation, Even When Showing Your Work)
by: Hurley, Isabelle, et al.
Published: (2024)
by: Hurley, Isabelle, et al.
Published: (2024)
Robots in the Middle: Evaluating LLMs in Dispute Resolution
by: Tan, Jinzhe, et al.
Published: (2024)
by: Tan, Jinzhe, et al.
Published: (2024)
Pragmatics beyond humans: meaning, communication, and LLMs
by: Gvoždiak, Vít
Published: (2025)
by: Gvoždiak, Vít
Published: (2025)
LLMs in Qualitative Research: Opportunities, Limitations, and Practical Considerations
by: Salgado, Henry, et al.
Published: (2026)
by: Salgado, Henry, et al.
Published: (2026)
Revisiting OPRO: The Limitations of Small-Scale LLMs as Optimizers
by: Zhang, Tuo, et al.
Published: (2024)
by: Zhang, Tuo, et al.
Published: (2024)
Personas with Attitudes: Controlling LLMs for Diverse Data Annotation
by: Fröhling, Leon, et al.
Published: (2024)
by: Fröhling, Leon, et al.
Published: (2024)
Can ChatGPT Read Who You Are?
by: Derner, Erik, et al.
Published: (2023)
by: Derner, Erik, et al.
Published: (2023)
Generating Educational Materials with Different Levels of Readability using LLMs
by: Huang, Chieh-Yang, et al.
Published: (2024)
by: Huang, Chieh-Yang, et al.
Published: (2024)
Plan-Then-Execute: An Empirical Study of User Trust and Team Performance When Using LLM Agents As A Daily Assistant
by: He, Gaole, et al.
Published: (2025)
by: He, Gaole, et al.
Published: (2025)
Chart-to-Experience: Benchmarking Multimodal LLMs for Predicting Experiential Impact of Charts
by: Kim, Seon Gyeom, et al.
Published: (2025)
by: Kim, Seon Gyeom, et al.
Published: (2025)
Characterizing Similarities and Divergences in Conversational Tones in Humans and LLMs by Sampling with People
by: Huang, Dun-Ming, et al.
Published: (2024)
by: Huang, Dun-Ming, et al.
Published: (2024)
Do LLMs Truly Benefit from Longer Context in Automatic Post-Editing?
by: Kim, Ahrii, et al.
Published: (2026)
by: Kim, Ahrii, et al.
Published: (2026)
Translation Analytics for Freelancers II: Benchmarking Local LLMs for Confidential Translation Workflows
by: Balashov, Yuri, et al.
Published: (2026)
by: Balashov, Yuri, et al.
Published: (2026)
The LLM Effect: Are Humans Truly Using LLMs, or Are They Being Influenced By Them Instead?
by: Choi, Alexander S., et al.
Published: (2024)
by: Choi, Alexander S., et al.
Published: (2024)
Risks and NLP Design: A Case Study on Procedural Document QA
by: Haduong, Nikita, et al.
Published: (2024)
by: Haduong, Nikita, et al.
Published: (2024)
WordCraft: Scaffolding the Keyword Method for L2 Vocabulary Learning with Multimodal LLMs
by: Shao, Yuheng, et al.
Published: (2026)
by: Shao, Yuheng, et al.
Published: (2026)
Leveraging Small LLMs for Argument Mining in Education: Argument Component Identification, Classification, and Assessment
by: Favero, Lucile, et al.
Published: (2025)
by: Favero, Lucile, et al.
Published: (2025)
ReSpark: Leveraging Previous Data Reports as References to Generate New Reports with LLMs
by: Tian, Yuan, et al.
Published: (2025)
by: Tian, Yuan, et al.
Published: (2025)
S-DAT: A Multilingual, GenAI-Driven Framework for Automated Divergent Thinking Assessment
by: Haase, Jennifer, et al.
Published: (2025)
by: Haase, Jennifer, et al.
Published: (2025)
Similar Items
-
LLMs Get Lost In Multi-Turn Conversation
by: Laban, Philippe, et al.
Published: (2025) -
Can AI writing be salvaged? Mitigating Idiosyncrasies and Improving Human-AI Alignment in the Writing Process through Edits
by: Chakrabarty, Tuhin, et al.
Published: (2024) -
Your Students Don't Use LLMs Like You Wish They Did
by: Kobler, Sebastian, et al.
Published: (2026) -
Art or Artifice? Large Language Models and the False Promise of Creativity
by: Chakrabarty, Tuhin, et al.
Published: (2023) -
Do We Talk to Robots Like Therapists, and Do They Respond Accordingly? Language Alignment in AI Emotional Support
by: Chiang, Sophie, et al.
Published: (2025)