Detection of adversarial intent in Human-AI teams using LLMs
Fuente:
arXiv
Guardado en:
| Autores principales: | Musaffar, Abed K., Singh, Ambuj, Bullo, Francesco |
|---|---|
| Formato: | Preprint |
| Publicado: |
2026
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Learning to Lie: Reinforcement Learning Attacks Damage Human-AI Teams and Teams of LLMs
por: Musaffar, Abed Kareem, et al.
Publicado: (2025)
por: Musaffar, Abed Kareem, et al.
Publicado: (2025)
The case for delegated AI autonomy for Human AI teaming in healthcare
por: Jia, Yan, et al.
Publicado: (2025)
por: Jia, Yan, et al.
Publicado: (2025)
Human-Centered Explainable AI for Security Enhancement: A Deep Intrusion Detection Framework
por: Ayan, Md Muntasir Jahid, et al.
Publicado: (2026)
por: Ayan, Md Muntasir Jahid, et al.
Publicado: (2026)
On the Utility of Accounting for Human Beliefs about AI Intention in Human-AI Collaboration
por: Yu, Guanghui, et al.
Publicado: (2024)
por: Yu, Guanghui, et al.
Publicado: (2024)
Toward Human-AI Complementarity Across Diverse Tasks
por: Xu, Yuzheng, et al.
Publicado: (2026)
por: Xu, Yuzheng, et al.
Publicado: (2026)
Limited but consistent gains in adversarial robustness by co-training object recognition models with human EEG
por: Guo, Manshan, et al.
Publicado: (2024)
por: Guo, Manshan, et al.
Publicado: (2024)
Human-in-the-Loop AI for Cheating Ring Detection
por: Shih, Yong-Siang, et al.
Publicado: (2024)
por: Shih, Yong-Siang, et al.
Publicado: (2024)
Almost AI, Almost Human: The Challenge of Detecting AI-Polished Writing
por: Saha, Shoumik, et al.
Publicado: (2025)
por: Saha, Shoumik, et al.
Publicado: (2025)
AI-Induced Human Responsibility (AIHR) in AI-Human teams
por: Nyilasy, Greg, et al.
Publicado: (2026)
por: Nyilasy, Greg, et al.
Publicado: (2026)
Human-AI Collaborative Uncertainty Quantification
por: Noorani, Sima, et al.
Publicado: (2025)
por: Noorani, Sima, et al.
Publicado: (2025)
CREW: Facilitating Human-AI Teaming Research
por: Zhang, Lingyu, et al.
Publicado: (2024)
por: Zhang, Lingyu, et al.
Publicado: (2024)
Decoding Human Emotions: Analyzing Multi-Channel EEG Data using LSTM Networks
por: Sateesh, Shyam K, et al.
Publicado: (2024)
por: Sateesh, Shyam K, et al.
Publicado: (2024)
AI Agents for Inventory Control: Human-LLM-OR Complementarity
por: Baek, Jackie, et al.
Publicado: (2026)
por: Baek, Jackie, et al.
Publicado: (2026)
Learning to Decide with AI Assistance under Human-Alignment
por: Benz, Nina Corvelo, et al.
Publicado: (2026)
por: Benz, Nina Corvelo, et al.
Publicado: (2026)
Rationalize: Shared Semantic Reasoning for Human-AI Alignment
por: Dasgupta, Aritra, et al.
Publicado: (2026)
por: Dasgupta, Aritra, et al.
Publicado: (2026)
Optimizing Delegation in Collaborative Human-AI Hybrid Teams
por: Fuchs, Andrew, et al.
Publicado: (2024)
por: Fuchs, Andrew, et al.
Publicado: (2024)
MedSyn: Enhancing Diagnostics with Human-AI Collaboration
por: Sayin, Burcu, et al.
Publicado: (2025)
por: Sayin, Burcu, et al.
Publicado: (2025)
A No Free Lunch Theorem for Human-AI Collaboration
por: Peng, Kenny, et al.
Publicado: (2024)
por: Peng, Kenny, et al.
Publicado: (2024)
Agentic AI for Commercial Insurance Underwriting with Adversarial Self-Critique
por: Roy, Joyjit, et al.
Publicado: (2026)
por: Roy, Joyjit, et al.
Publicado: (2026)
Human-Computer Interaction and Human-AI Collaboration in Advanced Air Mobility: A Comprehensive Review
por: Sagirli, Fatma Yamac, et al.
Publicado: (2024)
por: Sagirli, Fatma Yamac, et al.
Publicado: (2024)
ABScribe: Rapid Exploration & Organization of Multiple Writing Variations in Human-AI Co-Writing Tasks using Large Language Models
por: Reza, Mohi, et al.
Publicado: (2023)
por: Reza, Mohi, et al.
Publicado: (2023)
Evaluation of LLMs-based Hidden States as Author Representations for Psychological Human-Centered NLP Tasks
por: Soni, Nikita, et al.
Publicado: (2025)
por: Soni, Nikita, et al.
Publicado: (2025)
Improving Health Professionals' Onboarding with AI and XAI for Trustworthy Human-AI Collaborative Decision Making
por: Lee, Min Hun, et al.
Publicado: (2024)
por: Lee, Min Hun, et al.
Publicado: (2024)
Align When They Want, Complement When They Need! Human-Centered Ensembles for Adaptive Human-AI Collaboration
por: Amin, Hasan, et al.
Publicado: (2026)
por: Amin, Hasan, et al.
Publicado: (2026)
Epistemology gives a Future to Complementarity in Human-AI Interactions
por: Ferrario, Andrea, et al.
Publicado: (2026)
por: Ferrario, Andrea, et al.
Publicado: (2026)
Reversing the Lens: Using Explainable AI to Understand Human Expertise
por: Rahman, Roussel, et al.
Publicado: (2025)
por: Rahman, Roussel, et al.
Publicado: (2025)
LLMs May Not Be Human-Level Players, But They Can Be Testers: Measuring Game Difficulty with LLM Agents
por: Xiao, Chang, et al.
Publicado: (2024)
por: Xiao, Chang, et al.
Publicado: (2024)
From Accuracy to Readiness: Metrics and Benchmarks for Human-AI Decision-Making
por: Lee, Min Hun
Publicado: (2026)
por: Lee, Min Hun
Publicado: (2026)
Interactive Example-based Explanations to Improve Health Professionals' Onboarding with AI for Human-AI Collaborative Decision Making
por: Lee, Min Hun, et al.
Publicado: (2024)
por: Lee, Min Hun, et al.
Publicado: (2024)
AI, Meet Human: Learning Paradigms for Hybrid Decision Making Systems
por: Punzi, Clara, et al.
Publicado: (2024)
por: Punzi, Clara, et al.
Publicado: (2024)
The Model Mastery Lifecycle: A Framework for Designing Human-AI Interaction
por: Chignell, Mark, et al.
Publicado: (2024)
por: Chignell, Mark, et al.
Publicado: (2024)
LLMs as Policy-Agnostic Teammates: A Case Study in Human Proxy Design for Heterogeneous Agent Teams
por: Justus, Aju Ani, et al.
Publicado: (2025)
por: Justus, Aju Ani, et al.
Publicado: (2025)
Benchmarking System Dynamics AI Assistants: Cloud Versus Local LLMs on CLD Extraction and Discussion
por: Leitch, Terry
Publicado: (2026)
por: Leitch, Terry
Publicado: (2026)
Towards Uncertainty Aware Task Delegation and Human-AI Collaborative Decision-Making
por: Lee, Min Hun, et al.
Publicado: (2025)
por: Lee, Min Hun, et al.
Publicado: (2025)
Predictive AI Can Support Human Learning while Preserving Error Diversity
por: He, Vivianna Fang, et al.
Publicado: (2025)
por: He, Vivianna Fang, et al.
Publicado: (2025)
Cognitive Exoskeleton: Augmenting Human Cognition with an AI-Mediated Intelligent Visual Feedback
por: Xu, Songlin, et al.
Publicado: (2025)
por: Xu, Songlin, et al.
Publicado: (2025)
Co-Creative Learning via Metropolis-Hastings Interaction between Humans and AI
por: Okumura, Ryota, et al.
Publicado: (2025)
por: Okumura, Ryota, et al.
Publicado: (2025)
RuleEdit: Failure-Guided Human-AI Model Editing with Prospective Impact Preview
por: Lee, Min Hun, et al.
Publicado: (2026)
por: Lee, Min Hun, et al.
Publicado: (2026)
Towards User-Focused Research in Training Data Attribution for Human-Centered Explainable AI
por: Nguyen, Elisa, et al.
Publicado: (2024)
por: Nguyen, Elisa, et al.
Publicado: (2024)
Generative AI-Driven Human Digital Twin in IoT-Healthcare: A Comprehensive Survey
por: Chen, Jiayuan, et al.
Publicado: (2024)
por: Chen, Jiayuan, et al.
Publicado: (2024)
Ejemplares similares
-
Learning to Lie: Reinforcement Learning Attacks Damage Human-AI Teams and Teams of LLMs
por: Musaffar, Abed Kareem, et al.
Publicado: (2025) -
The case for delegated AI autonomy for Human AI teaming in healthcare
por: Jia, Yan, et al.
Publicado: (2025) -
Human-Centered Explainable AI for Security Enhancement: A Deep Intrusion Detection Framework
por: Ayan, Md Muntasir Jahid, et al.
Publicado: (2026) -
On the Utility of Accounting for Human Beliefs about AI Intention in Human-AI Collaboration
por: Yu, Guanghui, et al.
Publicado: (2024) -
Toward Human-AI Complementarity Across Diverse Tasks
por: Xu, Yuzheng, et al.
Publicado: (2026)