Learning to Lie: Reinforcement Learning Attacks Damage Human-AI Teams and Teams of LLMs

Fuente: arXiv
Enregistré dans:
Détails bibliographiques
Auteurs principaux: Musaffar, Abed Kareem, Gokhale, Anand, Zeng, Sirui, Tadayon, Rasta, Yan, Xifeng, Singh, Ambuj, Bullo, Francesco
Format: Preprint
Publié: 2025
Sujets:
Accès en ligne:
Tags: Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!