Gespeichert in:
| Hauptverfasser: | Volkova, Svitlana, Dupree, Will, Kao, Hsien-Te, Bautista, Peter, Ganberg, Gabe, Beaubien, Jeff, Cassani, Laura |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | https://arxiv.org/abs/2511.21749 |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Building Resilient Information Ecosystems: Large LLM-Generated Dataset of Persuasion Attacks
von: Kao, Hsien-Te, et al.
Veröffentlicht: (2025)
von: Kao, Hsien-Te, et al.
Veröffentlicht: (2025)
Cross-Disciplinary Knowledge Retrieval and Synthesis: A Compound AI Architecture for Scientific Discovery
von: Volkova, Svitlana, et al.
Veröffentlicht: (2025)
von: Volkova, Svitlana, et al.
Veröffentlicht: (2025)
Imperfectly Cooperative Human-AI Interactions: Comparing the Impacts of Human and AI Attributes in Simulated and User Studies
von: Cohen, Myke C., et al.
Veröffentlicht: (2026)
von: Cohen, Myke C., et al.
Veröffentlicht: (2026)
Exploring Big Five Personality and AI Capability Effects in LLM-Simulated Negotiation Dialogues
von: Cohen, Myke C., et al.
Veröffentlicht: (2025)
von: Cohen, Myke C., et al.
Veröffentlicht: (2025)
Towards Safer Online Spaces: Simulating and Assessing Intervention Strategies for Eating Disorder Discussions
von: Penafiel, Louis, et al.
Veröffentlicht: (2024)
von: Penafiel, Louis, et al.
Veröffentlicht: (2024)
Density-Guided Response Optimization: Community-Grounded Alignment via Implicit Acceptance Signals
von: Gerard, Patrick, et al.
Veröffentlicht: (2026)
von: Gerard, Patrick, et al.
Veröffentlicht: (2026)
Exploratory Models of Human-AI Teams: Leveraging Human Digital Twins to Investigate Trust Development
von: Nguyen, Daniel, et al.
Veröffentlicht: (2024)
von: Nguyen, Daniel, et al.
Veröffentlicht: (2024)
Community-Aligned Behavior Under Uncertainty: Evidence of Epistemic Stance Transfer in LLMs
von: Gerard, Patrick, et al.
Veröffentlicht: (2025)
von: Gerard, Patrick, et al.
Veröffentlicht: (2025)
Stop Tracking Me! Proactive Defense Against Attribute Inference Attack in LLMs
von: Yan, Dong, et al.
Veröffentlicht: (2026)
von: Yan, Dong, et al.
Veröffentlicht: (2026)
‘Who Is Afraid of Fairenesse or Wanton Ladies Appearing in Their Barenesse?’: Laughing at Female Desire in Early Modern English Reception of the Myth of the Trojan War☆
von: Evgeniia Ganberg
Veröffentlicht: (2024)
von: Evgeniia Ganberg
Veröffentlicht: (2024)
Uncovering the Persuasive Fingerprint of LLMs in Jailbreaking Attacks
von: Noughabi, Havva Alizadeh, et al.
Veröffentlicht: (2025)
von: Noughabi, Havva Alizadeh, et al.
Veröffentlicht: (2025)
Iterative Prompting with Persuasion Skills in Jailbreaking Large Language Models
von: Ke, Shih-Wen, et al.
Veröffentlicht: (2025)
von: Ke, Shih-Wen, et al.
Veröffentlicht: (2025)
Redefining Proactivity for Information Seeking Dialogue
von: Lee, Jing Yang, et al.
Veröffentlicht: (2024)
von: Lee, Jing Yang, et al.
Veröffentlicht: (2024)
MALicious INTent Dataset and Inoculating LLMs for Enhanced Disinformation Detection
von: Modzelewski, Arkadiusz, et al.
Veröffentlicht: (2026)
von: Modzelewski, Arkadiusz, et al.
Veröffentlicht: (2026)
Can AI-Generated Persuasion Be Detected? Persuaficial Benchmark and AI vs. Human Linguistic Differences
von: Modzelewski, Arkadiusz, et al.
Veröffentlicht: (2026)
von: Modzelewski, Arkadiusz, et al.
Veröffentlicht: (2026)
ValueScope: Unveiling Implicit Norms and Values via Return Potential Model of Social Interactions
von: Park, Chan Young, et al.
Veröffentlicht: (2024)
von: Park, Chan Young, et al.
Veröffentlicht: (2024)
Measuring and Improving Persuasiveness of Large Language Models
von: Singh, Somesh, et al.
Veröffentlicht: (2024)
von: Singh, Somesh, et al.
Veröffentlicht: (2024)
Evaluating OpenAI GPT Models for Translation of Endangered Uralic Languages: A Comparison of Reasoning and Non-Reasoning Architectures
von: Tereshchenko, Yehor, et al.
Veröffentlicht: (2025)
von: Tereshchenko, Yehor, et al.
Veröffentlicht: (2025)
Commercial Persuasion in AI-Mediated Conversations
von: Salvi, Francesco, et al.
Veröffentlicht: (2026)
von: Salvi, Francesco, et al.
Veröffentlicht: (2026)
Measuring Opinion Bias and Sycophancy via LLM-based Persuasion
von: Nogueira, Rodrigo, et al.
Veröffentlicht: (2026)
von: Nogueira, Rodrigo, et al.
Veröffentlicht: (2026)
LLM-Based Adversarial Persuasion Attacks on Fact-Checking Systems
von: Leite, João A., et al.
Veröffentlicht: (2026)
von: Leite, João A., et al.
Veröffentlicht: (2026)
A Survey on Proactive Defense Strategies Against Misinformation in Large Language Models
von: Liu, Shuliang, et al.
Veröffentlicht: (2025)
von: Liu, Shuliang, et al.
Veröffentlicht: (2025)
Defending Against Social Engineering Attacks in the Age of LLMs
von: Ai, Lin, et al.
Veröffentlicht: (2024)
von: Ai, Lin, et al.
Veröffentlicht: (2024)
You Need Better Attention Priors
von: Litman, Elon, et al.
Veröffentlicht: (2026)
von: Litman, Elon, et al.
Veröffentlicht: (2026)
Reviving Any-Subset Autoregressive Models with Principled Parallel Sampling and Speculative Decoding
von: Guo, Gabe, et al.
Veröffentlicht: (2025)
von: Guo, Gabe, et al.
Veröffentlicht: (2025)
The Levers of Political Persuasion with Conversational AI
von: Hackenburg, Kobi, et al.
Veröffentlicht: (2025)
von: Hackenburg, Kobi, et al.
Veröffentlicht: (2025)
Towards Detecting Persuasion on Social Media: From Model Development to Insights on Persuasion Strategies
von: Meguellati, Elyas, et al.
Veröffentlicht: (2025)
von: Meguellati, Elyas, et al.
Veröffentlicht: (2025)
Improving QA Model Performance with Cartographic Inoculation
von: Chen, Allen, et al.
Veröffentlicht: (2024)
von: Chen, Allen, et al.
Veröffentlicht: (2024)
PVP: An Image Dataset for Personalized Visual Persuasion with Persuasion Strategies, Viewer Characteristics, and Persuasiveness Ratings
von: Kim, Junseo, et al.
Veröffentlicht: (2025)
von: Kim, Junseo, et al.
Veröffentlicht: (2025)
AutoDefense: Multi-Agent LLM Defense against Jailbreak Attacks
von: Zeng, Yifan, et al.
Veröffentlicht: (2024)
von: Zeng, Yifan, et al.
Veröffentlicht: (2024)
Detecting Winning Arguments with Large Language Models and Persuasion Strategies
von: Labruna, Tiziano, et al.
Veröffentlicht: (2026)
von: Labruna, Tiziano, et al.
Veröffentlicht: (2026)
Measuring and Benchmarking Large Language Models' Capabilities to Generate Persuasive Language
von: Pauli, Amalie Brogaard, et al.
Veröffentlicht: (2024)
von: Pauli, Amalie Brogaard, et al.
Veröffentlicht: (2024)
Adversarial Attacks and Defense for Conversation Entailment Task
von: Yang, Zhenning, et al.
Veröffentlicht: (2024)
von: Yang, Zhenning, et al.
Veröffentlicht: (2024)
Teaching Models to Balance Resisting and Accepting Persuasion
von: Stengel-Eskin, Elias, et al.
Veröffentlicht: (2024)
von: Stengel-Eskin, Elias, et al.
Veröffentlicht: (2024)
Verification Required: The Impact of Information Credibility on AI Persuasion
von: Mahmud, Saaduddin, et al.
Veröffentlicht: (2026)
von: Mahmud, Saaduddin, et al.
Veröffentlicht: (2026)
AI for Service: Proactive Assistance with AI Glasses
von: Wen, Zichen, et al.
Veröffentlicht: (2025)
von: Wen, Zichen, et al.
Veröffentlicht: (2025)
PersuasiveToM: A Benchmark for Evaluating Machine Theory of Mind in Persuasive Dialogues
von: Yu, Fangxu, et al.
Veröffentlicht: (2025)
von: Yu, Fangxu, et al.
Veröffentlicht: (2025)
UniGuardian: A Unified Defense for Detecting Prompt Injection, Backdoor Attacks and Adversarial Attacks in Large Language Models
von: Lin, Huawei, et al.
Veröffentlicht: (2025)
von: Lin, Huawei, et al.
Veröffentlicht: (2025)
Defense Against Syntactic Textual Backdoor Attacks with Token Substitution
von: Li, Xinglin, et al.
Veröffentlicht: (2024)
von: Li, Xinglin, et al.
Veröffentlicht: (2024)
The Best Defense is Attack: Repairing Semantics in Textual Adversarial Examples
von: Yang, Heng, et al.
Veröffentlicht: (2023)
von: Yang, Heng, et al.
Veröffentlicht: (2023)
Ähnliche Einträge
-
Building Resilient Information Ecosystems: Large LLM-Generated Dataset of Persuasion Attacks
von: Kao, Hsien-Te, et al.
Veröffentlicht: (2025) -
Cross-Disciplinary Knowledge Retrieval and Synthesis: A Compound AI Architecture for Scientific Discovery
von: Volkova, Svitlana, et al.
Veröffentlicht: (2025) -
Imperfectly Cooperative Human-AI Interactions: Comparing the Impacts of Human and AI Attributes in Simulated and User Studies
von: Cohen, Myke C., et al.
Veröffentlicht: (2026) -
Exploring Big Five Personality and AI Capability Effects in LLM-Simulated Negotiation Dialogues
von: Cohen, Myke C., et al.
Veröffentlicht: (2025) -
Towards Safer Online Spaces: Simulating and Assessing Intervention Strategies for Eating Disorder Discussions
von: Penafiel, Louis, et al.
Veröffentlicht: (2024)