AREG: Adversarial Resource Extraction Game for Evaluating Persuasion and Resistance in Large Language Models
Fuente:
arXiv
Guardado en:
| Autores principales: | Sakhawat, Adib, Sadab, Fardeen |
|---|---|
| Formato: | Preprint |
| Publicado: |
2026
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
AIDG: A Formal Decomposition of Information Extraction and Containment Asymmetries in Multi-Turn LLM Dialogue
por: Sakhawat, Adib, et al.
Publicado: (2026)
por: Sakhawat, Adib, et al.
Publicado: (2026)
Persuasion Games using Large Language Models
por: Ramani, Ganesh Prasath, et al.
Publicado: (2024)
por: Ramani, Ganesh Prasath, et al.
Publicado: (2024)
Political Alignment in Large Language Models: A Multidimensional Audit of Psychometric Identity and Behavioral Bias
por: Sakhawat, Adib, et al.
Publicado: (2026)
por: Sakhawat, Adib, et al.
Publicado: (2026)
Coordinates of Capability: A Unified MTMM-Geometric Framework for LLM Evaluation
por: Sakhawat, Adib, et al.
Publicado: (2026)
por: Sakhawat, Adib, et al.
Publicado: (2026)
Gradient Masters at BLP-2025 Task 1: Advancing Low-Resource NLP for Bengali using Ensemble-Based Adversarial Training for Hate Speech Detection
por: Hoque, Syed Mohaiminul, et al.
Publicado: (2025)
por: Hoque, Syed Mohaiminul, et al.
Publicado: (2025)
On the Adaptive Psychological Persuasion of Large Language Models
por: Ju, Tianjie, et al.
Publicado: (2025)
por: Ju, Tianjie, et al.
Publicado: (2025)
Moral Persuasion in Large Language Models: Evaluating Susceptibility and Ethical Alignment
por: Huang, Allison, et al.
Publicado: (2024)
por: Huang, Allison, et al.
Publicado: (2024)
Evaluating from Benign to Dynamic Adversarial: A Squid Game for Large Language Models
por: Chen, Zijian, et al.
Publicado: (2025)
por: Chen, Zijian, et al.
Publicado: (2025)
Measuring and Improving Persuasiveness of Large Language Models
por: Singh, Somesh, et al.
Publicado: (2024)
por: Singh, Somesh, et al.
Publicado: (2024)
Detecting Winning Arguments with Large Language Models and Persuasion Strategies
por: Labruna, Tiziano, et al.
Publicado: (2026)
por: Labruna, Tiziano, et al.
Publicado: (2026)
Teaching Models to Balance Resisting and Accepting Persuasion
por: Stengel-Eskin, Elias, et al.
Publicado: (2024)
por: Stengel-Eskin, Elias, et al.
Publicado: (2024)
Benchmarking Political Persuasion Risks Across Frontier Large Language Models
por: Chen, Zhongren, et al.
Publicado: (2026)
por: Chen, Zhongren, et al.
Publicado: (2026)
Assessing Large Language Models for Medical QA: Zero-Shot and LLM-as-a-Judge Evaluation
por: Adib, Shefayat E Shams, et al.
Publicado: (2026)
por: Adib, Shefayat E Shams, et al.
Publicado: (2026)
Signature vs. Substance: Evaluating the Balance of Adversarial Resistance and Linguistic Quality in Watermarking Large Language Models
por: Guo, William, et al.
Publicado: (2025)
por: Guo, William, et al.
Publicado: (2025)
When Words Don't Mean What They Say: Figurative Understanding in Bengali Idioms
por: Sakhawat, Adib, et al.
Publicado: (2026)
por: Sakhawat, Adib, et al.
Publicado: (2026)
Measuring and Benchmarking Large Language Models' Capabilities to Generate Persuasive Language
por: Pauli, Amalie Brogaard, et al.
Publicado: (2024)
por: Pauli, Amalie Brogaard, et al.
Publicado: (2024)
Do Large Language Models Have a Planning Theory of Mind? Evidence from MindGames: a Multi-Step Persuasion Task
por: Moore, Jared, et al.
Publicado: (2025)
por: Moore, Jared, et al.
Publicado: (2025)
A Hybrid Theory and Data-driven Approach to Persuasion Detection with Large Language Models
por: Hoang, Gia Bao, et al.
Publicado: (2025)
por: Hoang, Gia Bao, et al.
Publicado: (2025)
Leveraging Open-Source Large Language Models for Clinical Information Extraction in Resource-Constrained Settings
por: Builtjes, Luc, et al.
Publicado: (2025)
por: Builtjes, Luc, et al.
Publicado: (2025)
Can You Trick the Grader? Adversarial Persuasion of LLM Judges
por: Hwang, Yerin, et al.
Publicado: (2025)
por: Hwang, Yerin, et al.
Publicado: (2025)
PersuasiveToM: A Benchmark for Evaluating Machine Theory of Mind in Persuasive Dialogues
por: Yu, Fangxu, et al.
Publicado: (2025)
por: Yu, Fangxu, et al.
Publicado: (2025)
Event Extraction in Large Language Model
por: Li, Bobo, et al.
Publicado: (2025)
por: Li, Bobo, et al.
Publicado: (2025)
Under the Influence: Quantifying Persuasion and Vigilance in Large Language Models
por: Robinson, Sasha, et al.
Publicado: (2026)
por: Robinson, Sasha, et al.
Publicado: (2026)
Iterative Prompting with Persuasion Skills in Jailbreaking Large Language Models
por: Ke, Shih-Wen, et al.
Publicado: (2025)
por: Ke, Shih-Wen, et al.
Publicado: (2025)
Evaluating and Adapting Large Language Models to Represent Folktales in Low-Resource Languages
por: Meaney, JA, et al.
Publicado: (2024)
por: Meaney, JA, et al.
Publicado: (2024)
Persuasion with Large Language Models: A Survey of Empirical Evidence, Study Methodologies, and Ethical Implications
por: Noels, Sander, et al.
Publicado: (2024)
por: Noels, Sander, et al.
Publicado: (2024)
A Framework to Assess the Persuasion Risks Large Language Model Chatbots Pose to Democratic Societies
por: Chen, Zhongren, et al.
Publicado: (2025)
por: Chen, Zhongren, et al.
Publicado: (2025)
Investigating Persuasion Techniques in Arabic: An Empirical Study Leveraging Large Language Models
por: Alzahrani, Abdurahmman, et al.
Publicado: (2024)
por: Alzahrani, Abdurahmman, et al.
Publicado: (2024)
The Dark Patterns of Personalized Persuasion in Large Language Models: Exposing Persuasive Linguistic Features for Big Five Personality Traits in LLMs Responses
por: Mieleszczenko-Kowszewicz, Wiktoria, et al.
Publicado: (2024)
por: Mieleszczenko-Kowszewicz, Wiktoria, et al.
Publicado: (2024)
MedicalBench: Evaluating Large Language Models Toward Improved Medical Concept Extraction
por: Yang, Zhichao, et al.
Publicado: (2026)
por: Yang, Zhichao, et al.
Publicado: (2026)
Evaluating Language Models' Evaluations of Games
por: Collins, Katherine M., et al.
Publicado: (2025)
por: Collins, Katherine M., et al.
Publicado: (2025)
DIRI: Adversarial Patient Reidentification with Large Language Models for Evaluating Clinical Text Anonymization
por: Morris, John X., et al.
Publicado: (2024)
por: Morris, John X., et al.
Publicado: (2024)
BenHalluEval: A Multi-Task Hallucination Evaluation Framework for Large Language Models on Bengali
por: Adib, Shefayat E Shams, et al.
Publicado: (2026)
por: Adib, Shefayat E Shams, et al.
Publicado: (2026)
BURExtract-Llama: An LLM for Clinical Concept Extraction in Breast Ultrasound Reports
por: Chen, Yuxuan, et al.
Publicado: (2024)
por: Chen, Yuxuan, et al.
Publicado: (2024)
LLMsPark: A Benchmark for Evaluating Large Language Models in Strategic Gaming Contexts
por: Chen, Junhao, et al.
Publicado: (2025)
por: Chen, Junhao, et al.
Publicado: (2025)
AMONGAGENTS: Evaluating Large Language Models in the Interactive Text-Based Social Deduction Game
por: Chi, Yizhou, et al.
Publicado: (2024)
por: Chi, Yizhou, et al.
Publicado: (2024)
Eka-Eval: An Evaluation Framework for Low-Resource Multilingual Large Language Models
por: Sinha, Samridhi Raj, et al.
Publicado: (2025)
por: Sinha, Samridhi Raj, et al.
Publicado: (2025)
RPGBENCH: Evaluating Large Language Models as Role-Playing Game Engines
por: Yu, Pengfei, et al.
Publicado: (2025)
por: Yu, Pengfei, et al.
Publicado: (2025)
Mind What You Ask For: Emotional and Rational Faces of Persuasion by Large Language Models
por: Mieleszczenko-Kowszewicz, Wiktoria, et al.
Publicado: (2025)
por: Mieleszczenko-Kowszewicz, Wiktoria, et al.
Publicado: (2025)
Do Vision-Language Models Understand Visual Persuasiveness?
por: Park, Gyuwon
Publicado: (2025)
por: Park, Gyuwon
Publicado: (2025)
Ejemplares similares
-
AIDG: A Formal Decomposition of Information Extraction and Containment Asymmetries in Multi-Turn LLM Dialogue
por: Sakhawat, Adib, et al.
Publicado: (2026) -
Persuasion Games using Large Language Models
por: Ramani, Ganesh Prasath, et al.
Publicado: (2024) -
Political Alignment in Large Language Models: A Multidimensional Audit of Psychometric Identity and Behavioral Bias
por: Sakhawat, Adib, et al.
Publicado: (2026) -
Coordinates of Capability: A Unified MTMM-Geometric Framework for LLM Evaluation
por: Sakhawat, Adib, et al.
Publicado: (2026) -
Gradient Masters at BLP-2025 Task 1: Advancing Low-Resource NLP for Bengali using Ensemble-Based Adversarial Training for Hate Speech Detection
por: Hoque, Syed Mohaiminul, et al.
Publicado: (2025)