Ignore All Previous Instructions: Jailbreaking as a de-escalatory peace building practise to resist LLM social media bots
Fuente:
arXiv
Saved in:
| Main Authors: | Day, Huw, Jezierska, Adrianna, Woodgate, Jessica |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Highlight All the Phrases: Enhancing LLM Transparency through Visual Factuality Indicators
by: Do, Hyo Jin, et al.
Published: (2025)
by: Do, Hyo Jin, et al.
Published: (2025)
Play Guessing Game with LLM: Indirect Jailbreak Attack with Implicit Clues
by: Chang, Zhiyuan, et al.
Published: (2024)
by: Chang, Zhiyuan, et al.
Published: (2024)
Transforming GenAI Policy to Prompting Instruction: An RCT of Scalable Prompting Interventions in a CS1 Course
by: Xiao, Ruiwei, et al.
Published: (2026)
by: Xiao, Ruiwei, et al.
Published: (2026)
ComplLLM: Fine-tuning LLMs to Discover Complementary Signals for Decision-making
by: Guo, Ziyang, et al.
Published: (2026)
by: Guo, Ziyang, et al.
Published: (2026)
Beyond prior knowledge: The predictive role of knowledge-building in Tutor Learning
by: Shahriar, Tasmia, et al.
Published: (2025)
by: Shahriar, Tasmia, et al.
Published: (2025)
How to build trust in answers given by Generative AI for specific, and vague, financial questions
by: Zarifis, Alex, et al.
Published: (2024)
by: Zarifis, Alex, et al.
Published: (2024)
Aria-UI: Visual Grounding for GUI Instructions
by: Yang, Yuhao, et al.
Published: (2024)
by: Yang, Yuhao, et al.
Published: (2024)
Lost in Instructions: Study of Blind Users' Experiences with DIY Manuals and AI-Rewritten Instructions for Assembly, Operation, and Troubleshooting of Tangible Products
by: Reddy, Monalika Padma, et al.
Published: (2026)
by: Reddy, Monalika Padma, et al.
Published: (2026)
Imagining and building wise machines: The centrality of AI metacognition
by: Johnson, Samuel G. B., et al.
Published: (2024)
by: Johnson, Samuel G. B., et al.
Published: (2024)
The influence of persona and conversational task on social interactions with a LLM-controlled embodied conversational agent
by: Kroczek, Leon O. H., et al.
Published: (2024)
by: Kroczek, Leon O. H., et al.
Published: (2024)
InstructPipe: Generating Visual Blocks Pipelines with Human Instructions and LLMs
by: Zhou, Zhongyi, et al.
Published: (2023)
by: Zhou, Zhongyi, et al.
Published: (2023)
Situated Instruction Following
by: Min, So Yeon, et al.
Published: (2024)
by: Min, So Yeon, et al.
Published: (2024)
XAI for All: Can Large Language Models Simplify Explainable AI?
by: Mavrepis, Philip, et al.
Published: (2024)
by: Mavrepis, Philip, et al.
Published: (2024)
Point and Instruct: Enabling Precise Image Editing by Unifying Direct Manipulation and Text Instructions
by: Helbling, Alec, et al.
Published: (2024)
by: Helbling, Alec, et al.
Published: (2024)
NoteIt: A System Converting Instructional Videos to Interactable Notes Through Multimodal Video Understanding
by: Zhao, Running, et al.
Published: (2025)
by: Zhao, Running, et al.
Published: (2025)
HelpViz: Automatic Generation of Contextual Visual MobileTutorials from Text-Based Instructions
by: Zhong, Mingyuan, et al.
Published: (2021)
by: Zhong, Mingyuan, et al.
Published: (2021)
Designing Gaze Analytics for ELA Instruction: A User-Centered Dashboard with Conversational AI Support
by: Davalos, Eduardo, et al.
Published: (2025)
by: Davalos, Eduardo, et al.
Published: (2025)
Not All Trust is the Same: Effects of Decision Workflow and Explanations in Human-AI Decision Making
by: Spillner, Laura, et al.
Published: (2026)
by: Spillner, Laura, et al.
Published: (2026)
Who Validates the Validators? Aligning LLM-Assisted Evaluation of LLM Outputs with Human Preferences
by: Shankar, Shreya, et al.
Published: (2024)
by: Shankar, Shreya, et al.
Published: (2024)
HCC Is All You Need: Alignment-The Sensible Kind Anyway-Is Just Human-Centered Computing
by: Gilbert, Eric
Published: (2024)
by: Gilbert, Eric
Published: (2024)
LLM Enhancement with Domain Expert Mental Model to Reduce LLM Hallucination with Causal Prompt Engineering
by: Kovalerchuk, Boris, et al.
Published: (2025)
by: Kovalerchuk, Boris, et al.
Published: (2025)
LLM-DaaS: LLM-driven Drone-as-a-Service Operations from Text User Requests
by: Wassim, Lillian, et al.
Published: (2024)
by: Wassim, Lillian, et al.
Published: (2024)
Underspecified Human Decision Experiments Considered Harmful
by: Hullman, Jessica, et al.
Published: (2024)
by: Hullman, Jessica, et al.
Published: (2024)
Is Conversational XAI All You Need? Human-AI Decision Making With a Conversational XAI Assistant
by: He, Gaole, et al.
Published: (2025)
by: He, Gaole, et al.
Published: (2025)
Knowing Your Uncertainty -- On the application of LLM in social sciences
by: Zhang, Bolun, et al.
Published: (2025)
by: Zhang, Bolun, et al.
Published: (2025)
Visual Fingerprints for LLM Generation Comparison
by: Alnouri, Amal, et al.
Published: (2026)
by: Alnouri, Amal, et al.
Published: (2026)
Determinants of LLM-assisted Decision-Making
by: Eigner, Eva, et al.
Published: (2024)
by: Eigner, Eva, et al.
Published: (2024)
"HOT" ChatGPT: The promise of ChatGPT in detecting and discriminating hateful, offensive, and toxic comments on social media
by: Li, Lingyao, et al.
Published: (2023)
by: Li, Lingyao, et al.
Published: (2023)
LLM Bazaar: A Service Design for Supporting Collaborative Learning with an LLM-Powered Multi-Party Collaboration Infrastructure
by: Wu, Zhen, et al.
Published: (2025)
by: Wu, Zhen, et al.
Published: (2025)
LLM-Assisted Visual Analytics: Opportunities and Challenges
by: Hutchinson, Maeve, et al.
Published: (2024)
by: Hutchinson, Maeve, et al.
Published: (2024)
Challenges & Opportunities with LLM-Assisted Visualization Retargeting
by: Snyder, Luke S., et al.
Published: (2025)
by: Snyder, Luke S., et al.
Published: (2025)
Secret Use of Large Language Model (LLM)
by: Zhang, Zhiping, et al.
Published: (2024)
by: Zhang, Zhiping, et al.
Published: (2024)
PolicyPad: Collaborative Prototyping of LLM Policies
by: Feng, K. J. Kevin, et al.
Published: (2025)
by: Feng, K. J. Kevin, et al.
Published: (2025)
Wearable Meets LLM for Stress Management: A Duoethnographic Study Integrating Wearable-Triggered Stressors and LLM Chatbots for Personalized Interventions
by: Neupane, Sameer, et al.
Published: (2025)
by: Neupane, Sameer, et al.
Published: (2025)
Hacc-Man: An Arcade Game for Jailbreaking LLMs
by: Valentim, Matheus, et al.
Published: (2024)
by: Valentim, Matheus, et al.
Published: (2024)
Unexploited Information Value in Human-AI Collaboration
by: Guo, Ziyang, et al.
Published: (2024)
by: Guo, Ziyang, et al.
Published: (2024)
Incremental XAI: Memorable Understanding of AI with Incremental Explanations
by: Bo, Jessica Y., et al.
Published: (2024)
by: Bo, Jessica Y., et al.
Published: (2024)
A Decision Theoretic Framework for Measuring AI Reliance
by: Guo, Ziyang, et al.
Published: (2024)
by: Guo, Ziyang, et al.
Published: (2024)
An Empirical Exploration of Trust Dynamics in LLM Supply Chains
by: Balayn, Agathe, et al.
Published: (2024)
by: Balayn, Agathe, et al.
Published: (2024)
Human Decision-Making with Persuasive and Narrative LLM Explanations
by: Marusich, Laura R., et al.
Published: (2026)
by: Marusich, Laura R., et al.
Published: (2026)
Similar Items
-
Highlight All the Phrases: Enhancing LLM Transparency through Visual Factuality Indicators
by: Do, Hyo Jin, et al.
Published: (2025) -
Play Guessing Game with LLM: Indirect Jailbreak Attack with Implicit Clues
by: Chang, Zhiyuan, et al.
Published: (2024) -
Transforming GenAI Policy to Prompting Instruction: An RCT of Scalable Prompting Interventions in a CS1 Course
by: Xiao, Ruiwei, et al.
Published: (2026) -
ComplLLM: Fine-tuning LLMs to Discover Complementary Signals for Decision-making
by: Guo, Ziyang, et al.
Published: (2026) -
Beyond prior knowledge: The predictive role of knowledge-building in Tutor Learning
by: Shahriar, Tasmia, et al.
Published: (2025)