Guardado en:
| Autores principales: | Wiegmann, Matti, Rakete, Jennifer, Wolska, Magdalena, Stein, Benno, Potthast, Martin |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | https://arxiv.org/abs/2404.09615 |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Understanding the Perceptions of Trigger Warning and Content Warning on Social Media Platforms in the U.S
por: Zhang, Xinyi, et al.
Publicado: (2025)
por: Zhang, Xinyi, et al.
Publicado: (2025)
Trigger Warnings
Publicado: (2026)
Publicado: (2026)
An Investigation of Warning Erroneous Chat Translations in Cross-lingual Communication
por: Li, Yunmeng, et al.
Publicado: (2024)
por: Li, Yunmeng, et al.
Publicado: (2024)
Analyzing the Safety of Japanese Large Language Models in Stereotype-Triggering Prompts
por: Nakanishi, Akito, et al.
Publicado: (2025)
por: Nakanishi, Akito, et al.
Publicado: (2025)
VIGIL: An Extensible System for Real-Time Detection and Mitigation of Cognitive Bias Triggers
por: Kang, Bo, et al.
Publicado: (2026)
por: Kang, Bo, et al.
Publicado: (2026)
BadFair: Backdoored Fairness Attacks with Group-conditioned Triggers
por: Xue, Jiaqi, et al.
Publicado: (2024)
por: Xue, Jiaqi, et al.
Publicado: (2024)
Task-Oriented Paraphrase Analytics
por: Gohsen, Marcel, et al.
Publicado: (2024)
por: Gohsen, Marcel, et al.
Publicado: (2024)
The Ontological Dissonance Hypothesis: AI-Triggered Delusional Ideation as Folie a Deux Technologique
por: Lipinska, Izabela, et al.
Publicado: (2025)
por: Lipinska, Izabela, et al.
Publicado: (2025)
Are Triggers Needed for Document-Level Event Extraction?
por: Shaar, Shaden, et al.
Publicado: (2024)
por: Shaar, Shaden, et al.
Publicado: (2024)
Causal Effects of Trigger Words in Social Media Discussions: A Large-Scale Case Study about UK Politics on Reddit
por: Antypas, Dimosthenis, et al.
Publicado: (2024)
por: Antypas, Dimosthenis, et al.
Publicado: (2024)
Investigating Adversarial Trigger Transfer in Large Language Models
por: Meade, Nicholas, et al.
Publicado: (2024)
por: Meade, Nicholas, et al.
Publicado: (2024)
Does the Source of a Warning Matter? Examining the Effectiveness of Veracity Warning Labels Across Warners
por: Horne, Benjamin D.
Publicado: (2024)
por: Horne, Benjamin D.
Publicado: (2024)
Universal Adversarial Triggers
por: Arockiaraj, Benedict Florance, et al.
Publicado: (2026)
por: Arockiaraj, Benedict Florance, et al.
Publicado: (2026)
Detecting Generated Native Ads in Conversational Search
por: Schmidt, Sebastian, et al.
Publicado: (2024)
por: Schmidt, Sebastian, et al.
Publicado: (2024)
Linguistic Hooks: Investigating The Role of Language Triggers in Phishing Emails Targeting African Refugees and Students
por: Menon, Mythili, et al.
Publicado: (2025)
por: Menon, Mythili, et al.
Publicado: (2025)
Argumentation in Waltz's "Emerging Structure of International Politics''
por: Wolska, Magdalena, et al.
Publicado: (2023)
por: Wolska, Magdalena, et al.
Publicado: (2023)
An Early Warning Model for Forced Displacement
por: Henningsen, Geraldine
Publicado: (2025)
por: Henningsen, Geraldine
Publicado: (2025)
Future Events as Backdoor Triggers: Investigating Temporal Vulnerabilities in LLMs
por: Price, Sara, et al.
Publicado: (2024)
por: Price, Sara, et al.
Publicado: (2024)
Can I guess where you are from? Modeling dialectal morphosyntactic similarities in Brazilian Portuguese
por: Siqueira, Manoel, et al.
Publicado: (2026)
por: Siqueira, Manoel, et al.
Publicado: (2026)
Storm Warning.
por: Filmer, Paul E.
Publicado: (1995)
por: Filmer, Paul E.
Publicado: (1995)
NatLan: Native Language Prompting Facilitates Knowledge Elicitation Through Language Trigger Provision and Domain Trigger Retention
por: Li, Baixuan, et al.
Publicado: (2024)
por: Li, Baixuan, et al.
Publicado: (2024)
Contextualising (Im)plausible Events Triggers Figurative Language
por: Eichel, Annerose, et al.
Publicado: (2026)
por: Eichel, Annerose, et al.
Publicado: (2026)
DeepInnovator: Triggering the Innovative Capabilities of LLMs
por: Fan, Tianyu, et al.
Publicado: (2026)
por: Fan, Tianyu, et al.
Publicado: (2026)
Investigating Symbolic Triggers of Hallucination in Gemma Models Across HaluEval and TruthfulQA
por: Lamba, Naveen, et al.
Publicado: (2025)
por: Lamba, Naveen, et al.
Publicado: (2025)
Toward Finding and Supporting Struggling Students in a Programming Course with an Early Warning System
por: Schantong, Belinda, et al.
Publicado: (2024)
por: Schantong, Belinda, et al.
Publicado: (2024)
Data Voids and Warning Banners on Google Search
por: Robertson, Ronald E., et al.
Publicado: (2025)
por: Robertson, Ronald E., et al.
Publicado: (2025)
Backdoor Attacks on Dense Retrieval via Public and Unintentional Triggers
por: Long, Quanyu, et al.
Publicado: (2024)
por: Long, Quanyu, et al.
Publicado: (2024)
Adaptive Deep Reasoning: Triggering Deep Thinking When Needed
por: Wang, Yunhao, et al.
Publicado: (2025)
por: Wang, Yunhao, et al.
Publicado: (2025)
Are Security Cues Static? Rethinking Warning and Trust Indicators for Life Transitions
por: Tabassum, Sarah
Publicado: (2026)
por: Tabassum, Sarah
Publicado: (2026)
Learning from Negative Examples: Why Warning-Framed Training Data Teaches What It Warns Against
por: Enkhbayar, Tsogt-Ochir
Publicado: (2025)
por: Enkhbayar, Tsogt-Ochir
Publicado: (2025)
Mid-Think: Training-Free Intermediate-Budget Reasoning via Token-Level Triggers
por: Yang, Wang, et al.
Publicado: (2026)
por: Yang, Wang, et al.
Publicado: (2026)
Language-Switching Triggers Take a Latent Detour Through Language Models
por: Kulumba, Francis, et al.
Publicado: (2026)
por: Kulumba, Francis, et al.
Publicado: (2026)
Trigger$^3$: Refining Query Correction via Adaptive Model Selector
por: Zhang, Kepu, et al.
Publicado: (2024)
por: Zhang, Kepu, et al.
Publicado: (2024)
TWeddit : A Dataset of Triggering Stories Predominantly Shared by Women on Reddit
por: Bandela, Shirlene Rose, et al.
Publicado: (2026)
por: Bandela, Shirlene Rose, et al.
Publicado: (2026)
Small Models are LLM Knowledge Triggers on Medical Tabular Prediction
por: Yan, Jiahuan, et al.
Publicado: (2024)
por: Yan, Jiahuan, et al.
Publicado: (2024)
Prompt as Triggers for Backdoor Attack: Examining the Vulnerability in Language Models
por: Zhao, Shuai, et al.
Publicado: (2023)
por: Zhao, Shuai, et al.
Publicado: (2023)
ExpSeek: Self-Triggered Experience Seeking for Web Agents
por: Zhang, Wenyuan, et al.
Publicado: (2026)
por: Zhang, Wenyuan, et al.
Publicado: (2026)
Coordinated Disclosure of Dual-Use Capabilities: An Early Warning System for Advanced AI
por: O'Brien, Joe, et al.
Publicado: (2024)
por: O'Brien, Joe, et al.
Publicado: (2024)
Earthquake Early Warning and Beyond: Systems Challenges in Smartphone-based Seismic Network
por: Kong, Qingkai, et al.
Publicado: (2019)
por: Kong, Qingkai, et al.
Publicado: (2019)
Multi-Trigger Poisoning Amplifies Backdoor Vulnerabilities in LLMs
por: Sivapiromrat, Sanhanat, et al.
Publicado: (2025)
por: Sivapiromrat, Sanhanat, et al.
Publicado: (2025)
Ejemplares similares
-
Understanding the Perceptions of Trigger Warning and Content Warning on Social Media Platforms in the U.S
por: Zhang, Xinyi, et al.
Publicado: (2025) -
Trigger Warnings
Publicado: (2026) -
An Investigation of Warning Erroneous Chat Translations in Cross-lingual Communication
por: Li, Yunmeng, et al.
Publicado: (2024) -
Analyzing the Safety of Japanese Large Language Models in Stereotype-Triggering Prompts
por: Nakanishi, Akito, et al.
Publicado: (2025) -
VIGIL: An Extensible System for Real-Time Detection and Mitigation of Cognitive Bias Triggers
por: Kang, Bo, et al.
Publicado: (2026)