TrojanStego: Your Language Model Can Secretly Be A Steganographic Privacy Leaking Agent
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Meier, Dominik, Wahle, Jan Philip, Röttger, Paul, Ruas, Terry, Gipp, Bela |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Towards Human Understanding of Paraphrase Types in Large Language Models
von: Meier, Dominik, et al.
Veröffentlicht: (2024)
von: Meier, Dominik, et al.
Veröffentlicht: (2024)
StegoStylo: Squelching Stylometric Scrutiny through Steganographic Stitching
von: Dilworth, Robert
Veröffentlicht: (2026)
von: Dilworth, Robert
Veröffentlicht: (2026)
Paraphrase Types for Generation and Detection
von: Wahle, Jan Philip, et al.
Veröffentlicht: (2023)
von: Wahle, Jan Philip, et al.
Veröffentlicht: (2023)
How Large Language Models are Transforming Machine-Paraphrased Plagiarism
von: Wahle, Jan Philip, et al.
Veröffentlicht: (2022)
von: Wahle, Jan Philip, et al.
Veröffentlicht: (2022)
Who Watches the Watchmen? Humans Disagree With Translation Metrics on Unseen Domains
von: Schmidt, Finn, et al.
Veröffentlicht: (2026)
von: Schmidt, Finn, et al.
Veröffentlicht: (2026)
Paraphrase Types Elicit Prompt Engineering Capabilities
von: Wahle, Jan Philip, et al.
Veröffentlicht: (2024)
von: Wahle, Jan Philip, et al.
Veröffentlicht: (2024)
Piecing Together Cross-Document Coreference Resolution Datasets: Systematic Dataset Analysis and Unification
von: Zhukova, Anastasia, et al.
Veröffentlicht: (2026)
von: Zhukova, Anastasia, et al.
Veröffentlicht: (2026)
What's under the hood: Investigating Automatic Metrics on Meeting Summarization
von: Kirstein, Frederic, et al.
Veröffentlicht: (2024)
von: Kirstein, Frederic, et al.
Veröffentlicht: (2024)
CADS: A Systematic Literature Review on the Challenges of Abstractive Dialogue Summarization
von: Kirstein, Frederic, et al.
Veröffentlicht: (2024)
von: Kirstein, Frederic, et al.
Veröffentlicht: (2024)
Text Generation: A Systematic Literature Review of Tasks, Evaluation, and Challenges
von: Becker, Jonas, et al.
Veröffentlicht: (2024)
von: Becker, Jonas, et al.
Veröffentlicht: (2024)
RedacBench: Can AI Erase Your Secrets?
von: Jeon, Hyunjun, et al.
Veröffentlicht: (2026)
von: Jeon, Hyunjun, et al.
Veröffentlicht: (2026)
You need to MIMIC to get FAME: Solving Meeting Transcript Scarcity with a Multi-Agent Conversations
von: Kirstein, Frederic, et al.
Veröffentlicht: (2025)
von: Kirstein, Frederic, et al.
Veröffentlicht: (2025)
TrojanRAG: Retrieval-Augmented Generation Can Be Backdoor Driver in Large Language Models
von: Cheng, Pengzhou, et al.
Veröffentlicht: (2024)
von: Cheng, Pengzhou, et al.
Veröffentlicht: (2024)
CiteAssist: A System for Automated Preprint Citation and BibTeX Generation
von: Kaesberg, Lars Benedikt, et al.
Veröffentlicht: (2024)
von: Kaesberg, Lars Benedikt, et al.
Veröffentlicht: (2024)
D3: A Massive Dataset of Scholarly Metadata for Analyzing the State of Computer Science Research
von: Wahle, Jan Philip, et al.
Veröffentlicht: (2022)
von: Wahle, Jan Philip, et al.
Veröffentlicht: (2022)
SPaRC: A Spatial Pathfinding Reasoning Challenge
von: Kaesberg, Lars Benedikt, et al.
Veröffentlicht: (2025)
von: Kaesberg, Lars Benedikt, et al.
Veröffentlicht: (2025)
Big Tech-Funded AI Papers Have Higher Citation Impact, Greater Insularity, and Larger Recency Bias
von: Gnewuch, Max Martin, et al.
Veröffentlicht: (2025)
von: Gnewuch, Max Martin, et al.
Veröffentlicht: (2025)
Voting or Consensus? Decision-Making in Multi-Agent Debate
von: Kaesberg, Lars Benedikt, et al.
Veröffentlicht: (2025)
von: Kaesberg, Lars Benedikt, et al.
Veröffentlicht: (2025)
Can LLMs Keep a Secret? Testing Privacy Implications of Language Models via Contextual Integrity Theory
von: Mireshghallah, Niloofar, et al.
Veröffentlicht: (2023)
von: Mireshghallah, Niloofar, et al.
Veröffentlicht: (2023)
We are Who We Cite: Bridges of Influence Between Natural Language Processing and Other Academic Fields
von: Wahle, Jan Philip, et al.
Veröffentlicht: (2023)
von: Wahle, Jan Philip, et al.
Veröffentlicht: (2023)
You Can't Trust Your Tag Neither: Privacy Leaks and Potential Legal Violations within the Google Tag Manager
von: Mertens, Gilles, et al.
Veröffentlicht: (2023)
von: Mertens, Gilles, et al.
Veröffentlicht: (2023)
The Sum Leaks More Than Its Parts: Compositional Privacy Risks and Mitigations in Multi-Agent Collaboration
von: Patil, Vaidehi, et al.
Veröffentlicht: (2025)
von: Patil, Vaidehi, et al.
Veröffentlicht: (2025)
Your LLM Agent Can Leak Your Data: Data Exfiltration via Backdoored Tool Use
von: Zhang, Wuyang, et al.
Veröffentlicht: (2026)
von: Zhang, Wuyang, et al.
Veröffentlicht: (2026)
StegoHound: A Novel Multi-Approaches Method for Efficient and Effective Identification and Extraction of Digital Evidence Masked by Steganographic Techniques in WAV and MP3 Files
von: Ghanem, Mohamed C., et al.
Veröffentlicht: (2023)
von: Ghanem, Mohamed C., et al.
Veröffentlicht: (2023)
MALLM: Multi-Agent Large Language Models Framework
von: Becker, Jonas, et al.
Veröffentlicht: (2025)
von: Becker, Jonas, et al.
Veröffentlicht: (2025)
Simple Prompt Injection Attacks Can Leak Personal Data Observed by LLM Agents During Task Execution
von: Alizadeh, Meysam, et al.
Veröffentlicht: (2025)
von: Alizadeh, Meysam, et al.
Veröffentlicht: (2025)
Leaking LoRa: An Evaluation of Password Leaks and Knowledge Storage in Large Language Models
von: Marinelli, Ryan, et al.
Veröffentlicht: (2025)
von: Marinelli, Ryan, et al.
Veröffentlicht: (2025)
A Character-based Diffusion Embedding Algorithm for Enhancing the Generation Quality of Generative Linguistic Steganographic Texts
von: Chen, Yingquan, et al.
Veröffentlicht: (2025)
von: Chen, Yingquan, et al.
Veröffentlicht: (2025)
Sharing The Secret: Distributed Privacy-Preserving Monitoring
von: Karimi, Mahyar, et al.
Veröffentlicht: (2026)
von: Karimi, Mahyar, et al.
Veröffentlicht: (2026)
Citation Amnesia: On The Recency Bias of NLP and Other Academic Fields
von: Wahle, Jan Philip, et al.
Veröffentlicht: (2024)
von: Wahle, Jan Philip, et al.
Veröffentlicht: (2024)
E-Trojans: Ransomware, Tracking, DoS, and Data Leaks on Battery-powered Embedded Systems
von: Casagrande, Marco, et al.
Veröffentlicht: (2024)
von: Casagrande, Marco, et al.
Veröffentlicht: (2024)
Testing the Generalization of Neural Language Models for COVID-19 Misinformation Detection
von: Wahle, Jan Philip, et al.
Veröffentlicht: (2021)
von: Wahle, Jan Philip, et al.
Veröffentlicht: (2021)
I Can Tell Your Secrets: Inferring Privacy Attributes from Mini-app Interaction History in Super-apps
von: Cai, Yifeng, et al.
Veröffentlicht: (2025)
von: Cai, Yifeng, et al.
Veröffentlicht: (2025)
Hidden in Plain Text: Emergence & Mitigation of Steganographic Collusion in LLMs
von: Mathew, Yohan, et al.
Veröffentlicht: (2024)
von: Mathew, Yohan, et al.
Veröffentlicht: (2024)
Mind the Gap Between Spatial Reasoning and Acting! Step-by-Step Evaluation of Agents With Spatial-Gym
von: Kaesberg, Lars Benedikt, et al.
Veröffentlicht: (2026)
von: Kaesberg, Lars Benedikt, et al.
Veröffentlicht: (2026)
Why Are My Prompts Leaked? Unraveling Prompt Extraction Threats in Customized Large Language Models
von: Liang, Zi, et al.
Veröffentlicht: (2024)
von: Liang, Zi, et al.
Veröffentlicht: (2024)
Agent Tools Orchestration Leaks More: Dataset, Benchmark, and Mitigation
von: Qiao, Yuxuan, et al.
Veröffentlicht: (2025)
von: Qiao, Yuxuan, et al.
Veröffentlicht: (2025)
Stay Focused: Problem Drift in Multi-Agent Debate
von: Becker, Jonas, et al.
Veröffentlicht: (2025)
von: Becker, Jonas, et al.
Veröffentlicht: (2025)
Sanitize Your Responses: Mitigating Privacy Leakage in Large Language Models
von: Fu, Wenjie, et al.
Veröffentlicht: (2025)
von: Fu, Wenjie, et al.
Veröffentlicht: (2025)
Do Phone-Use Agents Respect Your Privacy?
von: Tang, Zhengyang, et al.
Veröffentlicht: (2026)
von: Tang, Zhengyang, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
Towards Human Understanding of Paraphrase Types in Large Language Models
von: Meier, Dominik, et al.
Veröffentlicht: (2024) -
StegoStylo: Squelching Stylometric Scrutiny through Steganographic Stitching
von: Dilworth, Robert
Veröffentlicht: (2026) -
Paraphrase Types for Generation and Detection
von: Wahle, Jan Philip, et al.
Veröffentlicht: (2023) -
How Large Language Models are Transforming Machine-Paraphrased Plagiarism
von: Wahle, Jan Philip, et al.
Veröffentlicht: (2022) -
Who Watches the Watchmen? Humans Disagree With Translation Metrics on Unseen Domains
von: Schmidt, Finn, et al.
Veröffentlicht: (2026)