Rerouting LLM Routers
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Shafran, Avital, Schuster, Roei, Ristenpart, Thomas, Shmatikov, Vitaly |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Machine Against the RAG: Jamming Retrieval-Augmented Generation with Blocker Documents
von: Shafran, Avital, et al.
Veröffentlicht: (2024)
von: Shafran, Avital, et al.
Veröffentlicht: (2024)
Multi-Agent Systems Execute Arbitrary Malicious Code
von: Triedman, Harold, et al.
Veröffentlicht: (2025)
von: Triedman, Harold, et al.
Veröffentlicht: (2025)
Laundering AI Authority with Adversarial Examples
von: Zhang, Jie, et al.
Veröffentlicht: (2026)
von: Zhang, Jie, et al.
Veröffentlicht: (2026)
Beyond Labeling Oracles: What does it mean to steal ML models?
von: Shafran, Avital, et al.
Veröffentlicht: (2023)
von: Shafran, Avital, et al.
Veröffentlicht: (2023)
Adversarial Decoding: Generating Readable Documents for Adversarial Objectives
von: Zhang, Collin, et al.
Veröffentlicht: (2024)
von: Zhang, Collin, et al.
Veröffentlicht: (2024)
Breaking and Fixing Defenses Against Control-Flow Hijacking in Multi-Agent Systems
von: Jha, Rishi, et al.
Veröffentlicht: (2025)
von: Jha, Rishi, et al.
Veröffentlicht: (2025)
Adversarial Illusions in Multi-Modal Embeddings
von: Zhang, Tingwei, et al.
Veröffentlicht: (2023)
von: Zhang, Tingwei, et al.
Veröffentlicht: (2023)
Learned-Database Systems Security
von: Schuster, Roei, et al.
Veröffentlicht: (2022)
von: Schuster, Roei, et al.
Veröffentlicht: (2022)
Self-interpreting Adversarial Images
von: Zhang, Tingwei, et al.
Veröffentlicht: (2024)
von: Zhang, Tingwei, et al.
Veröffentlicht: (2024)
Differential Degradation Vulnerabilities in Censorship Circumvention Systems
von: Sun, Zhen, et al.
Veröffentlicht: (2024)
von: Sun, Zhen, et al.
Veröffentlicht: (2024)
Deep-Research Agents Can Be Poisoned via User-Generated Content
von: Zhang, Tingwei, et al.
Veröffentlicht: (2026)
von: Zhang, Tingwei, et al.
Veröffentlicht: (2026)
How to Steal Reasoning Without Reasoning Traces
von: Zhang, Tingwei, et al.
Veröffentlicht: (2026)
von: Zhang, Tingwei, et al.
Veröffentlicht: (2026)
Transcript Franking for Encrypted Messaging
von: Namavari, Armin, et al.
Veröffentlicht: (2025)
von: Namavari, Armin, et al.
Veröffentlicht: (2025)
A New Dataset and Methodology for Malicious URL Classification
von: Schvartzman, Ilan, et al.
Veröffentlicht: (2024)
von: Schvartzman, Ilan, et al.
Veröffentlicht: (2024)
Universal Zero-shot Embedding Inversion
von: Zhang, Collin, et al.
Veröffentlicht: (2025)
von: Zhang, Collin, et al.
Veröffentlicht: (2025)
RerouteGuard: Understanding and Mitigating Adversarial Risks for LLM Routing
von: Zhang, Wenhui, et al.
Veröffentlicht: (2026)
von: Zhang, Wenhui, et al.
Veröffentlicht: (2026)
Agent Meltdowns: The Road to Hell Is Paved with Helpful Agents
von: Jha, Rishi, et al.
Veröffentlicht: (2026)
von: Jha, Rishi, et al.
Veröffentlicht: (2026)
Route to Rome Attack: Directing LLM Routers to Expensive Models via Adversarial Suffix Optimization
von: Tang, Haochun, et al.
Veröffentlicht: (2026)
von: Tang, Haochun, et al.
Veröffentlicht: (2026)
Cascade: Token-Sharded Private LLM Inference
von: Thomas, Rahul, et al.
Veröffentlicht: (2025)
von: Thomas, Rahul, et al.
Veröffentlicht: (2025)
Noticing the Watcher: LLM Agents Can Infer CoT Monitoring from Blocking Feedback
von: Jiralerspong, Thomas, et al.
Veröffentlicht: (2026)
von: Jiralerspong, Thomas, et al.
Veröffentlicht: (2026)
Enabling Differentially Private Federated Learning for Speech Recognition: Benchmarks, Adaptive Optimizers and Gradient Clipping
von: Pelikan, Martin, et al.
Veröffentlicht: (2023)
von: Pelikan, Martin, et al.
Veröffentlicht: (2023)
Adversarial Hubness in Multi-Modal Retrieval
von: Zhang, Tingwei, et al.
Veröffentlicht: (2024)
von: Zhang, Tingwei, et al.
Veröffentlicht: (2024)
AttackLLM: LLM-based Attack Pattern Generation for an Industrial Control System
von: Ahmed, Chuadhry Mujeeb
Veröffentlicht: (2025)
von: Ahmed, Chuadhry Mujeeb
Veröffentlicht: (2025)
Encryption-Friendly LLM Architecture
von: Rho, Donghwan, et al.
Veröffentlicht: (2024)
von: Rho, Donghwan, et al.
Veröffentlicht: (2024)
Log Probability Tracking of LLM APIs
von: Chauvin, Timothée, et al.
Veröffentlicht: (2025)
von: Chauvin, Timothée, et al.
Veröffentlicht: (2025)
Is The Watermarking Of LLM-Generated Code Robust?
von: Suresh, Tarun, et al.
Veröffentlicht: (2024)
von: Suresh, Tarun, et al.
Veröffentlicht: (2024)
Good-Enough LLM Obfuscation (GELO)
von: Belikov, Anatoly, et al.
Veröffentlicht: (2026)
von: Belikov, Anatoly, et al.
Veröffentlicht: (2026)
PAE MobiLLM: Privacy-Aware and Efficient LLM Fine-Tuning on the Mobile Device via Additive Side-Tuning
von: Yang, Xingke, et al.
Veröffentlicht: (2025)
von: Yang, Xingke, et al.
Veröffentlicht: (2025)
Innovative tokenisation of structured data for LLM training
von: Karim, Kayvan, et al.
Veröffentlicht: (2025)
von: Karim, Kayvan, et al.
Veröffentlicht: (2025)
LLM-Generated Samples for Android Malware Detection
von: Rollinson, Nik, et al.
Veröffentlicht: (2025)
von: Rollinson, Nik, et al.
Veröffentlicht: (2025)
LLM Fingerprinting via Semantically Conditioned Watermarks
von: Gloaguen, Thibaud, et al.
Veröffentlicht: (2025)
von: Gloaguen, Thibaud, et al.
Veröffentlicht: (2025)
Exploiting Leakage in Password Managers via Injection Attacks
von: Fábrega, Andrés, et al.
Veröffentlicht: (2024)
von: Fábrega, Andrés, et al.
Veröffentlicht: (2024)
Adversarial Contrastive Learning for LLM Quantization Attacks
von: Song, Dinghong, et al.
Veröffentlicht: (2026)
von: Song, Dinghong, et al.
Veröffentlicht: (2026)
Token-Efficient Change Detection in LLM APIs
von: Chauvin, Timothée, et al.
Veröffentlicht: (2026)
von: Chauvin, Timothée, et al.
Veröffentlicht: (2026)
Memory-Induced Tool-Drift in LLM Agents
von: Dabas, Mahavir, et al.
Veröffentlicht: (2026)
von: Dabas, Mahavir, et al.
Veröffentlicht: (2026)
Order of Magnitude Speedups for LLM Membership Inference
von: Zhang, Rongting, et al.
Veröffentlicht: (2024)
von: Zhang, Rongting, et al.
Veröffentlicht: (2024)
LLM Watermarking Using Mixtures and Statistical-to-Computational Gaps
von: Abdalla, Pedro, et al.
Veröffentlicht: (2025)
von: Abdalla, Pedro, et al.
Veröffentlicht: (2025)
Fundamental Limitations in Pointwise Defences of LLM Finetuning APIs
von: Davies, Xander, et al.
Veröffentlicht: (2025)
von: Davies, Xander, et al.
Veröffentlicht: (2025)
Verifying LLM Inference to Detect Model Weight Exfiltration
von: Rinberg, Roy, et al.
Veröffentlicht: (2025)
von: Rinberg, Roy, et al.
Veröffentlicht: (2025)
Black-box Optimization of LLM Outputs by Asking for Directions
von: Zhang, Jie, et al.
Veröffentlicht: (2025)
von: Zhang, Jie, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Machine Against the RAG: Jamming Retrieval-Augmented Generation with Blocker Documents
von: Shafran, Avital, et al.
Veröffentlicht: (2024) -
Multi-Agent Systems Execute Arbitrary Malicious Code
von: Triedman, Harold, et al.
Veröffentlicht: (2025) -
Laundering AI Authority with Adversarial Examples
von: Zhang, Jie, et al.
Veröffentlicht: (2026) -
Beyond Labeling Oracles: What does it mean to steal ML models?
von: Shafran, Avital, et al.
Veröffentlicht: (2023) -
Adversarial Decoding: Generating Readable Documents for Adversarial Objectives
von: Zhang, Collin, et al.
Veröffentlicht: (2024)