OMNI-LEAK: Orchestrator Multi-Agent Network Induced Data Leakage
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Naik, Akshat, Culligan, Jay, Gal, Yarin, Torr, Philip, Aljundi, Rahaf, Paren, Alasdair, Bibi, Adel |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
ToolTweak: An Attack on Tool Selection in LLM-based Agents
von: Sneh, Jonathan, et al.
Veröffentlicht: (2025)
von: Sneh, Jonathan, et al.
Veröffentlicht: (2025)
BiasBusters: Uncovering and Mitigating Tool Selection Bias in Large Language Models
von: Blankenstein, Thierry, et al.
Veröffentlicht: (2025)
von: Blankenstein, Thierry, et al.
Veröffentlicht: (2025)
MIP against Agent: Malicious Image Patches Hijacking Multimodal OS Agents
von: Aichberger, Lukas, et al.
Veröffentlicht: (2025)
von: Aichberger, Lukas, et al.
Veröffentlicht: (2025)
Universal In-Context Approximation By Prompting Fully Recurrent Models
von: Petrov, Aleksandar, et al.
Veröffentlicht: (2024)
von: Petrov, Aleksandar, et al.
Veröffentlicht: (2024)
Memo: Training Memory-Efficient Embodied Agents with Reinforcement Learning
von: Gupta, Gunshi, et al.
Veröffentlicht: (2025)
von: Gupta, Gunshi, et al.
Veröffentlicht: (2025)
Shh, don't say that! Domain Certification in LLMs
von: Emde, Cornelius, et al.
Veröffentlicht: (2025)
von: Emde, Cornelius, et al.
Veröffentlicht: (2025)
Focus On This, Not That! Steering LLMs with Adaptive Feature Specification
von: Lamb, Tom A., et al.
Veröffentlicht: (2024)
von: Lamb, Tom A., et al.
Veröffentlicht: (2024)
Prompting a Pretrained Transformer Can Be a Universal Approximator
von: Petrov, Aleksandar, et al.
Veröffentlicht: (2024)
von: Petrov, Aleksandar, et al.
Veröffentlicht: (2024)
Bi-Factorial Preference Optimization: Balancing Safety-Helpfulness in Language Models
von: Zhang, Wenxuan, et al.
Veröffentlicht: (2024)
von: Zhang, Wenxuan, et al.
Veröffentlicht: (2024)
FORCE: Transferable Visual Jailbreaking Attacks via Feature Over-Reliance CorrEction
von: Lin, Runqi, et al.
Veröffentlicht: (2025)
von: Lin, Runqi, et al.
Veröffentlicht: (2025)
Detecting LLM Hallucination Through Layer-wise Information Deficiency: Analysis of Ambiguous Prompts and Unanswerable Questions
von: Kim, Hazel, et al.
Veröffentlicht: (2024)
von: Kim, Hazel, et al.
Veröffentlicht: (2024)
Cross-Layer Attention Probing for Fine-Grained Hallucination Detection
von: Suresh, Malavika, et al.
Veröffentlicht: (2025)
von: Suresh, Malavika, et al.
Veröffentlicht: (2025)
On the Coexistence and Ensembling of Watermarks
von: Petrov, Aleksandar, et al.
Veröffentlicht: (2025)
von: Petrov, Aleksandar, et al.
Veröffentlicht: (2025)
On Pretraining Data Diversity for Self-Supervised Learning
von: Hammoud, Hasan Abed Al Kader, et al.
Veröffentlicht: (2024)
von: Hammoud, Hasan Abed Al Kader, et al.
Veröffentlicht: (2024)
Efficient Few-Shot Continual Learning in Vision-Language Models
von: Panos, Aristeidis, et al.
Veröffentlicht: (2025)
von: Panos, Aristeidis, et al.
Veröffentlicht: (2025)
Unforgotten Safety: Preserving Safety Alignment of Large Language Models with Continual Learning
von: Alssum, Lama, et al.
Veröffentlicht: (2025)
von: Alssum, Lama, et al.
Veröffentlicht: (2025)
It's a TRAP! Task-Redirecting Agent Persuasion Benchmark for Web Agents
von: Korgul, Karolina, et al.
Veröffentlicht: (2025)
von: Korgul, Karolina, et al.
Veröffentlicht: (2025)
Simple Baselines are Competitive with Code Evolution
von: Gideoni, Yonatan, et al.
Veröffentlicht: (2026)
von: Gideoni, Yonatan, et al.
Veröffentlicht: (2026)
Ego: Embedding-Guided Personalization of Vision-Language Models
von: Seifi, Soroush, et al.
Veröffentlicht: (2026)
von: Seifi, Soroush, et al.
Veröffentlicht: (2026)
Rethinking Safety in LLM Fine-tuning: An Optimization Perspective
von: Kim, Minseon, et al.
Veröffentlicht: (2025)
von: Kim, Minseon, et al.
Veröffentlicht: (2025)
Language Models Change Facts Based on the Way You Talk
von: Kearney, Matthew, et al.
Veröffentlicht: (2025)
von: Kearney, Matthew, et al.
Veröffentlicht: (2025)
Do Multilingual LLMs Think In English?
von: Schut, Lisa, et al.
Veröffentlicht: (2025)
von: Schut, Lisa, et al.
Veröffentlicht: (2025)
In-Context Learning Learns Label Relationships but Is Not Conventional Learning
von: Kossen, Jannik, et al.
Veröffentlicht: (2023)
von: Kossen, Jannik, et al.
Veröffentlicht: (2023)
Stabilizing Policy Gradients for Sample-Efficient Reinforcement Learning in LLM Reasoning
von: Melo, Luckeciano C., et al.
Veröffentlicht: (2025)
von: Melo, Luckeciano C., et al.
Veröffentlicht: (2025)
Temporal-Difference Variational Continual Learning
von: Melo, Luckeciano C., et al.
Veröffentlicht: (2024)
von: Melo, Luckeciano C., et al.
Veröffentlicht: (2024)
Imperfect Vision Encoders: Efficient and Robust Tuning for Vision-Language Models
von: Panos, Aristeidis, et al.
Veröffentlicht: (2024)
von: Panos, Aristeidis, et al.
Veröffentlicht: (2024)
Illusory Attacks: Information-Theoretic Detectability Matters in Adversarial Attacks
von: Franzmeyer, Tim, et al.
Veröffentlicht: (2022)
von: Franzmeyer, Tim, et al.
Veröffentlicht: (2022)
SynthCLIP: Are We Ready for a Fully Synthetic CLIP Training?
von: Hammoud, Hasan Abed Al Kader, et al.
Veröffentlicht: (2024)
von: Hammoud, Hasan Abed Al Kader, et al.
Veröffentlicht: (2024)
Evaluating the Impact of Post-Training Quantization on Reliable VQA with Multimodal LLMs
von: Kurz, Paul Jonas, et al.
Veröffentlicht: (2026)
von: Kurz, Paul Jonas, et al.
Veröffentlicht: (2026)
Model Merging and Safety Alignment: One Bad Model Spoils the Bunch
von: Hammoud, Hasan Abed Al Kader, et al.
Veröffentlicht: (2024)
von: Hammoud, Hasan Abed Al Kader, et al.
Veröffentlicht: (2024)
SEM: Sparse Embedding Modulation for Post-Hoc Debiasing of Vision-Language Models
von: Guimard, Quentin, et al.
Veröffentlicht: (2026)
von: Guimard, Quentin, et al.
Veröffentlicht: (2026)
OMNI: Open-endedness via Models of human Notions of Interestingness
von: Zhang, Jenny, et al.
Veröffentlicht: (2023)
von: Zhang, Jenny, et al.
Veröffentlicht: (2023)
The Authorization-Execution Gap Is a Major Safety and Security Problem in Open-World Agents
von: Wu, Baoyuan, et al.
Veröffentlicht: (2026)
von: Wu, Baoyuan, et al.
Veröffentlicht: (2026)
FedMedICL: Towards Holistic Evaluation of Distribution Shifts in Federated Medical Imaging
von: Alhamoud, Kumail, et al.
Veröffentlicht: (2024)
von: Alhamoud, Kumail, et al.
Veröffentlicht: (2024)
Kernel Language Entropy: Fine-grained Uncertainty Quantification for LLMs from Semantic Similarities
von: Nikitin, Alexander, et al.
Veröffentlicht: (2024)
von: Nikitin, Alexander, et al.
Veröffentlicht: (2024)
The Phantom Menace: Unmasking Privacy Leakages in Vision-Language Models
von: Caldarella, Simone, et al.
Veröffentlicht: (2024)
von: Caldarella, Simone, et al.
Veröffentlicht: (2024)
OMNI-EPIC: Open-endedness via Models of human Notions of Interestingness with Environments Programmed in Code
von: Faldor, Maxence, et al.
Veröffentlicht: (2024)
von: Faldor, Maxence, et al.
Veröffentlicht: (2024)
Detecting Multi-Agent Collusion Through Multi-Agent Interpretability
von: Rose, Aaron, et al.
Veröffentlicht: (2026)
von: Rose, Aaron, et al.
Veröffentlicht: (2026)
MAD-Sherlock: Multi-Agent Debate for Visual Misinformation Detection
von: Lakara, Kumud, et al.
Veröffentlicht: (2024)
von: Lakara, Kumud, et al.
Veröffentlicht: (2024)
AgentOrchestra: Orchestrating Multi-Agent Intelligence with the Tool-Environment-Agent(TEA) Protocol
von: Zhang, Wentao, et al.
Veröffentlicht: (2025)
von: Zhang, Wentao, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
ToolTweak: An Attack on Tool Selection in LLM-based Agents
von: Sneh, Jonathan, et al.
Veröffentlicht: (2025) -
BiasBusters: Uncovering and Mitigating Tool Selection Bias in Large Language Models
von: Blankenstein, Thierry, et al.
Veröffentlicht: (2025) -
MIP against Agent: Malicious Image Patches Hijacking Multimodal OS Agents
von: Aichberger, Lukas, et al.
Veröffentlicht: (2025) -
Universal In-Context Approximation By Prompting Fully Recurrent Models
von: Petrov, Aleksandar, et al.
Veröffentlicht: (2024) -
Memo: Training Memory-Efficient Embodied Agents with Reinforcement Learning
von: Gupta, Gunshi, et al.
Veröffentlicht: (2025)