Saved in:
| Main Authors: | Seip, Dominik, Hein, Matthias |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | https://arxiv.org/abs/2604.08005 |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Visual Memory Injection Attacks for Multi-Turn Conversations
by: Schlarmann, Christian, et al.
Published: (2026)
by: Schlarmann, Christian, et al.
Published: (2026)
SafeRedirect: Defeating Internal Safety Collapse via Task-Completion Redirection in Frontier LLMs
by: Pan, Chao, et al.
Published: (2026)
by: Pan, Chao, et al.
Published: (2026)
LLM Unlearning via Neural Activation Redirection
by: Shen, William F., et al.
Published: (2025)
by: Shen, William F., et al.
Published: (2025)
Inference-Time Machine Unlearning via Gated Activation Redirection
by: Turani, Vinícius Conte, et al.
Published: (2026)
by: Turani, Vinícius Conte, et al.
Published: (2026)
Mahalanobis++: Improving OOD Detection via Feature Normalization
by: Mueller, Maximilian, et al.
Published: (2025)
by: Mueller, Maximilian, et al.
Published: (2025)
LoGex: Improved tail detection of extremely rare histopathology classes via guided diffusion
by: Mueller, Maximilian, et al.
Published: (2024)
by: Mueller, Maximilian, et al.
Published: (2024)
Bias of Stochastic Gradient Descent or the Architecture: Disentangling the Effects of Overparameterization of Neural Networks
by: Peleg, Amit, et al.
Published: (2024)
by: Peleg, Amit, et al.
Published: (2024)
Targeted Manipulation: Slope-Based Attacks on Financial Time-Series Data
by: Luszczynski, Dominik
Published: (2025)
by: Luszczynski, Dominik
Published: (2025)
Multi-Agent Computer Use
by: Koh, Jing Yu, et al.
Published: (2026)
by: Koh, Jing Yu, et al.
Published: (2026)
ReTrack: Data Unlearning in Diffusion Models through Redirecting the Denoising Trajectory
by: Shi, Qitan, et al.
Published: (2025)
by: Shi, Qitan, et al.
Published: (2025)
Grounding Computer Use Agents on Human Demonstrations
by: Feizi, Aarash, et al.
Published: (2025)
by: Feizi, Aarash, et al.
Published: (2025)
Redirection for Erasing Memory (REM): Towards a universal unlearning method for corrupted data
by: Schoepf, Stefan, et al.
Published: (2025)
by: Schoepf, Stefan, et al.
Published: (2025)
Efficient Agent Training for Computer Use
by: He, Yanheng, et al.
Published: (2025)
by: He, Yanheng, et al.
Published: (2025)
Spectral Guardrails for Agents in the Wild: Detecting Tool Use Hallucinations via Attention Topology
by: Noël, Valentin
Published: (2026)
by: Noël, Valentin
Published: (2026)
How to train your ViT for OOD Detection
by: Mueller, Maximilian, et al.
Published: (2024)
by: Mueller, Maximilian, et al.
Published: (2024)
VerificAgent: Domain-Specific Memory Verification for Scalable Oversight of Aligned Computer-Use Agents
by: Nguyen, Thong Q., et al.
Published: (2025)
by: Nguyen, Thong Q., et al.
Published: (2025)
MACTAS: Self-Attention-Based Inter-Agent Communication in Multi-Agent Reinforcement Learning with Action-Value Function Decomposition
by: Wojtala, Maciej, et al.
Published: (2025)
by: Wojtala, Maciej, et al.
Published: (2025)
Scaling Agents for Computer Use
by: Gonzalez-Pumariega, Gonzalo, et al.
Published: (2025)
by: Gonzalez-Pumariega, Gonzalo, et al.
Published: (2025)
FuseLIP: Multimodal Embeddings via Early Fusion of Discrete Tokens
by: Schlarmann, Christian, et al.
Published: (2025)
by: Schlarmann, Christian, et al.
Published: (2025)
RePOPE: Impact of Annotation Errors on the POPE Benchmark
by: Neuhaus, Yannic, et al.
Published: (2025)
by: Neuhaus, Yannic, et al.
Published: (2025)
OSWorld-Human: Benchmarking the Efficiency of Computer-Use Agents
by: Abhyankar, Reyna, et al.
Published: (2025)
by: Abhyankar, Reyna, et al.
Published: (2025)
Proportional Aggregation of Preferences for Sequential Decision Making
by: Chandak, Nikhil, et al.
Published: (2023)
by: Chandak, Nikhil, et al.
Published: (2023)
Parametrized Multi-Agent Routing via Deep Attention Models
by: Basiri, Salar, et al.
Published: (2025)
by: Basiri, Salar, et al.
Published: (2025)
Flow-Attentional Graph Neural Networks
by: Plettenberg, Pascal, et al.
Published: (2025)
by: Plettenberg, Pascal, et al.
Published: (2025)
Memory Injection Attacks on LLM Agents via Query-Only Interaction
by: Dong, Shen, et al.
Published: (2025)
by: Dong, Shen, et al.
Published: (2025)
ROGUE: Misaligned Agent Behavior Arising from Ordinary Computer Use
by: Tien, Jeremy, et al.
Published: (2026)
by: Tien, Jeremy, et al.
Published: (2026)
OS-Harm: A Benchmark for Measuring Safety of Computer Use Agents
by: Kuntz, Thomas, et al.
Published: (2025)
by: Kuntz, Thomas, et al.
Published: (2025)
Programming with Pixels: Can Computer-Use Agents do Software Engineering?
by: Aggarwal, Pranjal, et al.
Published: (2025)
by: Aggarwal, Pranjal, et al.
Published: (2025)
Dual-Agent Co-Training for Health Coaching via Implicit Adversarial Preference Optimization
by: Long, Da, et al.
Published: (2026)
by: Long, Da, et al.
Published: (2026)
LoRA-Ensemble: Efficient Uncertainty Modelling for Self-Attention Networks
by: Mühlematter, Dominik J., et al.
Published: (2024)
by: Mühlematter, Dominik J., et al.
Published: (2024)
Preference Poisoning Attacks on Reward Model Learning
by: Wu, Junlin, et al.
Published: (2024)
by: Wu, Junlin, et al.
Published: (2024)
Attack Smarter: Attention-Driven Fine-Grained Webpage Fingerprinting Attacks
by: Yuan, Yali, et al.
Published: (2025)
by: Yuan, Yali, et al.
Published: (2025)
Unlearning That Lasts: Utility-Preserving, Robust, and Almost Irreversible Forgetting in LLMs
by: Singh, Naman Deep, et al.
Published: (2025)
by: Singh, Naman Deep, et al.
Published: (2025)
Automated Membership Inference Attacks: Discovering MIA Signal Computations using LLM Agents
by: Tran, Toan, et al.
Published: (2026)
by: Tran, Toan, et al.
Published: (2026)
CUA-Gym: Scaling Verifiable Training Environments and Tasks for Computer-Use Agents
by: Wang, Bowen, et al.
Published: (2026)
by: Wang, Bowen, et al.
Published: (2026)
EPO: Hierarchical LLM Agents with Environment Preference Optimization
by: Zhao, Qi, et al.
Published: (2024)
by: Zhao, Qi, et al.
Published: (2024)
Computationally lightweight classifiers with frequentist bounds on predictions
by: Murali, Shreeram, et al.
Published: (2026)
by: Murali, Shreeram, et al.
Published: (2026)
VideoAgentTrek: Computer Use Pretraining from Unlabeled Videos
by: Lu, Dunjie, et al.
Published: (2025)
by: Lu, Dunjie, et al.
Published: (2025)
WebSTAR: Scalable Data Synthesis for Computer Use Agents with Step-Level Filtering
by: He, Yifei, et al.
Published: (2025)
by: He, Yifei, et al.
Published: (2025)
Advancing Compositional Awareness in CLIP with Efficient Fine-Tuning
by: Peleg, Amit, et al.
Published: (2025)
by: Peleg, Amit, et al.
Published: (2025)
Similar Items
-
Visual Memory Injection Attacks for Multi-Turn Conversations
by: Schlarmann, Christian, et al.
Published: (2026) -
SafeRedirect: Defeating Internal Safety Collapse via Task-Completion Redirection in Frontier LLMs
by: Pan, Chao, et al.
Published: (2026) -
LLM Unlearning via Neural Activation Redirection
by: Shen, William F., et al.
Published: (2025) -
Inference-Time Machine Unlearning via Gated Activation Redirection
by: Turani, Vinícius Conte, et al.
Published: (2026) -
Mahalanobis++: Improving OOD Detection via Feature Normalization
by: Mueller, Maximilian, et al.
Published: (2025)