ProjGuard: Safety Monitoring for Computer-Use Agents via Low-Dimensional Projections
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Contreras, Kebin, Hinojosa, Carlos, Bacca, Jorge, Ghanem, Bernard |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
SAVeS: Steering Safety Judgments in Vision-Language Models via Semantic Cues
von: Hinojosa, Carlos, et al.
Veröffentlicht: (2026)
von: Hinojosa, Carlos, et al.
Veröffentlicht: (2026)
Autoregressive High-Order Finite Difference Modulo Imaging: High-Dynamic Range for Computer Vision Applications
von: Monroy, Brayan, et al.
Veröffentlicht: (2025)
von: Monroy, Brayan, et al.
Veröffentlicht: (2025)
High Dynamic Range Modulo Imaging for Robust Object Detection in Autonomous Driving
von: Contreras, Kebin, et al.
Veröffentlicht: (2025)
von: Contreras, Kebin, et al.
Veröffentlicht: (2025)
Multimodal Safety Evaluation in Generative Agent Social Simulations
von: Vera, Alhim, et al.
Veröffentlicht: (2025)
von: Vera, Alhim, et al.
Veröffentlicht: (2025)
ColorMAE: Exploring data-independent masking strategies in Masked AutoEncoders
von: Hinojosa, Carlos, et al.
Veröffentlicht: (2024)
von: Hinojosa, Carlos, et al.
Veröffentlicht: (2024)
See the past: Time-Reversed Scene Reconstruction from Thermal Traces Using Visual Language Models
von: Contreras, Kebin, et al.
Veröffentlicht: (2025)
von: Contreras, Kebin, et al.
Veröffentlicht: (2025)
Projection-Based Correction for Enhancing Deep Inverse Networks
von: Bacca, Jorge
Veröffentlicht: (2025)
von: Bacca, Jorge
Veröffentlicht: (2025)
Learning Semantic Segmentation with Query Points Supervision on Aerial Images
von: Rivier, Santiago, et al.
Veröffentlicht: (2023)
von: Rivier, Santiago, et al.
Veröffentlicht: (2023)
MambaStyle: Efficient StyleGAN Inversion for Real Image Editing with State-Space Models
von: Lopez, Jhon, et al.
Veröffentlicht: (2025)
von: Lopez, Jhon, et al.
Veröffentlicht: (2025)
Privacy-preserving Optics for Enhancing Protection in Face De-identification
von: Lopez, Jhon, et al.
Veröffentlicht: (2024)
von: Lopez, Jhon, et al.
Veröffentlicht: (2024)
LiftProj: Space Lifting and Projection-Based Panorama Stitching
von: Jia, Yuan, et al.
Veröffentlicht: (2025)
von: Jia, Yuan, et al.
Veröffentlicht: (2025)
CAPTAIN: Semantic Feature Injection for Memorization Mitigation in Text-to-Image Diffusion Models
von: Zhang, Tong, et al.
Veröffentlicht: (2025)
von: Zhang, Tong, et al.
Veröffentlicht: (2025)
MD-ProjTex: Texturing 3D Shapes with Multi-Diffusion Projection
von: Yildirim, Ahmet Burak, et al.
Veröffentlicht: (2025)
von: Yildirim, Ahmet Burak, et al.
Veröffentlicht: (2025)
BraveGuard: From Open-World Threats to Safer Computer-Use Agents
von: Feng, Yunhao, et al.
Veröffentlicht: (2026)
von: Feng, Yunhao, et al.
Veröffentlicht: (2026)
UWB-PostureGuard: A Privacy-Preserving RF Sensing System for Continuous Ergonomic Sitting Posture Monitoring
von: Li, Haotang, et al.
Veröffentlicht: (2025)
von: Li, Haotang, et al.
Veröffentlicht: (2025)
ProjFlow: Projection Sampling with Flow Matching for Zero-Shot Exact Spatial Motion Control
von: Watanabe, Akihisa, et al.
Veröffentlicht: (2026)
von: Watanabe, Akihisa, et al.
Veröffentlicht: (2026)
UnMix-NeRF: Spectral Unmixing Meets Neural Radiance Fields
von: Perez, Fabian, et al.
Veröffentlicht: (2025)
von: Perez, Fabian, et al.
Veröffentlicht: (2025)
CO2Wounds-V2: Extended Chronic Wounds Dataset From Leprosy Patients
von: Sanchez, Karen, et al.
Veröffentlicht: (2024)
von: Sanchez, Karen, et al.
Veröffentlicht: (2024)
Can One Safety Loop Guard Them All? Agentic Guard Rails for Federated Computing
von: Veeraragavan, Narasimha Raghavan, et al.
Veröffentlicht: (2025)
von: Veeraragavan, Narasimha Raghavan, et al.
Veröffentlicht: (2025)
ProjDevBench: Benchmarking AI Coding Agents on End-to-End Project Development
von: Lu, Pengrui, et al.
Veröffentlicht: (2026)
von: Lu, Pengrui, et al.
Veröffentlicht: (2026)
ProbGuard: Probabilistic Runtime Monitoring for LLM Agent Safety
von: Wang, Haoyu, et al.
Veröffentlicht: (2025)
von: Wang, Haoyu, et al.
Veröffentlicht: (2025)
Turning the Spell Around: Lightweight Alignment Amplification via Rank-One Safety Injection
von: Shairah, Harethah Abu, et al.
Veröffentlicht: (2025)
von: Shairah, Harethah Abu, et al.
Veröffentlicht: (2025)
Uncertainty Visualization via Low-Dimensional Posterior Projections
von: Yair, Omer, et al.
Veröffentlicht: (2023)
von: Yair, Omer, et al.
Veröffentlicht: (2023)
ProjLens: Unveiling the Role of Projectors in Multimodal Model Safety
von: Wang, Kun, et al.
Veröffentlicht: (2026)
von: Wang, Kun, et al.
Veröffentlicht: (2026)
ProjQ: Project-and-Quantize for Adapter-Aware LLM Compression
von: Yu, Wneya, et al.
Veröffentlicht: (2026)
von: Yu, Wneya, et al.
Veröffentlicht: (2026)
Ultra-Low-Dimensional Prompt Tuning via Random Projection
von: Wu, Zijun, et al.
Veröffentlicht: (2025)
von: Wu, Zijun, et al.
Veröffentlicht: (2025)
CURE: Curriculum-guided Multi-task Training for Reliable Anatomy Grounded Report Generation
von: Messina, Pablo, et al.
Veröffentlicht: (2026)
von: Messina, Pablo, et al.
Veröffentlicht: (2026)
Unforgotten Safety: Preserving Safety Alignment of Large Language Models with Continual Learning
von: Alssum, Lama, et al.
Veröffentlicht: (2025)
von: Alssum, Lama, et al.
Veröffentlicht: (2025)
HarmonyGuard: Toward Safety and Utility in Web Agents via Adaptive Policy Enhancement and Dual-Objective Optimization
von: Chen, Yurun, et al.
Veröffentlicht: (2025)
von: Chen, Yurun, et al.
Veröffentlicht: (2025)
The mechanization of science illustrated by the Lean formalization of the multi-graded Proj construction
von: Mayeux, Arnaud, et al.
Veröffentlicht: (2025)
von: Mayeux, Arnaud, et al.
Veröffentlicht: (2025)
DialogGuard: Multi-Agent Psychosocial Safety Evaluation of Sensitive LLM Responses
von: Luo, Han, et al.
Veröffentlicht: (2025)
von: Luo, Han, et al.
Veröffentlicht: (2025)
Formalizing multi-graded Brenner-Schröer Proj schemes and dilatations of rings in Lean4
von: Mayeux, Arnaud, et al.
Veröffentlicht: (2026)
von: Mayeux, Arnaud, et al.
Veröffentlicht: (2026)
Reasoning Vectors: Transferring Chain-of-Thought Capabilities via Task Arithmetic
von: Zbeeb, Mohammad, et al.
Veröffentlicht: (2025)
von: Zbeeb, Mohammad, et al.
Veröffentlicht: (2025)
CultureGuard: Towards Culturally-Aware Dataset and Guard Model for Multilingual Safety Applications
von: Joshi, Raviraj, et al.
Veröffentlicht: (2025)
von: Joshi, Raviraj, et al.
Veröffentlicht: (2025)
ConceptGuard: Neuro-Symbolic Safety Guardrails via Sparse Interpretable Jailbreak Concepts
von: Aswal, Darpan, et al.
Veröffentlicht: (2025)
von: Aswal, Darpan, et al.
Veröffentlicht: (2025)
Hadamard Row-Wise Generation Algorithm
von: Monroy, Brayan, et al.
Veröffentlicht: (2024)
von: Monroy, Brayan, et al.
Veröffentlicht: (2024)
Video Self-Stitching Graph Network for Temporal Action Localization
von: Zhao, Chen, et al.
Veröffentlicht: (2020)
von: Zhao, Chen, et al.
Veröffentlicht: (2020)
On multi-graded Proj schemes
von: Mayeux, Arnaud, et al.
Veröffentlicht: (2023)
von: Mayeux, Arnaud, et al.
Veröffentlicht: (2023)
FineLIP: Extending CLIP's Reach via Fine-Grained Alignment with Longer Text Inputs
von: Asokan, Mothilal, et al.
Veröffentlicht: (2025)
von: Asokan, Mothilal, et al.
Veröffentlicht: (2025)
Behind the Magic, MERLIM: Multi-modal Evaluation Benchmark for Large Image-Language Models
von: Villa, Andrés, et al.
Veröffentlicht: (2023)
von: Villa, Andrés, et al.
Veröffentlicht: (2023)
Ähnliche Einträge
-
SAVeS: Steering Safety Judgments in Vision-Language Models via Semantic Cues
von: Hinojosa, Carlos, et al.
Veröffentlicht: (2026) -
Autoregressive High-Order Finite Difference Modulo Imaging: High-Dynamic Range for Computer Vision Applications
von: Monroy, Brayan, et al.
Veröffentlicht: (2025) -
High Dynamic Range Modulo Imaging for Robust Object Detection in Autonomous Driving
von: Contreras, Kebin, et al.
Veröffentlicht: (2025) -
Multimodal Safety Evaluation in Generative Agent Social Simulations
von: Vera, Alhim, et al.
Veröffentlicht: (2025) -
ColorMAE: Exploring data-independent masking strategies in Masked AutoEncoders
von: Hinojosa, Carlos, et al.
Veröffentlicht: (2024)