HarmonyGuard: Toward Safety and Utility in Web Agents via Adaptive Policy Enhancement and Dual-Objective Optimization
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Chen, Yurun, Hu, Xavier, Liu, Yuhan, Yin, Keting, Li, Juncheng, Zhang, Zhuosheng, Zhang, Shengyu |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Evaluating the Robustness of Multimodal Agents Against Active Environmental Injection Attacks
von: Chen, Yurun, et al.
Veröffentlicht: (2025)
von: Chen, Yurun, et al.
Veröffentlicht: (2025)
SafePred: A Predictive Guardrail for Computer-Using Agents via World Models
von: Chen, Yurun, et al.
Veröffentlicht: (2026)
von: Chen, Yurun, et al.
Veröffentlicht: (2026)
Graph2Eval: Automatic Multimodal Task Generation for Agents via Knowledge Graphs
von: Chen, Yurun, et al.
Veröffentlicht: (2025)
von: Chen, Yurun, et al.
Veröffentlicht: (2025)
EcoAgent: An Efficient Device-Cloud Collaborative Multi-Agent Framework for Mobile Automation
von: Yi, Biao, et al.
Veröffentlicht: (2025)
von: Yi, Biao, et al.
Veröffentlicht: (2025)
AccKV: Towards Efficient Audio-Video LLMs Inference via Adaptive-Focusing and Cross-Calibration KV Cache Optimization
von: Jiang, Zhonghua, et al.
Veröffentlicht: (2025)
von: Jiang, Zhonghua, et al.
Veröffentlicht: (2025)
Compress to Focus: Efficient Coordinate Compression for Policy Optimization in Multi-Turn GUI Agents
von: Song, Yurun, et al.
Veröffentlicht: (2026)
von: Song, Yurun, et al.
Veröffentlicht: (2026)
World-Model-Augmented Web Agents with Action Correction
von: Shen, Zhouzhou, et al.
Veröffentlicht: (2026)
von: Shen, Zhouzhou, et al.
Veröffentlicht: (2026)
Personalized Federated Learning with Adaptive Feature Aggregation and Knowledge Transfer
von: Yin, Keting, et al.
Veröffentlicht: (2024)
von: Yin, Keting, et al.
Veröffentlicht: (2024)
Learning to Adapt: Self-Improving Web Agent via Cognitive-Aware Exploration
von: Chen, Weile, et al.
Veröffentlicht: (2026)
von: Chen, Weile, et al.
Veröffentlicht: (2026)
GUI-PRA: Process Reward Agent for GUI Tasks
von: Xiong, Tao, et al.
Veröffentlicht: (2025)
von: Xiong, Tao, et al.
Veröffentlicht: (2025)
AdaptiveGuard: Towards Adaptive Runtime Safety for LLM-Powered Software
von: Yang, Rui, et al.
Veröffentlicht: (2025)
von: Yang, Rui, et al.
Veröffentlicht: (2025)
Safety Optimized Reinforcement Learning via Multi-Objective Policy Optimization
von: Honari, Homayoun, et al.
Veröffentlicht: (2024)
von: Honari, Homayoun, et al.
Veröffentlicht: (2024)
Multimodal LLM-Guided Semantic Correction in Text-to-Image Diffusion
von: Lv, Zheqi, et al.
Veröffentlicht: (2025)
von: Lv, Zheqi, et al.
Veröffentlicht: (2025)
Gradient-Adaptive Policy Optimization: Towards Multi-Objective Alignment of Large Language Models
von: Li, Chengao, et al.
Veröffentlicht: (2025)
von: Li, Chengao, et al.
Veröffentlicht: (2025)
HTPO: Towards Exploration-Exploitation Balanced Policy Optimization via Hierarchical Token-level Objective Control
von: Yao, Xincheng, et al.
Veröffentlicht: (2026)
von: Yao, Xincheng, et al.
Veröffentlicht: (2026)
You Only Look at Screens: Multimodal Chain-of-Action Agents
von: Zhang, Zhuosheng, et al.
Veröffentlicht: (2023)
von: Zhang, Zhuosheng, et al.
Veröffentlicht: (2023)
Mixture of Reasonings: Teach Large Language Models to Reason with Adaptive Strategies
von: Xiong, Tao, et al.
Veröffentlicht: (2025)
von: Xiong, Tao, et al.
Veröffentlicht: (2025)
WebAgentGuard: A Reasoning-Driven Guard Model for Detecting Prompt Injection Attacks in Web Agents
von: Chen, Yulin, et al.
Veröffentlicht: (2026)
von: Chen, Yulin, et al.
Veröffentlicht: (2026)
Cognitive Duality for Adaptive Web Agents
von: Liu, Jiarun, et al.
Veröffentlicht: (2025)
von: Liu, Jiarun, et al.
Veröffentlicht: (2025)
Improving LLM Safety Alignment with Dual-Objective Optimization
von: Zhao, Xuandong, et al.
Veröffentlicht: (2025)
von: Zhao, Xuandong, et al.
Veröffentlicht: (2025)
WebGuard: Building a Generalizable Guardrail for Web Agents
von: Zheng, Boyuan, et al.
Veröffentlicht: (2025)
von: Zheng, Boyuan, et al.
Veröffentlicht: (2025)
Plan-over-Graph: Towards Parallelable LLM Agent Schedule
von: Zhang, Shiqi, et al.
Veröffentlicht: (2025)
von: Zhang, Shiqi, et al.
Veröffentlicht: (2025)
A Rolling Stone Gathers No Moss: Adaptive Policy Optimization for Stable Self-Evaluation in Large Multimodal Models
von: Wang, Wenkai, et al.
Veröffentlicht: (2025)
von: Wang, Wenkai, et al.
Veröffentlicht: (2025)
OS-Sentinel: Towards Safety-Enhanced Mobile GUI Agents via Hybrid Validation in Realistic Workflows
von: Sun, Qiushi, et al.
Veröffentlicht: (2025)
von: Sun, Qiushi, et al.
Veröffentlicht: (2025)
OS-Kairos: Adaptive Interaction for MLLM-Powered GUI Agents
von: Cheng, Pengzhou, et al.
Veröffentlicht: (2025)
von: Cheng, Pengzhou, et al.
Veröffentlicht: (2025)
Reinforce LLM Reasoning through Multi-Agent Reflection
von: Yuan, Yurun, et al.
Veröffentlicht: (2025)
von: Yuan, Yurun, et al.
Veröffentlicht: (2025)
SnapGuard: Lightweight Prompt Injection Detection for Screenshot-Based Web Agents
von: Du, Mengyao, et al.
Veröffentlicht: (2026)
von: Du, Mengyao, et al.
Veröffentlicht: (2026)
InfiGUI-G1: Advancing GUI Grounding with Adaptive Exploration Policy Optimization
von: Liu, Yuhang, et al.
Veröffentlicht: (2025)
von: Liu, Yuhang, et al.
Veröffentlicht: (2025)
CoCo-Agent: A Comprehensive Cognitive MLLM Agent for Smartphone GUI Automation
von: Ma, Xinbei, et al.
Veröffentlicht: (2024)
von: Ma, Xinbei, et al.
Veröffentlicht: (2024)
GuardAgent: Safeguard LLM Agents by a Guard Agent via Knowledge-Enabled Reasoning
von: Xiang, Zhen, et al.
Veröffentlicht: (2024)
von: Xiang, Zhen, et al.
Veröffentlicht: (2024)
Novel Object Synthesis via Adaptive Text-Image Harmony
von: Xiong, Zeren, et al.
Veröffentlicht: (2024)
von: Xiong, Zeren, et al.
Veröffentlicht: (2024)
ConsisGuard: Aligning Safety Deliberation with Policy Enforcement in LLM Guardrails
von: Wang, Yan, et al.
Veröffentlicht: (2026)
von: Wang, Yan, et al.
Veröffentlicht: (2026)
An Adaptive KKT-Based Indicator for Convergence Assessment in Multi-Objective Optimization
von: Santos, Thiago, et al.
Veröffentlicht: (2026)
von: Santos, Thiago, et al.
Veröffentlicht: (2026)
AGPO: Adaptive Group Policy Optimization with Dual Statistical Feedback
von: Hu, Miaobo, et al.
Veröffentlicht: (2026)
von: Hu, Miaobo, et al.
Veröffentlicht: (2026)
Fast-Slow Co-advancing Optimizer: Toward Harmonious Adversarial Training of GAN
von: Wang, Lin, et al.
Veröffentlicht: (2025)
von: Wang, Lin, et al.
Veröffentlicht: (2025)
Poly-Guard: Massive Multi-Domain Safety Policy-Grounded Guardrail Dataset
von: Kang, Mintong, et al.
Veröffentlicht: (2025)
von: Kang, Mintong, et al.
Veröffentlicht: (2025)
CultureGuard: Towards Culturally-Aware Dataset and Guard Model for Multilingual Safety Applications
von: Joshi, Raviraj, et al.
Veröffentlicht: (2025)
von: Joshi, Raviraj, et al.
Veröffentlicht: (2025)
Adaptive Social Learning via Mode Policy Optimization for Language Agents
von: Wang, Minzheng, et al.
Veröffentlicht: (2025)
von: Wang, Minzheng, et al.
Veröffentlicht: (2025)
GLaPE: Gold Label-agnostic Prompt Evaluation and Optimization for Large Language Model
von: Zhang, Xuanchang, et al.
Veröffentlicht: (2024)
von: Zhang, Xuanchang, et al.
Veröffentlicht: (2024)
VeriGuard: Enhancing LLM Agent Safety via Verified Code Generation
von: Miculicich, Lesly, et al.
Veröffentlicht: (2025)
von: Miculicich, Lesly, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Evaluating the Robustness of Multimodal Agents Against Active Environmental Injection Attacks
von: Chen, Yurun, et al.
Veröffentlicht: (2025) -
SafePred: A Predictive Guardrail for Computer-Using Agents via World Models
von: Chen, Yurun, et al.
Veröffentlicht: (2026) -
Graph2Eval: Automatic Multimodal Task Generation for Agents via Knowledge Graphs
von: Chen, Yurun, et al.
Veröffentlicht: (2025) -
EcoAgent: An Efficient Device-Cloud Collaborative Multi-Agent Framework for Mobile Automation
von: Yi, Biao, et al.
Veröffentlicht: (2025) -
AccKV: Towards Efficient Audio-Video LLMs Inference via Adaptive-Focusing and Cross-Calibration KV Cache Optimization
von: Jiang, Zhonghua, et al.
Veröffentlicht: (2025)