TreeTeaming: Autonomous Red-Teaming of Vision-Language Models via Hierarchical Strategy Exploration
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Li, Chunxiao, Li, Lijun, Shao, Jing |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Red-Teaming Segment Anything Model
von: Jankowski, Krzysztof, et al.
Veröffentlicht: (2024)
von: Jankowski, Krzysztof, et al.
Veröffentlicht: (2024)
BlueSuffix: Reinforced Blue Teaming for Vision-Language Models Against Jailbreak Attacks
von: Zhao, Yunhan, et al.
Veröffentlicht: (2024)
von: Zhao, Yunhan, et al.
Veröffentlicht: (2024)
Visual Exclusivity Attacks: Automatic Multimodal Red Teaming via Agentic Planning
von: Zhang, Yunbei, et al.
Veröffentlicht: (2026)
von: Zhang, Yunbei, et al.
Veröffentlicht: (2026)
Red Teaming Visual Language Models
von: Li, Mukai, et al.
Veröffentlicht: (2024)
von: Li, Mukai, et al.
Veröffentlicht: (2024)
Red-Teaming Text-to-Image Models via In-Context Experience Replay and Semantic-Preserving Prompt Rewriting
von: Chin, Zhi-Yi, et al.
Veröffentlicht: (2024)
von: Chin, Zhi-Yi, et al.
Veröffentlicht: (2024)
Analyzing Hierarchical Structure in Vision Models with Sparse Autoencoders
von: Olson, Matthew Lyle, et al.
Veröffentlicht: (2025)
von: Olson, Matthew Lyle, et al.
Veröffentlicht: (2025)
HarmBench: A Standardized Evaluation Framework for Automated Red Teaming and Robust Refusal
von: Mazeika, Mantas, et al.
Veröffentlicht: (2024)
von: Mazeika, Mantas, et al.
Veröffentlicht: (2024)
Red-Teaming Text-to-Image Systems by Rule-based Preference Modeling
von: Cao, Yichuan, et al.
Veröffentlicht: (2025)
von: Cao, Yichuan, et al.
Veröffentlicht: (2025)
Rethinking Pruning for Vision-Language Models: Strategies for Effective Sparsity and Performance Restoration
von: He, Shwai, et al.
Veröffentlicht: (2024)
von: He, Shwai, et al.
Veröffentlicht: (2024)
Towards Compatible Fine-tuning for Vision-Language Model Updates
von: Wang, Zhengbo, et al.
Veröffentlicht: (2024)
von: Wang, Zhengbo, et al.
Veröffentlicht: (2024)
Universal Camouflage Attack on Vision-Language Models for Autonomous Driving
von: Kong, Dehong, et al.
Veröffentlicht: (2025)
von: Kong, Dehong, et al.
Veröffentlicht: (2025)
Coherent Multi-Agent Trajectory Forecasting in Team Sports with CausalTraj
von: Teoh, Wei Zhen
Veröffentlicht: (2025)
von: Teoh, Wei Zhen
Veröffentlicht: (2025)
Uncovering Linguistic Fragility in Vision-Language-Action Models via Diversity-Aware Red Teaming
von: Tong, Baoshun, et al.
Veröffentlicht: (2026)
von: Tong, Baoshun, et al.
Veröffentlicht: (2026)
The Illusion of Progress? A Critical Look at Test-Time Adaptation for Vision-Language Models
von: Sheng, Lijun, et al.
Veröffentlicht: (2025)
von: Sheng, Lijun, et al.
Veröffentlicht: (2025)
AutoTrust: Benchmarking Trustworthiness in Large Vision Language Models for Autonomous Driving
von: Xing, Shuo, et al.
Veröffentlicht: (2024)
von: Xing, Shuo, et al.
Veröffentlicht: (2024)
Active Zero: Self-Evolving Vision-Language Models through Active Environment Exploration
von: He, Jinghan, et al.
Veröffentlicht: (2026)
von: He, Jinghan, et al.
Veröffentlicht: (2026)
Hierarchical Safety Realignment: Lightweight Restoration of Safety in Pruned Large Vision-Language Models
von: Li, Yue, et al.
Veröffentlicht: (2025)
von: Li, Yue, et al.
Veröffentlicht: (2025)
Toward Autonomous Laboratory Safety Monitoring with Vision Language Models: Learning to See Hazards Through Scene Structure
von: Chakraborty, Trishna, et al.
Veröffentlicht: (2026)
von: Chakraborty, Trishna, et al.
Veröffentlicht: (2026)
AgrI Challenge: A Data-Centric AI Competition for Cross-Team Validation in Agricultural Vision
von: Brahimi, Mohammed, et al.
Veröffentlicht: (2026)
von: Brahimi, Mohammed, et al.
Veröffentlicht: (2026)
Iwin Transformer: Hierarchical Vision Transformer using Interleaved Windows
von: Huo, Simin, et al.
Veröffentlicht: (2025)
von: Huo, Simin, et al.
Veröffentlicht: (2025)
Continual Learning with Vision-Language Models via Semantic-Geometry Preservation
von: He, Chiyuan, et al.
Veröffentlicht: (2026)
von: He, Chiyuan, et al.
Veröffentlicht: (2026)
VLM-AD: End-to-End Autonomous Driving through Vision-Language Model Supervision
von: Xu, Yi, et al.
Veröffentlicht: (2024)
von: Xu, Yi, et al.
Veröffentlicht: (2024)
Vision-Language Models are Strong Noisy Label Detectors
von: Wei, Tong, et al.
Veröffentlicht: (2024)
von: Wei, Tong, et al.
Veröffentlicht: (2024)
Hierarchical Uncertainty Exploration via Feedforward Posterior Trees
von: Nehme, Elias, et al.
Veröffentlicht: (2024)
von: Nehme, Elias, et al.
Veröffentlicht: (2024)
Imagine, Verify, Execute: Memory-guided Agentic Exploration with Vision-Language Models
von: Lee, Seungjae, et al.
Veröffentlicht: (2025)
von: Lee, Seungjae, et al.
Veröffentlicht: (2025)
From Understanding to Engagement: Personalized pharmacy Video Clips via Vision Language Models (VLMs)
von: Mishra, Suyash, et al.
Veröffentlicht: (2026)
von: Mishra, Suyash, et al.
Veröffentlicht: (2026)
IMITATE: Clinical Prior Guided Hierarchical Vision-Language Pre-training
von: Liu, Che, et al.
Veröffentlicht: (2023)
von: Liu, Che, et al.
Veröffentlicht: (2023)
FRISM: Fine-Grained Reasoning Injection via Subspace-Level Model Merging for Vision-Language Models
von: Huang, Chenyu, et al.
Veröffentlicht: (2026)
von: Huang, Chenyu, et al.
Veröffentlicht: (2026)
GRPO-TTA: Test-Time Visual Tuning for Vision-Language Models via GRPO-Driven Reinforcement Learning
von: Li, Yujun, et al.
Veröffentlicht: (2026)
von: Li, Yujun, et al.
Veröffentlicht: (2026)
Learning Self-Correction in Vision-Language Models via Rollout Augmentation
von: Ding, Yi, et al.
Veröffentlicht: (2026)
von: Ding, Yi, et al.
Veröffentlicht: (2026)
EmboTeam: Grounding LLM Reasoning into Reactive Behavior Trees via PDDL for Embodied Multi-Robot Collaboration
von: Zeng, Haishan, et al.
Veröffentlicht: (2026)
von: Zeng, Haishan, et al.
Veröffentlicht: (2026)
MM-Zero: Self-Evolving Multi-Model Vision Language Models From Zero Data
von: Li, Zongxia, et al.
Veröffentlicht: (2026)
von: Li, Zongxia, et al.
Veröffentlicht: (2026)
Post-hoc Probabilistic Vision-Language Models
von: Baumann, Anton, et al.
Veröffentlicht: (2024)
von: Baumann, Anton, et al.
Veröffentlicht: (2024)
Adversarial Nibbler: An Open Red-Teaming Method for Identifying Diverse Harms in Text-to-Image Generation
von: Quaye, Jessica, et al.
Veröffentlicht: (2024)
von: Quaye, Jessica, et al.
Veröffentlicht: (2024)
Re-Align: Aligning Vision Language Models via Retrieval-Augmented Direct Preference Optimization
von: Xing, Shuo, et al.
Veröffentlicht: (2025)
von: Xing, Shuo, et al.
Veröffentlicht: (2025)
Tree of Attributes Prompt Learning for Vision-Language Models
von: Ding, Tong, et al.
Veröffentlicht: (2024)
von: Ding, Tong, et al.
Veröffentlicht: (2024)
Application of Vision-Language Model to Pedestrians Behavior and Scene Understanding in Autonomous Driving
von: Gao, Haoxiang, et al.
Veröffentlicht: (2025)
von: Gao, Haoxiang, et al.
Veröffentlicht: (2025)
Assessing the Visual Enumeration Abilities of Specialized Counting Architectures and Vision-Language Models
von: Hou, Kuinan, et al.
Veröffentlicht: (2025)
von: Hou, Kuinan, et al.
Veröffentlicht: (2025)
Cross-Modal Adapter: Parameter-Efficient Transfer Learning Approach for Vision-Language Models
von: Yang, Juncheng, et al.
Veröffentlicht: (2024)
von: Yang, Juncheng, et al.
Veröffentlicht: (2024)
Adapting Vision-Language Models Without Labels: A Comprehensive Survey
von: Dong, Hao, et al.
Veröffentlicht: (2025)
von: Dong, Hao, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Red-Teaming Segment Anything Model
von: Jankowski, Krzysztof, et al.
Veröffentlicht: (2024) -
BlueSuffix: Reinforced Blue Teaming for Vision-Language Models Against Jailbreak Attacks
von: Zhao, Yunhan, et al.
Veröffentlicht: (2024) -
Visual Exclusivity Attacks: Automatic Multimodal Red Teaming via Agentic Planning
von: Zhang, Yunbei, et al.
Veröffentlicht: (2026) -
Red Teaming Visual Language Models
von: Li, Mukai, et al.
Veröffentlicht: (2024) -
Red-Teaming Text-to-Image Models via In-Context Experience Replay and Semantic-Preserving Prompt Rewriting
von: Chin, Zhi-Yi, et al.
Veröffentlicht: (2024)