VideoSTF: Stress-Testing Output Repetition in Video Large Language Models
Fuente:
arXiv
Saved in:
| Main Authors: | Cao, Yuxin, Song, Wei, Xu, Shangzhi, Xue, Jingling, Dong, Jin Song |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Failures to Surface Harmful Contents in Video Large Language Models
by: Cao, Yuxin, et al.
Published: (2025)
by: Cao, Yuxin, et al.
Published: (2025)
Video Watermarking: Safeguarding Your Video from (Unauthorized) Annotations by Video-based LLMs
by: Li, Jinmin, et al.
Published: (2024)
by: Li, Jinmin, et al.
Published: (2024)
VVRec: Reconstruction Attacks on DL-based Volumetric Video Upstreaming via Latent Diffusion Model with Gamma Distribution
by: Lu, Rui, et al.
Published: (2025)
by: Lu, Rui, et al.
Published: (2025)
Universally Unfiltered and Unseen:Input-Agnostic Multimodal Jailbreaks against Text-to-Image Model Safeguards
by: Yan, Song, et al.
Published: (2025)
by: Yan, Song, et al.
Published: (2025)
Natural Language Induced Adversarial Images
by: Zhu, Xiaopei, et al.
Published: (2024)
by: Zhu, Xiaopei, et al.
Published: (2024)
Test-Time Backdoor Attacks on Multimodal Large Language Models
by: Lu, Dong, et al.
Published: (2024)
by: Lu, Dong, et al.
Published: (2024)
Privis: Towards Content-Aware Secure Volumetric Video Delivery
by: Hu, Kaiyuan, et al.
Published: (2025)
by: Hu, Kaiyuan, et al.
Published: (2025)
Improving Adversarial Transferability of Vision-Language Pre-training Models through Collaborative Multimodal Interaction
by: Fu, Jiyuan, et al.
Published: (2024)
by: Fu, Jiyuan, et al.
Published: (2024)
Is It Really You? Exploring Biometric Verification Scenarios in Photorealistic Talking-Head Avatar Videos
by: Pedrouzo-Rodriguez, Laura, et al.
Published: (2025)
by: Pedrouzo-Rodriguez, Laura, et al.
Published: (2025)
A Multi-task Adversarial Attack Against Face Authentication
by: Wang, Hanrui, et al.
Published: (2024)
by: Wang, Hanrui, et al.
Published: (2024)
BadCM: Invisible Backdoor Attack Against Cross-Modal Learning
by: Zhang, Zheng, et al.
Published: (2024)
by: Zhang, Zheng, et al.
Published: (2024)
DeepfakeBench-MM: A Comprehensive Benchmark for Multimodal Deepfake Detection
by: Zhao, Kangran, et al.
Published: (2025)
by: Zhao, Kangran, et al.
Published: (2025)
Blind Deep-Learning-Based Image Watermarking Robust Against Geometric Transformations
by: Mareen, Hannes, et al.
Published: (2024)
by: Mareen, Hannes, et al.
Published: (2024)
Provably Secure Robust Image Steganography via Cross-Modal Error Correction
by: Qi, Yuang, et al.
Published: (2024)
by: Qi, Yuang, et al.
Published: (2024)
From Attack to Protection: Leveraging Watermarking Attack Network for Advanced Add-on Watermarking
by: Nam, Seung-Hun, et al.
Published: (2020)
by: Nam, Seung-Hun, et al.
Published: (2020)
CoreMark: Toward Robust and Universal Text Watermarking Technique
by: Meng, Jiale, et al.
Published: (2025)
by: Meng, Jiale, et al.
Published: (2025)
Wallcamera: Reinventing the Wheel?
by: Bourquard, Aurélien, et al.
Published: (2024)
by: Bourquard, Aurélien, et al.
Published: (2024)
ByteNet: Rethinking Multimedia File Fragment Classification through Visual Perspectives
by: Liu, Wenyang, et al.
Published: (2024)
by: Liu, Wenyang, et al.
Published: (2024)
Reversible Video Steganography Using Quick Response Codes and Modified ElGamal Cryptosystem
by: Mstafa, Ramadhan J.
Published: (2025)
by: Mstafa, Ramadhan J.
Published: (2025)
Poisoning Prompt-Guided Sampling in Video Large Language Models
by: Cao, Yuxin, et al.
Published: (2025)
by: Cao, Yuxin, et al.
Published: (2025)
DKiS: Decay weight invertible image steganography with private key
by: Yang, Hang, et al.
Published: (2023)
by: Yang, Hang, et al.
Published: (2023)
VA3: Virtually Assured Amplification Attack on Probabilistic Copyright Protection for Text-to-Image Generative Models
by: Li, Xiang, et al.
Published: (2023)
by: Li, Xiang, et al.
Published: (2023)
qAttCNN - Self Attention Mechanism for Video QoE Prediction in Encrypted Traffic
by: Sidorov, Michael, et al.
Published: (2026)
by: Sidorov, Michael, et al.
Published: (2026)
Benchmarking Large Multimodal Models against Common Corruptions
by: Zhang, Jiawei, et al.
Published: (2024)
by: Zhang, Jiawei, et al.
Published: (2024)
Cross-Layer Encrypted Semantic Communication Framework for Panoramic Video Transmission
by: Gao, Haixiao, et al.
Published: (2024)
by: Gao, Haixiao, et al.
Published: (2024)
LogoStyleFool: Vitiating Video Recognition Systems via Logo Style Transfer
by: Cao, Yuxin, et al.
Published: (2023)
by: Cao, Yuxin, et al.
Published: (2023)
Structured Visual Narratives Undermine Safety Alignment in Multimodal Large Language Models
by: Tan, Rui Yang, et al.
Published: (2026)
by: Tan, Rui Yang, et al.
Published: (2026)
Cert-LAS: Toward Certified Model Ownership Verification for Text-to-Image Diffusion Models via Layer-Adaptive Smoothing
by: Qi, Leyi, et al.
Published: (2026)
by: Qi, Leyi, et al.
Published: (2026)
TGIF2: Extended Text-Guided Inpainting Forgery Dataset & Benchmark
by: Mareen, Hannes, et al.
Published: (2026)
by: Mareen, Hannes, et al.
Published: (2026)
TGIF: Text-Guided Inpainting Forgery Dataset
by: Mareen, Hannes, et al.
Published: (2024)
by: Mareen, Hannes, et al.
Published: (2024)
SWIFT: Semantic Watermarking for Image Forgery Thwarting
by: Evennou, Gautier, et al.
Published: (2024)
by: Evennou, Gautier, et al.
Published: (2024)
Beyond Text: Multimodal Jailbreaking of Vision-Language and Audio Models through Perceptually Simple Transformations
by: Kumar, Divyanshu, et al.
Published: (2025)
by: Kumar, Divyanshu, et al.
Published: (2025)
Towards Effective User Attribution for Latent Diffusion Models via Watermark-Informed Blending
by: Pan, Yongyang, et al.
Published: (2024)
by: Pan, Yongyang, et al.
Published: (2024)
SemCovert: Secure and Covert Video Transmission via Deep Semantic-Level Hiding
by: Cao, Zhihan, et al.
Published: (2025)
by: Cao, Zhihan, et al.
Published: (2025)
StyleFool: Fooling Video Classification Systems via Style Transfer
by: Cao, Yuxin, et al.
Published: (2022)
by: Cao, Yuxin, et al.
Published: (2022)
Multimodal Unlearnable Examples: Protecting Data against Multimodal Contrastive Learning
by: Liu, Xinwei, et al.
Published: (2024)
by: Liu, Xinwei, et al.
Published: (2024)
Security Analysis of Thumbnail-Preserving Image Encryption and a New Framework
by: Xie, Dong, et al.
Published: (2025)
by: Xie, Dong, et al.
Published: (2025)
SilhouetteTell: Practical Video Identification Leveraging Blurred Recordings of Video Subtitles
by: Huang, Guanchong, et al.
Published: (2025)
by: Huang, Guanchong, et al.
Published: (2025)
PRISM-XR: Empowering Privacy-Aware XR Collaboration with Multimodal Large Language Models
by: Chen, Jiangong, et al.
Published: (2026)
by: Chen, Jiangong, et al.
Published: (2026)
SEA: Low-Resource Safety Alignment for Multimodal Large Language Models via Synthetic Embeddings
by: Lu, Weikai, et al.
Published: (2025)
by: Lu, Weikai, et al.
Published: (2025)
Similar Items
-
Failures to Surface Harmful Contents in Video Large Language Models
by: Cao, Yuxin, et al.
Published: (2025) -
Video Watermarking: Safeguarding Your Video from (Unauthorized) Annotations by Video-based LLMs
by: Li, Jinmin, et al.
Published: (2024) -
VVRec: Reconstruction Attacks on DL-based Volumetric Video Upstreaming via Latent Diffusion Model with Gamma Distribution
by: Lu, Rui, et al.
Published: (2025) -
Universally Unfiltered and Unseen:Input-Agnostic Multimodal Jailbreaks against Text-to-Image Model Safeguards
by: Yan, Song, et al.
Published: (2025) -
Natural Language Induced Adversarial Images
by: Zhu, Xiaopei, et al.
Published: (2024)