Pop-Up Distractions Reveal Bag-of-Events Behavior in Video Large Language Models
Fuente:
arXiv
Saved in:
| Main Authors: | Chew, Oscar, Honcharenko, Serhii, Chen, Qian-Hui, Lu, Patricia, Zaveri, Dishant, Doan, Khoa D., Huang, Kuan-Hao |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Is CLIP Cross-Eyed? Revealing and Mitigating Center Bias in the CLIP Family
by: Chew, Oscar, et al.
Published: (2026)
by: Chew, Oscar, et al.
Published: (2026)
A Two-stage Transformer Framework for Temporal Localization of Distracted Driver Behaviors
by: Doan, Gia-Bao, et al.
Published: (2026)
by: Doan, Gia-Bao, et al.
Published: (2026)
Remember This Event That Year? Assessing Temporal Information and Reasoning in Large Language Models
by: Beniwal, Himanshu, et al.
Published: (2024)
by: Beniwal, Himanshu, et al.
Published: (2024)
Vision-Language Models can Identify Distracted Driver Behavior from Naturalistic Videos
by: Hasan, Md Zahid, et al.
Published: (2023)
by: Hasan, Md Zahid, et al.
Published: (2023)
PEPPER: Perception-Guided Perturbation for Robust Backdoor Defense in Text-to-Image Diffusion Models
by: Chew, Oscar, et al.
Published: (2025)
by: Chew, Oscar, et al.
Published: (2025)
No More Distractions: an Adaptive Up-Sampling Algorithm to Reduce Data Artifacts
by: Chen, Han
Published: (2024)
by: Chen, Han
Published: (2024)
Understanding and Mitigating Spurious Correlations in Text Classification with Neighborhood Analysis
by: Chew, Oscar, et al.
Published: (2023)
by: Chew, Oscar, et al.
Published: (2023)
Biometrics and Behavior Analysis for Detecting Distractions in e-Learning
by: Becerra, Álvaro, et al.
Published: (2024)
by: Becerra, Álvaro, et al.
Published: (2024)
Dynamical analysis of a parameter-aware reservoir computer
by: Sisodia, Dishant, et al.
Published: (2024)
by: Sisodia, Dishant, et al.
Published: (2024)
Distract Large Language Models for Automatic Jailbreak Attack
by: Xiao, Zeguan, et al.
Published: (2024)
by: Xiao, Zeguan, et al.
Published: (2024)
Seeing but Not Thinking: Routing Distraction in Multimodal Mixture-of-Experts
by: Xu, Haolei, et al.
Published: (2026)
by: Xu, Haolei, et al.
Published: (2026)
Distraction is All You Need for Multimodal Large Language Model Jailbreaking
by: Yang, Zuopeng, et al.
Published: (2025)
by: Yang, Zuopeng, et al.
Published: (2025)
Snakes and Ladders: Two Steps Up for VideoMamba
by: Lu, Hui, et al.
Published: (2024)
by: Lu, Hui, et al.
Published: (2024)
Infrastructuring Pop-Up Cities with "Social Layer": Designing Serendipitous Co-Livings for Temporary Intentional Communities
by: Ji, Danwen, et al.
Published: (2025)
by: Ji, Danwen, et al.
Published: (2025)
Analysis of Distracted Pedestrians Crossing Behavior: An Immersive Virtual Reality Application
by: Sulle, Methusela, et al.
Published: (2025)
by: Sulle, Methusela, et al.
Published: (2025)
Class-Aware Contrastive Optimization for Imbalanced Text Classification
by: Khvatskii, Grigorii, et al.
Published: (2024)
by: Khvatskii, Grigorii, et al.
Published: (2024)
Scaling Up Video Summarization Pretraining with Large Language Models
by: Argaw, Dawit Mureja, et al.
Published: (2024)
by: Argaw, Dawit Mureja, et al.
Published: (2024)
DistractMIA: Black-Box Membership Inference on Vision-Language Models via Semantic Distraction
by: Tang, Hongyi, et al.
Published: (2026)
by: Tang, Hongyi, et al.
Published: (2026)
EventVL: Understand Event Streams via Multimodal Large Language Model
by: Li, Pengteng, et al.
Published: (2025)
by: Li, Pengteng, et al.
Published: (2025)
Reducing Distraction in Long-Context Language Models by Focused Learning
by: Wu, Zijun, et al.
Published: (2024)
by: Wu, Zijun, et al.
Published: (2024)
TRACE: A Metrologically-Grounded Engineering Framework for Trustworthy Agentic AI Systems in Operationally Critical Domains
by: Zabolotnii, Serhii
Published: (2026)
by: Zabolotnii, Serhii
Published: (2026)
LLM-SRBench: A New Benchmark for Scientific Equation Discovery with Large Language Models
by: Shojaee, Parshin, et al.
Published: (2025)
by: Shojaee, Parshin, et al.
Published: (2025)
Unveiling Concept Attribution in Diffusion Models
by: Nguyen, Quang H., et al.
Published: (2024)
by: Nguyen, Quang H., et al.
Published: (2024)
Steering Vector Fields for Context-Aware Inference-Time Control in Large Language Models
by: Li, Jiaqian, et al.
Published: (2026)
by: Li, Jiaqian, et al.
Published: (2026)
VERHallu: Evaluating and Mitigating Event Relation Hallucination in Video Large Language Models
by: Zhang, Zefan, et al.
Published: (2026)
by: Zhang, Zefan, et al.
Published: (2026)
LLMVA-GEBC: Large Language Model with Video Adapter for Generic Event Boundary Captioning
by: Tang, Yolo Yunlong, et al.
Published: (2023)
by: Tang, Yolo Yunlong, et al.
Published: (2023)
EstemPMM: Polynomial Maximization Method for Non-Gaussian Regression and Time Series in R
by: Zabolotnii, Serhii
Published: (2026)
by: Zabolotnii, Serhii
Published: (2026)
LLM StructCore: Schema-Guided Reasoning Condensation and Deterministic Compilation
by: Zabolotnii, Serhii
Published: (2026)
by: Zabolotnii, Serhii
Published: (2026)
The Role of Exploration Modules in Small Language Models for Knowledge Graph Question Answering
by: Cheng, Yi-Jie, et al.
Published: (2025)
by: Cheng, Yi-Jie, et al.
Published: (2025)
LyCon: Lyrics Reconstruction from the Bag-of-Words Using Large Language Models
by: Kim, Haven, et al.
Published: (2024)
by: Kim, Haven, et al.
Published: (2024)
Empowering Large Language Model for Continual Video Question Answering with Collaborative Prompting
by: Cai, Chen, et al.
Published: (2024)
by: Cai, Chen, et al.
Published: (2024)
SHA256 at SemEval-2025 Task 4: Selective Amnesia -- Constrained Unlearning for Large Language Models via Knowledge Isolation
by: Agrawal, Saransh, et al.
Published: (2025)
by: Agrawal, Saransh, et al.
Published: (2025)
A Novel Dataset for Video-Based Neurodivergent Classification Leveraging Extra-Stimulatory Behavior
by: Serna-Aguilera, Manuel, et al.
Published: (2024)
by: Serna-Aguilera, Manuel, et al.
Published: (2024)
Iterative LLM-Based Generation and Refinement of Distracting Conditions in Math Word Problems
by: Yang, Kaiqi, et al.
Published: (2025)
by: Yang, Kaiqi, et al.
Published: (2025)
FLAT: Latent-Driven Arbitrary-Target Backdoor Attacks in Federated Learning
by: Nguyen, Tuan, et al.
Published: (2025)
by: Nguyen, Tuan, et al.
Published: (2025)
Attacking Vision-Language Computer Agents via Pop-ups
by: Zhang, Yanzhe, et al.
Published: (2024)
by: Zhang, Yanzhe, et al.
Published: (2024)
Event Extraction in Large Language Model
by: Li, Bobo, et al.
Published: (2025)
by: Li, Bobo, et al.
Published: (2025)
EventSTU: Event-Guided Efficient Spatio-Temporal Understanding for Video Large Language Models
by: Xu, Wenhao, et al.
Published: (2025)
by: Xu, Wenhao, et al.
Published: (2025)
Venomancer: Towards Imperceptible and Target-on-Demand Backdoor Attacks in Federated Learning
by: Nguyen, Son, et al.
Published: (2024)
by: Nguyen, Son, et al.
Published: (2024)
Spiking-DD: Neuromorphic Event Camera based Driver Distraction Detection with Spiking Neural Network
by: Shariff, Waseem, et al.
Published: (2024)
by: Shariff, Waseem, et al.
Published: (2024)
Similar Items
-
Is CLIP Cross-Eyed? Revealing and Mitigating Center Bias in the CLIP Family
by: Chew, Oscar, et al.
Published: (2026) -
A Two-stage Transformer Framework for Temporal Localization of Distracted Driver Behaviors
by: Doan, Gia-Bao, et al.
Published: (2026) -
Remember This Event That Year? Assessing Temporal Information and Reasoning in Large Language Models
by: Beniwal, Himanshu, et al.
Published: (2024) -
Vision-Language Models can Identify Distracted Driver Behavior from Naturalistic Videos
by: Hasan, Md Zahid, et al.
Published: (2023) -
PEPPER: Perception-Guided Perturbation for Robust Backdoor Defense in Text-to-Image Diffusion Models
by: Chew, Oscar, et al.
Published: (2025)