AD-MIR: Bridging the Gap from Perception to Persuasion in Advertising Video Understanding via Structured Reasoning
Fuente:
arXiv
Guardado en:
| Autores principales: | Xu, Binxiao, Feng, Junyu, Lin, Xiaopeng, Li, Haodong, Feng, Zhiyuan, Zeng, Bohan, Lu, Shaolin, Lu, Ming, She, Qi, Zhang, Wentao |
|---|---|
| Formato: | Preprint |
| Publicado: |
2026
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Jarvis: Towards Personalized AI Assistant via Personal KV-Cache Retrieval
por: Xu, Binxiao, et al.
Publicado: (2025)
por: Xu, Binxiao, et al.
Publicado: (2025)
M2A: Multimodal Memory Agent with Dual-Layer Hybrid Memory for Long-Term Personalized Interactions
por: Feng, Junyu, et al.
Publicado: (2026)
por: Feng, Junyu, et al.
Publicado: (2026)
Small-Large Collaboration: Training-efficient Concept Personalization for Large VLM using a Meta Personalized Small VLM
por: Yang, Sihan, et al.
Publicado: (2025)
por: Yang, Sihan, et al.
Publicado: (2025)
Bridging the Perception Gap in Image Super-Resolution Evaluation
por: Su, Shaolin, et al.
Publicado: (2025)
por: Su, Shaolin, et al.
Publicado: (2025)
TimeSearch: Hierarchical Video Search with Spotlight and Reflection for Human-like Long Video Understanding
por: Pan, Junwen, et al.
Publicado: (2025)
por: Pan, Junwen, et al.
Publicado: (2025)
Uni-Synergy: Bridging Understanding and Generation for Personalized Reasoning via Co-operative Reinforcement Learning
por: Shen, Zijun, et al.
Publicado: (2026)
por: Shen, Zijun, et al.
Publicado: (2026)
VCU-Bridge: Hierarchical Visual Connotation Understanding via Semantic Bridging
por: Zhong, Ming, et al.
Publicado: (2025)
por: Zhong, Ming, et al.
Publicado: (2025)
Decoupling Perception from Reasoning for Hallucination-Resistant Video Understanding
por: Pu, Bowei, et al.
Publicado: (2025)
por: Pu, Bowei, et al.
Publicado: (2025)
TraceAV-Bench: Benchmarking Multi-Hop Trajectory Reasoning over Long Audio-Visual Videos
por: Feng, Hengyi, et al.
Publicado: (2026)
por: Feng, Hengyi, et al.
Publicado: (2026)
Synthetic Video Enhances Physical Fidelity in Video Synthesis
por: Zhao, Qi, et al.
Publicado: (2025)
por: Zhao, Qi, et al.
Publicado: (2025)
Streaming Video Diffusion: Online Video Editing with Diffusion Models
por: Chen, Feng, et al.
Publicado: (2024)
por: Chen, Feng, et al.
Publicado: (2024)
Active Perception Agent for Omnimodal Audio-Video Understanding
por: Tao, Keda, et al.
Publicado: (2025)
por: Tao, Keda, et al.
Publicado: (2025)
OmniAD: Detect and Understand Industrial Anomaly via Multimodal Reasoning
por: Zhao, Shifang, et al.
Publicado: (2025)
por: Zhao, Shifang, et al.
Publicado: (2025)
The Persuasion Machine: Void Architecture in Advertising and Ad-Tech
por: Eckert, Anthony
Publicado: (2026)
por: Eckert, Anthony
Publicado: (2026)
PEARL: Personalized Streaming Video Understanding Model
por: Zheng, Yuanhong, et al.
Publicado: (2026)
por: Zheng, Yuanhong, et al.
Publicado: (2026)
GeoPQA: Bridging the Visual Perception Gap in MLLMs for Geometric Reasoning
por: Chen, Guizhen, et al.
Publicado: (2025)
por: Chen, Guizhen, et al.
Publicado: (2025)
Persuasive Calibration
por: Feng, Yiding, et al.
Publicado: (2025)
por: Feng, Yiding, et al.
Publicado: (2025)
Understand, Solve and Translate: Bridging the Multilingual Mathematical Reasoning Gap
por: Ko, Hyunwoo, et al.
Publicado: (2025)
por: Ko, Hyunwoo, et al.
Publicado: (2025)
TimeSearch-R: Adaptive Temporal Search for Long-Form Video Understanding via Self-Verification Reinforcement Learning
por: Pan, Junwen, et al.
Publicado: (2025)
por: Pan, Junwen, et al.
Publicado: (2025)
AdsQA: Towards Advertisement Video Understanding
por: Long, Xinwei, et al.
Publicado: (2025)
por: Long, Xinwei, et al.
Publicado: (2025)
Reasoning to Align: Implicit Reasoning in Diffusion Transformers for Video Editing
por: Li, Yan, et al.
Publicado: (2026)
por: Li, Yan, et al.
Publicado: (2026)
Bridging the Embodiment Gap: Disentangled Cross-Embodiment Video Editing
por: Li, Zhiyuan, et al.
Publicado: (2026)
por: Li, Zhiyuan, et al.
Publicado: (2026)
Bridging the Transparency Gap: Exploring Multi-Stakeholder Preferences for Targeted Advertisement Explanations
por: Zilbershtein, Dina, et al.
Publicado: (2024)
por: Zilbershtein, Dina, et al.
Publicado: (2024)
VCSearch: Bridging the Gap Between Well-Defined and Ill-Defined Problems in Mathematical Reasoning
por: Tian, Shi-Yu, et al.
Publicado: (2024)
por: Tian, Shi-Yu, et al.
Publicado: (2024)
The Persuasion Paradox: When LLM Explanations Fail to Improve Human-AI Team Performance
por: Cohen, Ruth, et al.
Publicado: (2026)
por: Cohen, Ruth, et al.
Publicado: (2026)
Scone: Bridging Composition and Distinction in Subject-Driven Image Generation via Unified Understanding-Generation Modeling
por: Wang, Yuran, et al.
Publicado: (2025)
por: Wang, Yuran, et al.
Publicado: (2025)
Bridging the Gap between Human Motion and Action Semantics via Kinematic Phrases
por: Liu, Xinpeng, et al.
Publicado: (2023)
por: Liu, Xinpeng, et al.
Publicado: (2023)
CogFlow: Bridging Perception and Reasoning through Knowledge Internalization for Visual Mathematical Problem Solving
por: Chen, Shuhang, et al.
Publicado: (2026)
por: Chen, Shuhang, et al.
Publicado: (2026)
A Hybrid Class‐B/C Mode‐Switching VCO With 80% Current Efficiency and 202.2 dBc/Hz FoMT
por: Yue Yin, et al.
Publicado: (2025)
por: Yue Yin, et al.
Publicado: (2025)
Blockchain-Based Ad Auctions and Bayesian Persuasion: An Analysis of Advertiser Behavior
por: Li, Xinyu
Publicado: (2024)
por: Li, Xinyu
Publicado: (2024)
MedSynapse-V: Bridging Visual Perception and Clinical Intuition via Latent Memory Evolution
por: Zhu, Chunzheng, et al.
Publicado: (2026)
por: Zhu, Chunzheng, et al.
Publicado: (2026)
Perception, Understanding and Reasoning, A Multimodal Benchmark for Video Fake News Detection
por: Yakun, Cui, et al.
Publicado: (2025)
por: Yakun, Cui, et al.
Publicado: (2025)
VideoPro: Adaptive Program Reasoning for Long Video Understanding
por: Li, Chenglin, et al.
Publicado: (2025)
por: Li, Chenglin, et al.
Publicado: (2025)
VidBridge-R1: Bridging QA and Captioning for RL-based Video Understanding Models with Intermediate Proxy Tasks
por: Chen, Xinlong, et al.
Publicado: (2025)
por: Chen, Xinlong, et al.
Publicado: (2025)
FineQuest: Adaptive Knowledge-Assisted Sports Video Understanding via Agent-of-Thoughts Reasoning
por: Chen, Haodong, et al.
Publicado: (2025)
por: Chen, Haodong, et al.
Publicado: (2025)
In the Eye of MLLM: Benchmarking Egocentric Video Intent Understanding with Gaze-Guided Prompting
por: Peng, Taiying, et al.
Publicado: (2025)
por: Peng, Taiying, et al.
Publicado: (2025)
Bridging Perception and Reasoning: Token Reweighting for RLVR in Multimodal LLMs
por: Lu, Jinda, et al.
Publicado: (2026)
por: Lu, Jinda, et al.
Publicado: (2026)
Beyond the Jingle: A Thematic Analysis of Creative Persuasion in Nigerian Telecommunications Advertising
por: Augustine Anyasi C., Akaeze
Publicado: (2025)
por: Augustine Anyasi C., Akaeze
Publicado: (2025)
Playful but Persuasive: Deceptive Designs and Advertising Strategies in Popular Mobile Apps for Children
por: Krahl, Hannah, et al.
Publicado: (2025)
por: Krahl, Hannah, et al.
Publicado: (2025)
Personality Modeling for Persuasion of Misinformation using AI Agent
por: Lou, Qianmin, et al.
Publicado: (2025)
por: Lou, Qianmin, et al.
Publicado: (2025)
Ejemplares similares
-
Jarvis: Towards Personalized AI Assistant via Personal KV-Cache Retrieval
por: Xu, Binxiao, et al.
Publicado: (2025) -
M2A: Multimodal Memory Agent with Dual-Layer Hybrid Memory for Long-Term Personalized Interactions
por: Feng, Junyu, et al.
Publicado: (2026) -
Small-Large Collaboration: Training-efficient Concept Personalization for Large VLM using a Meta Personalized Small VLM
por: Yang, Sihan, et al.
Publicado: (2025) -
Bridging the Perception Gap in Image Super-Resolution Evaluation
por: Su, Shaolin, et al.
Publicado: (2025) -
TimeSearch: Hierarchical Video Search with Spotlight and Reflection for Human-like Long Video Understanding
por: Pan, Junwen, et al.
Publicado: (2025)