HPNet: Dynamic Trajectory Forecasting with Historical Prediction Attention
Fuente:
arXiv
Salvato in:
| Autori principali: | Tang, Xiaolong, Kan, Meina, Shan, Shiguang, Ji, Zhilong, Bai, Jinfeng, Chen, Xilin |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Neural Gate: Mitigating Privacy Risks in LVLMs via Neuron-Level Gradient Gating
di: Cao, Xiangkui, et al.
Pubblicazione: (2026)
di: Cao, Xiangkui, et al.
Pubblicazione: (2026)
Jodi: Unification of Visual Generation and Understanding via Joint Modeling
di: Xu, Yifeng, et al.
Pubblicazione: (2025)
di: Xu, Yifeng, et al.
Pubblicazione: (2025)
Dynamic Attention Analysis for Backdoor Detection in Text-to-Image Diffusion Models
di: Wang, Zhongqi, et al.
Pubblicazione: (2025)
di: Wang, Zhongqi, et al.
Pubblicazione: (2025)
OSI: One-step Inversion Excels in Extracting Diffusion Watermarks
di: Chen, Yuwei, et al.
Pubblicazione: (2026)
di: Chen, Yuwei, et al.
Pubblicazione: (2026)
Towards Robust Semantic Segmentation against Patch-based Attack via Attention Refinement
di: Yuan, Zheng, et al.
Pubblicazione: (2024)
di: Yuan, Zheng, et al.
Pubblicazione: (2024)
EntropyScan: Towards Model-level Backdoor Detection in LVLMs via Visual Attention Entropy
di: Ge, Xuanyu, et al.
Pubblicazione: (2026)
di: Ge, Xuanyu, et al.
Pubblicazione: (2026)
JoPano: Unified Panorama Generation via Joint Modeling
di: Feng, Wancheng, et al.
Pubblicazione: (2025)
di: Feng, Wancheng, et al.
Pubblicazione: (2025)
Dual Attention Guided Defense Against Malicious Edits
di: Zhang, Jie, et al.
Pubblicazione: (2025)
di: Zhang, Jie, et al.
Pubblicazione: (2025)
CtrLoRA: An Extensible and Efficient Framework for Controllable Image Generation
di: Xu, Yifeng, et al.
Pubblicazione: (2024)
di: Xu, Yifeng, et al.
Pubblicazione: (2024)
Semantic or Covariate? A Study on the Intractable Case of Out-of-Distribution Detection
di: Long, Xingming, et al.
Pubblicazione: (2024)
di: Long, Xingming, et al.
Pubblicazione: (2024)
T2IShield: Defending Against Backdoors on Text-to-Image Diffusion Models
di: Wang, Zhongqi, et al.
Pubblicazione: (2024)
di: Wang, Zhongqi, et al.
Pubblicazione: (2024)
FullLoRA: Efficiently Boosting the Robustness of Pretrained Vision Transformers
di: Yuan, Zheng, et al.
Pubblicazione: (2024)
di: Yuan, Zheng, et al.
Pubblicazione: (2024)
Rethinking the Evaluation of Out-of-Distribution Detection: A Sorites Paradox
di: Long, Xingming, et al.
Pubblicazione: (2024)
di: Long, Xingming, et al.
Pubblicazione: (2024)
GLip: A Global-Local Integrated Progressive Framework for Robust Visual Speech Recognition
di: Wang, Tianyue, et al.
Pubblicazione: (2025)
di: Wang, Tianyue, et al.
Pubblicazione: (2025)
VOPE: Revisiting Hallucination of Vision-Language Models in Voluntary Imagination Task
di: Long, Xingming, et al.
Pubblicazione: (2025)
di: Long, Xingming, et al.
Pubblicazione: (2025)
Assimilation Matters: Model-level Backdoor Detection in Vision-Language Pretrained Models
di: Wang, Zhongqi, et al.
Pubblicazione: (2025)
di: Wang, Zhongqi, et al.
Pubblicazione: (2025)
Trigger without Trace: Towards Stealthy Backdoor Attack on Text-to-Image Diffusion Models
di: Zhang, Jie, et al.
Pubblicazione: (2025)
di: Zhang, Jie, et al.
Pubblicazione: (2025)
Task-adaptive Q-Face
di: Sun, Haomiao, et al.
Pubblicazione: (2024)
di: Sun, Haomiao, et al.
Pubblicazione: (2024)
Contrastive Spectral Rectification: Test-Time Defense towards Zero-shot Adversarial Robustness of CLIP
di: Nie, Sen, et al.
Pubblicazione: (2026)
di: Nie, Sen, et al.
Pubblicazione: (2026)
Component-Based Out-of-Distribution Detection
di: Liu, Wenrui, et al.
Pubblicazione: (2026)
di: Liu, Wenrui, et al.
Pubblicazione: (2026)
EfficientMT: Efficient Temporal Adaptation for Motion Transfer in Text-to-Video Diffusion Models
di: Cai, Yufei, et al.
Pubblicazione: (2025)
di: Cai, Yufei, et al.
Pubblicazione: (2025)
V-Attack: Targeting Disentangled Value Features for Controllable Adversarial Attacks on LVLMs
di: Nie, Sen, et al.
Pubblicazione: (2025)
di: Nie, Sen, et al.
Pubblicazione: (2025)
ACT Now: Preempting LVLM Hallucinations via Adaptive Context Integration
di: Yan, Bei, et al.
Pubblicazione: (2026)
di: Yan, Bei, et al.
Pubblicazione: (2026)
Dysca: A Dynamic and Scalable Benchmark for Evaluating Perception Ability of LVLMs
di: Zhang, Jie, et al.
Pubblicazione: (2024)
di: Zhang, Jie, et al.
Pubblicazione: (2024)
DIVE: Inverting Conditional Diffusion Models for Discriminative Tasks
di: Li, Yinqi, et al.
Pubblicazione: (2025)
di: Li, Yinqi, et al.
Pubblicazione: (2025)
MM-MoralBench: A MultiModal Moral Evaluation Benchmark for Large Vision-Language Models
di: Yan, Bei, et al.
Pubblicazione: (2024)
di: Yan, Bei, et al.
Pubblicazione: (2024)
UMFC: Unsupervised Multi-Domain Feature Calibration for Vision-Language Models
di: Liang, Jiachen, et al.
Pubblicazione: (2024)
di: Liang, Jiachen, et al.
Pubblicazione: (2024)
un$^2$CLIP: Improving CLIP's Visual Detail Capturing Ability via Inverting unCLIP
di: Li, Yinqi, et al.
Pubblicazione: (2025)
di: Li, Yinqi, et al.
Pubblicazione: (2025)
Revisiting Logit Distributions for Reliable Out-of-Distribution Detection
di: Liang, Jiachen, et al.
Pubblicazione: (2025)
di: Liang, Jiachen, et al.
Pubblicazione: (2025)
What Makes VLMs Robust? Towards Reconciling Robustness and Accuracy in Vision-Language Models
di: Nie, Sen, et al.
Pubblicazione: (2026)
di: Nie, Sen, et al.
Pubblicazione: (2026)
T2VAttack: Adversarial Attack on Text-to-Video Diffusion Models
di: Li, Changzhen, et al.
Pubblicazione: (2025)
di: Li, Changzhen, et al.
Pubblicazione: (2025)
Towards Transferable Defense Against Malicious Image Edits
di: Zhang, Jie, et al.
Pubblicazione: (2025)
di: Zhang, Jie, et al.
Pubblicazione: (2025)
Measuring the Measurers: Quality Evaluation of Hallucination Benchmarks for Large Vision-Language Models
di: Yan, Bei, et al.
Pubblicazione: (2024)
di: Yan, Bei, et al.
Pubblicazione: (2024)
UniPose: A Unified Multimodal Framework for Human Pose Comprehension, Generation and Editing
di: Li, Yiheng, et al.
Pubblicazione: (2024)
di: Li, Yiheng, et al.
Pubblicazione: (2024)
INFACT: A Diagnostic Benchmark for Induced Faithfulness and Factuality Hallucinations in Video-LLMs
di: Yang, Junqi, et al.
Pubblicazione: (2026)
di: Yang, Junqi, et al.
Pubblicazione: (2026)
MasterWeaver: Taming Editability and Face Identity for Personalized Text-to-Image Generation
di: Wei, Yuxiang, et al.
Pubblicazione: (2024)
di: Wei, Yuxiang, et al.
Pubblicazione: (2024)
Explicit Relational Reasoning Network for Scene Text Detection
di: Su, Yuchen, et al.
Pubblicazione: (2024)
di: Su, Yuchen, et al.
Pubblicazione: (2024)
Clothes-Changing Person Re-Identification with Feasibility-Aware Intermediary Matching
di: Zhao, Jiahe, et al.
Pubblicazione: (2024)
di: Zhao, Jiahe, et al.
Pubblicazione: (2024)
Learning Separable Hidden Unit Contributions for Speaker-Adaptive Lip-Reading
di: Luo, Songtao, et al.
Pubblicazione: (2023)
di: Luo, Songtao, et al.
Pubblicazione: (2023)
Contrastive Learning of Person-independent Representations for Facial Action Unit Detection
di: Li, Yong, et al.
Pubblicazione: (2024)
di: Li, Yong, et al.
Pubblicazione: (2024)
Documenti analoghi
-
Neural Gate: Mitigating Privacy Risks in LVLMs via Neuron-Level Gradient Gating
di: Cao, Xiangkui, et al.
Pubblicazione: (2026) -
Jodi: Unification of Visual Generation and Understanding via Joint Modeling
di: Xu, Yifeng, et al.
Pubblicazione: (2025) -
Dynamic Attention Analysis for Backdoor Detection in Text-to-Image Diffusion Models
di: Wang, Zhongqi, et al.
Pubblicazione: (2025) -
OSI: One-step Inversion Excels in Extracting Diffusion Watermarks
di: Chen, Yuwei, et al.
Pubblicazione: (2026) -
Towards Robust Semantic Segmentation against Patch-based Attack via Attention Refinement
di: Yuan, Zheng, et al.
Pubblicazione: (2024)