MoCLIP: Motion-Aware Fine-Tuning and Distillation of CLIP for Human Motion Generation
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Maldonado, Gabriel, Pazho, Armin Danesh, Noghre, Ghazal Alinezhad, Katariya, Vinit, Tabkhi, Hamed |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
VT-Former: An Exploratory Study on Vehicle Trajectory Prediction for Highway Surveillance through Graph Isomorphism and Transformer
von: Pazho, Armin Danesh, et al.
Veröffentlicht: (2023)
von: Pazho, Armin Danesh, et al.
Veröffentlicht: (2023)
Adversarially-Refined VQ-GAN with Dense Motion Tokenization for Spatio-Temporal Heatmaps
von: Maldonado, Gabriel, et al.
Veröffentlicht: (2025)
von: Maldonado, Gabriel, et al.
Veröffentlicht: (2025)
MoFM: A Large-Scale Human Motion Foundation Model
von: Baharani, Mohammadreza, et al.
Veröffentlicht: (2025)
von: Baharani, Mohammadreza, et al.
Veröffentlicht: (2025)
Human-Centric Video Anomaly Detection Through Spatio-Temporal Pose Tokenization and Transformer
von: Noghre, Ghazal Alinezhad, et al.
Veröffentlicht: (2024)
von: Noghre, Ghazal Alinezhad, et al.
Veröffentlicht: (2024)
A Survey on Video Anomaly Detection via Deep Learning: Human, Vehicle, and Environment
von: Noghre, Ghazal Alinezhad, et al.
Veröffentlicht: (2025)
von: Noghre, Ghazal Alinezhad, et al.
Veröffentlicht: (2025)
An Exploratory Study on Human-Centric Video Anomaly Detection through Variational Autoencoders and Trajectory Prediction
von: Noghre, Ghazal Alinezhad, et al.
Veröffentlicht: (2024)
von: Noghre, Ghazal Alinezhad, et al.
Veröffentlicht: (2024)
Towards Adaptive Human-centric Video Anomaly Detection: A Comprehensive Framework and A New Benchmark
von: Pazho, Armin Danesh, et al.
Veröffentlicht: (2024)
von: Pazho, Armin Danesh, et al.
Veröffentlicht: (2024)
ALFred: An Active Learning Framework for Real-world Semi-supervised Anomaly Detection with Adaptive Thresholds
von: Yao, Shanle, et al.
Veröffentlicht: (2025)
von: Yao, Shanle, et al.
Veröffentlicht: (2025)
Evaluating the Effectiveness of Video Anomaly Detection in the Wild: Online Learning and Inference for Real-world Deployment
von: Yao, Shanle, et al.
Veröffentlicht: (2024)
von: Yao, Shanle, et al.
Veröffentlicht: (2024)
Shopformer: Transformer-Based Framework for Detecting Shoplifting via Human Pose
von: Rashvand, Narges, et al.
Veröffentlicht: (2025)
von: Rashvand, Narges, et al.
Veröffentlicht: (2025)
Exploring Pose-Based Anomaly Detection for Retail Security: A Real-World Shoplifting Dataset and Benchmark
von: Rashvand, Narges, et al.
Veröffentlicht: (2025)
von: Rashvand, Narges, et al.
Veröffentlicht: (2025)
MoCLIP-Lite: Efficient Video Recognition by Fusing CLIP with Motion Vectors
von: Huang, Binhua, et al.
Veröffentlicht: (2025)
von: Huang, Binhua, et al.
Veröffentlicht: (2025)
From Lab to Field: Real-World Evaluation of an AI-Driven Smart Video Solution to Enhance Community Safety
von: Yao, Shanle, et al.
Veröffentlicht: (2023)
von: Yao, Shanle, et al.
Veröffentlicht: (2023)
Are Multimodal LLMs Ready for Surveillance? A Reality Check on Zero-Shot Anomaly Detection in the Wild
von: Yao, Shanle, et al.
Veröffentlicht: (2026)
von: Yao, Shanle, et al.
Veröffentlicht: (2026)
From Frames to Events: Rethinking Evaluation in Human-Centric Video Anomaly Detection
von: Rashvand, Narges, et al.
Veröffentlicht: (2026)
von: Rashvand, Narges, et al.
Veröffentlicht: (2026)
EdgeVTP: Exploration of Latency-efficient Trajectory Prediction for Edge-based Embedded Vision Applications
von: Kim, Seungjin, et al.
Veröffentlicht: (2026)
von: Kim, Seungjin, et al.
Veröffentlicht: (2026)
Intelligent CCTV for Urban Design: AI-Based Analysis of Soft Infrastructure at Intersections
von: Katariya, Vinit, et al.
Veröffentlicht: (2026)
von: Katariya, Vinit, et al.
Veröffentlicht: (2026)
AnimalMotionCLIP: Embedding motion in CLIP for Animal Behavior Analysis
von: Zhong, Enmin, et al.
Veröffentlicht: (2025)
von: Zhong, Enmin, et al.
Veröffentlicht: (2025)
Advancing Compositional Awareness in CLIP with Efficient Fine-Tuning
von: Peleg, Amit, et al.
Veröffentlicht: (2025)
von: Peleg, Amit, et al.
Veröffentlicht: (2025)
CLIP-RD: Relative Distillation for Efficient CLIP Knowledge Distillation
von: Chung, Jeannie, et al.
Veröffentlicht: (2026)
von: Chung, Jeannie, et al.
Veröffentlicht: (2026)
DGS-Net: Distillation-Guided Gradient Surgery for CLIP Fine-Tuning in AI-Generated Image Detection
von: Yan, Jiazhen, et al.
Veröffentlicht: (2025)
von: Yan, Jiazhen, et al.
Veröffentlicht: (2025)
CLIP-KD: An Empirical Study of CLIP Model Distillation
von: Yang, Chuanguang, et al.
Veröffentlicht: (2023)
von: Yang, Chuanguang, et al.
Veröffentlicht: (2023)
MoP-CLIP: A Mixture of Prompt-Tuned CLIP Models for Domain Incremental Learning
von: Nicolas, Julien, et al.
Veröffentlicht: (2023)
von: Nicolas, Julien, et al.
Veröffentlicht: (2023)
DetailCLIP: Detail-Oriented CLIP for Fine-Grained Tasks
von: Monsefi, Amin Karimi, et al.
Veröffentlicht: (2024)
von: Monsefi, Amin Karimi, et al.
Veröffentlicht: (2024)
CLIP-CID: Efficient CLIP Distillation via Cluster-Instance Discrimination
von: Yang, Kaicheng, et al.
Veröffentlicht: (2024)
von: Yang, Kaicheng, et al.
Veröffentlicht: (2024)
Contrast-Aware Calibration for Fine-Tuned CLIP: Leveraging Image-Text Alignment
von: Lv, Song-Lin, et al.
Veröffentlicht: (2025)
von: Lv, Song-Lin, et al.
Veröffentlicht: (2025)
LatteCLIP: Unsupervised CLIP Fine-Tuning via LMM-Synthetic Texts
von: Cao, Anh-Quan, et al.
Veröffentlicht: (2024)
von: Cao, Anh-Quan, et al.
Veröffentlicht: (2024)
EasyTune: Efficient Step-Aware Fine-Tuning for Diffusion-Based Motion Generation
von: Tan, Xiaofeng, et al.
Veröffentlicht: (2026)
von: Tan, Xiaofeng, et al.
Veröffentlicht: (2026)
MotionRFT: Unified Reinforcement Fine-Tuning for Text-to-Motion Generation
von: Tan, Xiaofeng, et al.
Veröffentlicht: (2026)
von: Tan, Xiaofeng, et al.
Veröffentlicht: (2026)
Omni-NegCLIP: Enhancing CLIP with Front-Layer Contrastive Fine-Tuning for Comprehensive Negation Understanding
von: Xu, Jingqi
Veröffentlicht: (2026)
von: Xu, Jingqi
Veröffentlicht: (2026)
MoMaps: Semantics-Aware Scene Motion Generation with Motion Maps
von: Lei, Jiahui, et al.
Veröffentlicht: (2025)
von: Lei, Jiahui, et al.
Veröffentlicht: (2025)
MedP-CLIP: Medical CLIP with Region-Aware Prompt Integration
von: Peng, Jiahui, et al.
Veröffentlicht: (2026)
von: Peng, Jiahui, et al.
Veröffentlicht: (2026)
BadCLIP: Trigger-Aware Prompt Learning for Backdoor Attacks on CLIP
von: Bai, Jiawang, et al.
Veröffentlicht: (2023)
von: Bai, Jiawang, et al.
Veröffentlicht: (2023)
Robotic-CLIP: Fine-tuning CLIP on Action Data for Robotic Applications
von: Nguyen, Nghia, et al.
Veröffentlicht: (2024)
von: Nguyen, Nghia, et al.
Veröffentlicht: (2024)
TNG-CLIP:Training-Time Negation Data Generation for Negation Awareness of CLIP
von: Cai, Yuliang, et al.
Veröffentlicht: (2025)
von: Cai, Yuliang, et al.
Veröffentlicht: (2025)
MotionCharacter: Fine-Grained Motion Controllable Human Video Generation
von: Fang, Haopeng, et al.
Veröffentlicht: (2024)
von: Fang, Haopeng, et al.
Veröffentlicht: (2024)
Prompt Tuning for CLIP on the Pretrained Manifold
von: Yang, Xi, et al.
Veröffentlicht: (2026)
von: Yang, Xi, et al.
Veröffentlicht: (2026)
Enhancing CLIP with CLIP: Exploring Pseudolabeling for Limited-Label Prompt Tuning
von: Menghini, Cristina, et al.
Veröffentlicht: (2023)
von: Menghini, Cristina, et al.
Veröffentlicht: (2023)
CLIP-FTI: Fine-Grained Face Template Inversion via CLIP-Driven Attribute Conditioning
von: Dai, Longchen, et al.
Veröffentlicht: (2025)
von: Dai, Longchen, et al.
Veröffentlicht: (2025)
TSCLIP: Robust CLIP Fine-Tuning for Worldwide Cross-Regional Traffic Sign Recognition
von: Zhao, Guoyang, et al.
Veröffentlicht: (2024)
von: Zhao, Guoyang, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
VT-Former: An Exploratory Study on Vehicle Trajectory Prediction for Highway Surveillance through Graph Isomorphism and Transformer
von: Pazho, Armin Danesh, et al.
Veröffentlicht: (2023) -
Adversarially-Refined VQ-GAN with Dense Motion Tokenization for Spatio-Temporal Heatmaps
von: Maldonado, Gabriel, et al.
Veröffentlicht: (2025) -
MoFM: A Large-Scale Human Motion Foundation Model
von: Baharani, Mohammadreza, et al.
Veröffentlicht: (2025) -
Human-Centric Video Anomaly Detection Through Spatio-Temporal Pose Tokenization and Transformer
von: Noghre, Ghazal Alinezhad, et al.
Veröffentlicht: (2024) -
A Survey on Video Anomaly Detection via Deep Learning: Human, Vehicle, and Environment
von: Noghre, Ghazal Alinezhad, et al.
Veröffentlicht: (2025)