Saved in:
| Main Authors: | Ahmar, Wassim El, Kolhatkar, Dhanvin, Nowruzi, Farzan, Laganiere, Robert |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | https://arxiv.org/abs/2411.12943 |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
FSMDet: Vision-guided feature diffusion for fully sparse 3D detector
by: Liu, Tianran, et al.
Published: (2024)
by: Liu, Tianran, et al.
Published: (2024)
MOT FCG++: Enhanced Representation of Spatio-temporal Motion and Appearance Features
by: Fang, Yanzhao
Published: (2024)
by: Fang, Yanzhao
Published: (2024)
ID-Sim: An Identity-Focused Similarity Metric
by: Chae, Julia, et al.
Published: (2026)
by: Chae, Julia, et al.
Published: (2026)
MAML MOT: Multiple Object Tracking based on Meta-Learning
by: Chen, Jiayi, et al.
Published: (2024)
by: Chen, Jiayi, et al.
Published: (2024)
Hyperbolic Space Learning Method Leveraging Temporal Motion Priors for Human Mesh Recovery
by: Zhang, Xiang, et al.
Published: (2025)
by: Zhang, Xiang, et al.
Published: (2025)
Leveraging Generative AI Models to Explore Human Identity
by: Yeo, Yunha, et al.
Published: (2025)
by: Yeo, Yunha, et al.
Published: (2025)
ThermalGaussian: Thermal 3D Gaussian Splatting
by: Lu, Rongfeng, et al.
Published: (2024)
by: Lu, Rongfeng, et al.
Published: (2024)
S3MOT: Monocular 3D Object Tracking with Selective State Space Model
by: Yan, Zhuohao, et al.
Published: (2025)
by: Yan, Zhuohao, et al.
Published: (2025)
Physics-Aware Diffusion for LiDAR Point Cloud Densification
by: Zhang, Zeping, et al.
Published: (2026)
by: Zhang, Zeping, et al.
Published: (2026)
X-UniMotion: Animating Human Images with Expressive, Unified and Identity-Agnostic Motion Latents
by: Song, Guoxian, et al.
Published: (2025)
by: Song, Guoxian, et al.
Published: (2025)
The Progression of Transformers from Language to Vision to MOT: A Literature Review on Multi-Object Tracking with Transformers
by: Kamboj, Abhi
Published: (2024)
by: Kamboj, Abhi
Published: (2024)
Domain Adaptive SAR Wake Detection: Leveraging Similarity Filtering and Memory Guidance
by: Gao, He, et al.
Published: (2025)
by: Gao, He, et al.
Published: (2025)
ReL-SAR: Representation Learning for Skeleton Action Recognition with Convolutional Transformers and BYOL
by: Naimi, Safwen, et al.
Published: (2024)
by: Naimi, Safwen, et al.
Published: (2024)
Thermal Imaging-based Real-time Fall Detection using Motion Flow and Attention-enhanced Convolutional Recurrent Architecture
by: Silver, Christopher, et al.
Published: (2025)
by: Silver, Christopher, et al.
Published: (2025)
Rethinking Plant Disease Diagnosis: Bridging the Academic-Practical Gap with Vision Transformers and Zero-Shot Learning
by: Benabbas, Wassim, et al.
Published: (2025)
by: Benabbas, Wassim, et al.
Published: (2025)
Eye Sclera for Fair Face Image Quality Assessment
by: Kabbani, Wassim, et al.
Published: (2025)
by: Kabbani, Wassim, et al.
Published: (2025)
StableMorph: High-Quality Face Morph Generation with Stable Diffusion
by: Kabbani, Wassim, et al.
Published: (2025)
by: Kabbani, Wassim, et al.
Published: (2025)
Hybrid Attention for Robust RGB-T Pedestrian Detection in Real-World Conditions
by: Rathinam, Arunkumar, et al.
Published: (2024)
by: Rathinam, Arunkumar, et al.
Published: (2024)
Case-Enhanced Vision Transformer: Improving Explanations of Image Similarity with a ViT-based Similarity Metric
by: Zhao, Ziwei, et al.
Published: (2024)
by: Zhao, Ziwei, et al.
Published: (2024)
DualReal: Adaptive Joint Training for Lossless Identity-Motion Fusion in Video Customization
by: Wang, Wenchuan, et al.
Published: (2025)
by: Wang, Wenchuan, et al.
Published: (2025)
DreamMotion: Space-Time Self-Similar Score Distillation for Zero-Shot Video Editing
by: Jeong, Hyeonho, et al.
Published: (2024)
by: Jeong, Hyeonho, et al.
Published: (2024)
Suicide Risk Assessment from AI-powered Video Surveillance: An Interpretable Framework for Prevention in Metro Stations
by: Naimi, Safwen, et al.
Published: (2026)
by: Naimi, Safwen, et al.
Published: (2026)
When Exploration Comes for Free with Mixture-Greedy: Do we need UCB in Diversity-Aware Multi-Armed Bandits?
by: Nia, Bahar Dibaei, et al.
Published: (2026)
by: Nia, Bahar Dibaei, et al.
Published: (2026)
ReWind: Understanding Long Videos with Instructed Learnable Memory
by: Diko, Anxhelo, et al.
Published: (2024)
by: Diko, Anxhelo, et al.
Published: (2024)
AniTalker: Animate Vivid and Diverse Talking Faces through Identity-Decoupled Facial Motion Encoding
by: Liu, Tao, et al.
Published: (2024)
by: Liu, Tao, et al.
Published: (2024)
Proto-OOD: Enhancing OOD Object Detection with Prototype Feature Similarity
by: Chen, Junkun, et al.
Published: (2024)
by: Chen, Junkun, et al.
Published: (2024)
Box-QAymo: Box-Referring VQA Dataset for Autonomous Driving
by: Etchegaray, Djamahl, et al.
Published: (2025)
by: Etchegaray, Djamahl, et al.
Published: (2025)
Doc-CoB: Enhancing Document Understanding with Visual Chain-of-Boxes Reasoning
by: Mo, Ye, et al.
Published: (2025)
by: Mo, Ye, et al.
Published: (2025)
LightMotion: A Light and Tuning-free Method for Simulating Camera Motion in Video Generation
by: Song, Quanjian, et al.
Published: (2025)
by: Song, Quanjian, et al.
Published: (2025)
PromptSplit: Revealing Prompt-Level Disagreement in Generative Models
by: Lotfian, Mehdi, et al.
Published: (2026)
by: Lotfian, Mehdi, et al.
Published: (2026)
Segment Any RGB-Thermal Model with Language-aided Distillation
by: Xing, Dong, et al.
Published: (2025)
by: Xing, Dong, et al.
Published: (2025)
AnyTSR: Any-Scale Thermal Super-Resolution for UAV
by: Li, Mengyuan, et al.
Published: (2025)
by: Li, Mengyuan, et al.
Published: (2025)
Intelligent Known and Novel Aircraft Recognition -- A Shift from Classification to Similarity Learning for Combat Identification
by: Saeed, Ahmad, et al.
Published: (2024)
by: Saeed, Ahmad, et al.
Published: (2024)
Evaluating Facial Expression Recognition Datasets for Deep Learning: A Benchmark Study with Novel Similarity Metrics
by: Gaya-Morey, F. Xavier, et al.
Published: (2025)
by: Gaya-Morey, F. Xavier, et al.
Published: (2025)
Box-Free Model Watermarks Are Prone to Black-Box Removal Attacks
by: An, Haonan, et al.
Published: (2024)
by: An, Haonan, et al.
Published: (2024)
IDNet: A Novel Dataset for Identity Document Analysis and Fraud Detection
by: Guan, Hong, et al.
Published: (2024)
by: Guan, Hong, et al.
Published: (2024)
AnyThermal: Towards Learning Universal Representations for Thermal Perception
by: Maheshwari, Parv, et al.
Published: (2026)
by: Maheshwari, Parv, et al.
Published: (2026)
Semantic Similarity Score for Measuring Visual Similarity at Semantic Level
by: Fan, Senran, et al.
Published: (2024)
by: Fan, Senran, et al.
Published: (2024)
BoxTuning: Directly Injecting the Object Box for Multimodal Model Fine-Tuning
by: Qian, Zekun, et al.
Published: (2026)
by: Qian, Zekun, et al.
Published: (2026)
Leveraging Large Models to Evaluate Novel Content: A Case Study on Advertisement Creativity
by: Hou, Zhaoyi Joey, et al.
Published: (2025)
by: Hou, Zhaoyi Joey, et al.
Published: (2025)
Similar Items
-
FSMDet: Vision-guided feature diffusion for fully sparse 3D detector
by: Liu, Tianran, et al.
Published: (2024) -
MOT FCG++: Enhanced Representation of Spatio-temporal Motion and Appearance Features
by: Fang, Yanzhao
Published: (2024) -
ID-Sim: An Identity-Focused Similarity Metric
by: Chae, Julia, et al.
Published: (2026) -
MAML MOT: Multiple Object Tracking based on Meta-Learning
by: Chen, Jiayi, et al.
Published: (2024) -
Hyperbolic Space Learning Method Leveraging Temporal Motion Priors for Human Mesh Recovery
by: Zhang, Xiang, et al.
Published: (2025)