Saved in:
| Main Authors: | Wu, Yihao, Zhao, Di, Zhang, Jingfeng, Koh, Yun Sing |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | https://arxiv.org/abs/2410.22927 |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
MetaWild: A Multimodal Dataset for Animal Re-Identification with Environmental Metadata
by: Li, Yuzhuo, et al.
Published: (2025)
by: Li, Yuzhuo, et al.
Published: (2025)
Animal Re-Identification on Microcontrollers
by: Chen, Yubo, et al.
Published: (2025)
by: Chen, Yubo, et al.
Published: (2025)
Sequence Transferability and Task Order Selection in Continual Learning
by: Nguyen, Thinh, et al.
Published: (2025)
by: Nguyen, Thinh, et al.
Published: (2025)
Knowledge-Guided Failure Prediction: Detecting When Object Detectors Miss Safety-Critical Objects
by: Zimmermann, Jakob Paul, et al.
Published: (2026)
by: Zimmermann, Jakob Paul, et al.
Published: (2026)
Identity documents recognition and detection using semantic segmentation with convolutional neural network
by: Kozlenko, Mykola, et al.
Published: (2025)
by: Kozlenko, Mykola, et al.
Published: (2025)
SPARF: Large-Scale Learning of 3D Sparse Radiance Fields from Few Input Images
by: Hamdi, Abdullah, et al.
Published: (2022)
by: Hamdi, Abdullah, et al.
Published: (2022)
Towards Optimal Convolutional Transfer Learning Architectures for Breast Lesion Classification and ACL Tear Detection
by: Frees, Daniel, et al.
Published: (2025)
by: Frees, Daniel, et al.
Published: (2025)
Harmony: A Joint Self-Supervised and Weakly-Supervised Framework for Learning General Purpose Visual Representations
by: Baharoon, Mohammed, et al.
Published: (2024)
by: Baharoon, Mohammed, et al.
Published: (2024)
CNNtention: Can CNNs do better with Attention?
by: Kapila, Nikhil, et al.
Published: (2024)
by: Kapila, Nikhil, et al.
Published: (2024)
Explainable Image Similarity: Integrating Siamese Networks and Grad-CAM
by: Livieris, Ioannis E., et al.
Published: (2023)
by: Livieris, Ioannis E., et al.
Published: (2023)
Spatial-ViLT: Enhancing Visual Spatial Reasoning through Multi-Task Learning
by: Islam, Chashi Mahiul, et al.
Published: (2025)
by: Islam, Chashi Mahiul, et al.
Published: (2025)
A review of Recent Techniques for Person Re-Identification
by: Asperti, Andrea, et al.
Published: (2025)
by: Asperti, Andrea, et al.
Published: (2025)
RefineFormer3D: Efficient 3D Medical Image Segmentation via Adaptive Multi-Scale Transformer with Cross Attention Fusion
by: Tyagi, Kavyansh, et al.
Published: (2026)
by: Tyagi, Kavyansh, et al.
Published: (2026)
Less Detail, Better Answers: Degradation-Driven Prompting for VQA
by: Han, Haoxuan, et al.
Published: (2026)
by: Han, Haoxuan, et al.
Published: (2026)
Search Multilayer Perceptron-Based Fusion for Efficient and Accurate Siamese Tracking
by: Shen, Tianqi, et al.
Published: (2026)
by: Shen, Tianqi, et al.
Published: (2026)
SegQC: a segmentation network-based framework for multi-metric segmentation quality control and segmentation error detection in volumetric medical images
by: Specktor-Fadida, Bella, et al.
Published: (2024)
by: Specktor-Fadida, Bella, et al.
Published: (2024)
3D Reconstruction from Sketches
by: Talwar, Abhimanyu, et al.
Published: (2025)
by: Talwar, Abhimanyu, et al.
Published: (2025)
Instance Segmentation for Point Sets
by: Talwar, Abhimanyu, et al.
Published: (2025)
by: Talwar, Abhimanyu, et al.
Published: (2025)
Overcoming Catastrophic Forgetting in Federated Class-Incremental Learning via Federated Global Twin Generator
by: Nguyen, Thinh, et al.
Published: (2024)
by: Nguyen, Thinh, et al.
Published: (2024)
LayerAct: Advanced Activation Mechanism for Robust Inference of CNNs
by: Yoon, Kihyuk, et al.
Published: (2023)
by: Yoon, Kihyuk, et al.
Published: (2023)
Fast Data Aware Neural Architecture Search via Supernet Accelerated Evaluation
by: Njor, Emil, et al.
Published: (2025)
by: Njor, Emil, et al.
Published: (2025)
Model-agnostic Adversarial Attack and Defense for Vision-Language-Action Models
by: Xu, Haochuan, et al.
Published: (2025)
by: Xu, Haochuan, et al.
Published: (2025)
TALON: Test-time Adaptive Learning for On-the-Fly Category Discovery
by: Wu, Yanan, et al.
Published: (2026)
by: Wu, Yanan, et al.
Published: (2026)
Privacy-Preserving Low-Rank Adaptation against Membership Inference Attacks for Latent Diffusion Models
by: Luo, Zihao, et al.
Published: (2024)
by: Luo, Zihao, et al.
Published: (2024)
Modular Deep Learning Framework for Assistive Perception: Gaze, Affect, and Speaker Identification
by: Anchan, Akshit Pramod, et al.
Published: (2025)
by: Anchan, Akshit Pramod, et al.
Published: (2025)
LADI v2: Multi-label Dataset and Classifiers for Low-Altitude Disaster Imagery
by: Scheele, Samuel, et al.
Published: (2024)
by: Scheele, Samuel, et al.
Published: (2024)
Classifying Healthy and Defective Fruits with a Multi-Input Architecture and CNN Models
by: Chuquimarca, Luis, et al.
Published: (2024)
by: Chuquimarca, Luis, et al.
Published: (2024)
Enhancing Apple's Defect Classification: Insights from Visible Spectrum and Narrow Spectral Band Imaging
by: Coello, Omar, et al.
Published: (2024)
by: Coello, Omar, et al.
Published: (2024)
Unsupervised Anomaly Detection Using Diffusion Trend Analysis for Display Inspection
by: Kim, Eunwoo, et al.
Published: (2024)
by: Kim, Eunwoo, et al.
Published: (2024)
On the Inherent Robustness of One-Stage Object Detection against Out-of-Distribution Data
by: Martinez-Seras, Aitor, et al.
Published: (2024)
by: Martinez-Seras, Aitor, et al.
Published: (2024)
Nearest Neighbor Projection Removal Adversarial Training
by: Singh, Himanshu, et al.
Published: (2025)
by: Singh, Himanshu, et al.
Published: (2025)
MotionFollower: Editing Video Motion via Lightweight Score-Guided Diffusion
by: Tu, Shuyuan, et al.
Published: (2024)
by: Tu, Shuyuan, et al.
Published: (2024)
Inclusive AI for Group Interactions: Predicting Gaze-Direction Behaviors in People with Intellectual and Developmental Disabilities
by: Huang, Giulia, et al.
Published: (2026)
by: Huang, Giulia, et al.
Published: (2026)
AMANet: Advancing SAR Ship Detection with Adaptive Multi-Hierarchical Attention Network
by: Ma, Xiaolin, et al.
Published: (2024)
by: Ma, Xiaolin, et al.
Published: (2024)
MvBody: Multi-View-Based Hybrid Transformer Using Optical 3D Body Scan for Explainable Cesarean Section Prediction
by: Cheng, Ruting, et al.
Published: (2025)
by: Cheng, Ruting, et al.
Published: (2025)
OptiRoulette Optimizer: A New Stochastic Meta-Optimizer for up to 5.3x Faster Convergence
by: Mastromichalakis, Stamatis
Published: (2026)
by: Mastromichalakis, Stamatis
Published: (2026)
Visual-Text Cross Alignment: Refining the Similarity Score in Vision-Language Models
by: Li, Jinhao, et al.
Published: (2024)
by: Li, Jinhao, et al.
Published: (2024)
The MSR-Video to Text Dataset with Clean Annotations
by: Chen, Haoran, et al.
Published: (2021)
by: Chen, Haoran, et al.
Published: (2021)
Large Language Models Powered Context-aware Motion Prediction in Autonomous Driving
by: Zheng, Xiaoji, et al.
Published: (2024)
by: Zheng, Xiaoji, et al.
Published: (2024)
Towards Multilingual Audio-Visual Question Answering
by: Phukan, Orchid Chetia, et al.
Published: (2024)
by: Phukan, Orchid Chetia, et al.
Published: (2024)
Similar Items
-
MetaWild: A Multimodal Dataset for Animal Re-Identification with Environmental Metadata
by: Li, Yuzhuo, et al.
Published: (2025) -
Animal Re-Identification on Microcontrollers
by: Chen, Yubo, et al.
Published: (2025) -
Sequence Transferability and Task Order Selection in Continual Learning
by: Nguyen, Thinh, et al.
Published: (2025) -
Knowledge-Guided Failure Prediction: Detecting When Object Detectors Miss Safety-Critical Objects
by: Zimmermann, Jakob Paul, et al.
Published: (2026) -
Identity documents recognition and detection using semantic segmentation with convolutional neural network
by: Kozlenko, Mykola, et al.
Published: (2025)