Task-based Loss Functions in Computer Vision: A Comprehensive Review
Fuente:
arXiv
Guardado en:
| Autores principales: | Elharrouss, Omar, Mahmood, Yasir, Bechqito, Yassine, Serhani, Mohamed Adel, Badidi, Elarbi, Riffi, Jamal, Tairi, Hamid |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Transformer-based Image and Video Inpainting: Current Challenges and Future Directions
por: Elharrouss, Omar, et al.
Publicado: (2024)
por: Elharrouss, Omar, et al.
Publicado: (2024)
Backbones-Review: Feature Extraction Networks for Deep Learning and Deep Reinforcement Learning Approaches
por: Elharrouss, Omar, et al.
Publicado: (2022)
por: Elharrouss, Omar, et al.
Publicado: (2022)
On the Generalizability of Iterative Patch Selection for Memory-Efficient High-Resolution Image Classification
por: Riffi-Aslett, Max, et al.
Publicado: (2024)
por: Riffi-Aslett, Max, et al.
Publicado: (2024)
Applications of Knowledge Distillation in Remote Sensing: A Survey
por: Himeur, Yassine, et al.
Publicado: (2024)
por: Himeur, Yassine, et al.
Publicado: (2024)
PDC-ViT : Source Camera Identification using Pixel Difference Convolution and Vision Transformer
por: Elharrouss, Omar, et al.
Publicado: (2025)
por: Elharrouss, Omar, et al.
Publicado: (2025)
Drone-type-Set: Drone types detection benchmark for drone detection and tracking
por: AlDosari, Kholoud, et al.
Publicado: (2024)
por: AlDosari, Kholoud, et al.
Publicado: (2024)
U-NetMN and SegNetMN: Modified U-Net and SegNet models for bimodal SAR image segmentation
por: Kzadri, Marwane, et al.
Publicado: (2025)
por: Kzadri, Marwane, et al.
Publicado: (2025)
A Comprehensive Review of Knowledge Distillation in Computer Vision
por: Habib, Gousia, et al.
Publicado: (2024)
por: Habib, Gousia, et al.
Publicado: (2024)
A Comprehensive Review on Computer Vision Analysis of Aerial Data
por: Tetarwal, Vivek, et al.
Publicado: (2024)
por: Tetarwal, Vivek, et al.
Publicado: (2024)
3D objects and scenes classification, recognition, segmentation, and reconstruction using 3D point cloud data: A review
por: Elharrouss, Omar, et al.
Publicado: (2023)
por: Elharrouss, Omar, et al.
Publicado: (2023)
Exploring Variational Autoencoders for Medical Image Generation: A Comprehensive Study
por: Rais, Khadija, et al.
Publicado: (2024)
por: Rais, Khadija, et al.
Publicado: (2024)
Advanced Deep Learning and Large Language Models: Comprehensive Insights for Cancer Detection
por: Habchi, Yassine, et al.
Publicado: (2025)
por: Habchi, Yassine, et al.
Publicado: (2025)
Computer Vision-Based Early Detection of Container Loss at Sea
por: Lall, Vishakha, et al.
Publicado: (2026)
por: Lall, Vishakha, et al.
Publicado: (2026)
End-to-end Inception-Unet based Generative Adversarial Networks for Snow and Rain Removals
por: Kajo, Ibrahim, et al.
Publicado: (2024)
por: Kajo, Ibrahim, et al.
Publicado: (2024)
A Physics-Informed Loss Function for Boundary-Consistent and Robust Artery Segmentation in DSA Sequences
por: Irfan, Muhammad, et al.
Publicado: (2025)
por: Irfan, Muhammad, et al.
Publicado: (2025)
A Review of Transformer-Based Models for Computer Vision Tasks: Capturing Global Context and Spatial Relationships
por: Pereira, Gracile Astlin, et al.
Publicado: (2024)
por: Pereira, Gracile Astlin, et al.
Publicado: (2024)
Reasoning in Computer Vision: Taxonomy, Models, Tasks, and Methodologies
por: Sarkar, Ayushman, et al.
Publicado: (2025)
por: Sarkar, Ayushman, et al.
Publicado: (2025)
Optimizing Hyper parameters in CNN for Soil Classification using PSO and Whale Optimization Algorithm
por: Ibrahim, Yasir Nooruldeen, et al.
Publicado: (2025)
por: Ibrahim, Yasir Nooruldeen, et al.
Publicado: (2025)
Beyond Anonymization: Object Scrubbing for Privacy-Preserving 2D and 3D Vision Tasks
por: Ertan, Murat Bilgehan, et al.
Publicado: (2025)
por: Ertan, Murat Bilgehan, et al.
Publicado: (2025)
Fuzzy Theory in Computer Vision: A Review
por: Yerkin, Adilet, et al.
Publicado: (2025)
por: Yerkin, Adilet, et al.
Publicado: (2025)
A Weighted Vision Transformer-Based Multi-Task Learning Framework for Predicting ADAS-Cog Scores
por: Hamid, Nur Amirah Abd, et al.
Publicado: (2025)
por: Hamid, Nur Amirah Abd, et al.
Publicado: (2025)
When Multi-Task Learning Meets Partial Supervision: A Computer Vision Review
por: Fontana, Maxime, et al.
Publicado: (2023)
por: Fontana, Maxime, et al.
Publicado: (2023)
Continual Learning on a Diet: Learning from Sparsely Labeled Streams Under Constrained Computation
por: Zhang, Wenxuan, et al.
Publicado: (2024)
por: Zhang, Wenxuan, et al.
Publicado: (2024)
Tracking and Mapping in Medical Computer Vision: A Review
por: Schmidt, Adam, et al.
Publicado: (2023)
por: Schmidt, Adam, et al.
Publicado: (2023)
SGW-based Multi-Task Learning in Vision Tasks
por: Zhang, Ruiyuan, et al.
Publicado: (2024)
por: Zhang, Ruiyuan, et al.
Publicado: (2024)
Multidimensional Task Learning: A Unified Tensor Framework for Computer Vision Tasks
por: Ichi, Alaa El, et al.
Publicado: (2026)
por: Ichi, Alaa El, et al.
Publicado: (2026)
Face Pyramid Vision Transformer
por: Islam, Khawar, et al.
Publicado: (2022)
por: Islam, Khawar, et al.
Publicado: (2022)
Benchmarking Large Vision-Language Models on Fine-Grained Image Tasks: A Comprehensive Evaluation
por: Yu, Hong-Tao, et al.
Publicado: (2025)
por: Yu, Hong-Tao, et al.
Publicado: (2025)
Deep Learning-Based 3D Instance and Semantic Segmentation: A Review
por: Yasir, Siddiqui Muhammad, et al.
Publicado: (2024)
por: Yasir, Siddiqui Muhammad, et al.
Publicado: (2024)
Olympus: A Universal Task Router for Computer Vision Tasks
por: Lin, Yuanze, et al.
Publicado: (2024)
por: Lin, Yuanze, et al.
Publicado: (2024)
CLIP-DPO: Vision-Language Models as a Source of Preference for Fixing Hallucinations in LVLMs
por: Ouali, Yassine, et al.
Publicado: (2024)
por: Ouali, Yassine, et al.
Publicado: (2024)
A Probabilistic Rotation Representation for Symmetric Shapes With an Efficiently Computable Bingham Loss Function
por: Sato, Hiroya, et al.
Publicado: (2023)
por: Sato, Hiroya, et al.
Publicado: (2023)
Deep Learning for Computer Vision based Activity Recognition and Fall Detection of the Elderly: a Systematic Review
por: Gaya-Morey, F. Xavier, et al.
Publicado: (2024)
por: Gaya-Morey, F. Xavier, et al.
Publicado: (2024)
A New Simple Vision Algorithm for Detecting the Enzymic Browning Defects in Golden Delicious Apples
por: Balanji, Hamid Majidi
Publicado: (2021)
por: Balanji, Hamid Majidi
Publicado: (2021)
A Comprehensive Review of YOLO Architectures in Computer Vision: From YOLOv1 to YOLOv8 and YOLO-NAS
por: Terven, Juan, et al.
Publicado: (2023)
por: Terven, Juan, et al.
Publicado: (2023)
Leveraging Perceptual Scores for Dataset Pruning in Computer Vision Tasks
por: Singh, Raghavendra
Publicado: (2024)
por: Singh, Raghavendra
Publicado: (2024)
Morphology-Aware KOA Classification: Integrating Graph Priors with Vision Models
por: Tliba, Marouane, et al.
Publicado: (2025)
por: Tliba, Marouane, et al.
Publicado: (2025)
Generalized Robust Fundus Photography-based Vision Loss Estimation for High Myopia
por: Yan, Zipei, et al.
Publicado: (2024)
por: Yan, Zipei, et al.
Publicado: (2024)
Tacchi 2.0: A Low Computational Cost and Comprehensive Dynamic Contact Simulator for Vision-based Tactile Sensors
por: Sun, Yuhao, et al.
Publicado: (2025)
por: Sun, Yuhao, et al.
Publicado: (2025)
Vision-Free Retrieval: Rethinking Multimodal Search with Textual Scene Descriptions
por: Ntinou, Ioanna, et al.
Publicado: (2025)
por: Ntinou, Ioanna, et al.
Publicado: (2025)
Ejemplares similares
-
Transformer-based Image and Video Inpainting: Current Challenges and Future Directions
por: Elharrouss, Omar, et al.
Publicado: (2024) -
Backbones-Review: Feature Extraction Networks for Deep Learning and Deep Reinforcement Learning Approaches
por: Elharrouss, Omar, et al.
Publicado: (2022) -
On the Generalizability of Iterative Patch Selection for Memory-Efficient High-Resolution Image Classification
por: Riffi-Aslett, Max, et al.
Publicado: (2024) -
Applications of Knowledge Distillation in Remote Sensing: A Survey
por: Himeur, Yassine, et al.
Publicado: (2024) -
PDC-ViT : Source Camera Identification using Pixel Difference Convolution and Vision Transformer
por: Elharrouss, Omar, et al.
Publicado: (2025)