A Comprehensive Review on Computer Vision Analysis of Aerial Data
Fuente:
arXiv
Guardado en:
| Autores principales: | Tetarwal, Vivek, Kumar, Sandeep |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
A Hybrid Ensemble Learning Framework for Image-Based Solar Panel Classification
por: Tetarwal, Vivek, et al.
Publicado: (2025)
por: Tetarwal, Vivek, et al.
Publicado: (2025)
Leveraging Perceptual Scores for Dataset Pruning in Computer Vision Tasks
por: Singh, Raghavendra
Publicado: (2024)
por: Singh, Raghavendra
Publicado: (2024)
Histogram Driven Amplitude Embedding for Qubit Efficient Quantum Image Compression
por: Tomar, Sahil, et al.
Publicado: (2025)
por: Tomar, Sahil, et al.
Publicado: (2025)
Task-aware Distributed Source Coding under Dynamic Bandwidth
por: Li, Po-han, et al.
Publicado: (2023)
por: Li, Po-han, et al.
Publicado: (2023)
TokenCom: Vision-Language Model for Multimodal and Multitask Token Communications
por: Jiang, Feibo, et al.
Publicado: (2026)
por: Jiang, Feibo, et al.
Publicado: (2026)
Recursive Vision Transformer with Dynamic Depth and Width Adjustment for Resource-Efficient Image Semantic Communication
por: Zhang, Zhilong, et al.
Publicado: (2026)
por: Zhang, Zhilong, et al.
Publicado: (2026)
Challenges and Solutions in Selecting Optimal Lossless Data Compression Algorithms
por: Rahman, Md. Atiqur, et al.
Publicado: (2025)
por: Rahman, Md. Atiqur, et al.
Publicado: (2025)
VIBE: Annotation-Free Video-to-Text Information Bottleneck Evaluation for TL;DR
por: Chen, Shenghui, et al.
Publicado: (2025)
por: Chen, Shenghui, et al.
Publicado: (2025)
Enhancing 3D Robotic Vision Robustness by Minimizing Adversarial Mutual Information through a Curriculum Training Approach
por: Darabi, Nastaran, et al.
Publicado: (2024)
por: Darabi, Nastaran, et al.
Publicado: (2024)
A New Statistical Approach to the Performance Analysis of Vision-based Localization
por: Hu, Haozhou, et al.
Publicado: (2025)
por: Hu, Haozhou, et al.
Publicado: (2025)
A Comprehensive Review of Knowledge Distillation in Computer Vision
por: Habib, Gousia, et al.
Publicado: (2024)
por: Habib, Gousia, et al.
Publicado: (2024)
Vision Transformer-based Semantic Communications With Importance-Aware Quantization
por: Park, Joohyuk, et al.
Publicado: (2024)
por: Park, Joohyuk, et al.
Publicado: (2024)
Exploiting Information Redundancy in Attention Maps for Extreme Quantization of Vision Transformers
por: Maisonnave, Lucas, et al.
Publicado: (2025)
por: Maisonnave, Lucas, et al.
Publicado: (2025)
A Novel Image Similarity Metric for Scene Composition Structure
por: Haque, Md Redwanul, et al.
Publicado: (2025)
por: Haque, Md Redwanul, et al.
Publicado: (2025)
Robust Sparse Signal Recovery with Outliers: A Hard Thresholding Pursuit Approach Based on LAD
por: Xu, Jiao, et al.
Publicado: (2026)
por: Xu, Jiao, et al.
Publicado: (2026)
Quantum walk inspired JPEG compression of images
por: Verma, Abhishek, et al.
Publicado: (2026)
por: Verma, Abhishek, et al.
Publicado: (2026)
Directional Confusions Reveal Divergent Inductive Biases Through Rate-Distortion Geometry in Human and Machine Vision
por: Caglar, Leyla Roksan, et al.
Publicado: (2026)
por: Caglar, Leyla Roksan, et al.
Publicado: (2026)
Alternating Minimization Schemes for Computing Rate-Distortion-Perception Functions with $f$-Divergence Perception Constraints
por: Serra, Giuseppe, et al.
Publicado: (2024)
por: Serra, Giuseppe, et al.
Publicado: (2024)
HawkRover: An Autonomous mmWave Vehicular Communication Testbed with Multi-sensor Fusion and Deep Learning
por: Zhu, Ethan, et al.
Publicado: (2024)
por: Zhu, Ethan, et al.
Publicado: (2024)
Spatial Channel State Information Prediction with Generative AI: Towards Holographic Communication and Digital Radio Twin
por: Zhang, Lihao, et al.
Publicado: (2024)
por: Zhang, Lihao, et al.
Publicado: (2024)
Shot Segmentation Based on Von Neumann Entropy for Key Frame Extraction
por: Zhang, Xueqing, et al.
Publicado: (2024)
por: Zhang, Xueqing, et al.
Publicado: (2024)
Track Initialization and Re-Identification for~3D Multi-View Multi-Object Tracking
por: Van Ma, Linh, et al.
Publicado: (2024)
por: Van Ma, Linh, et al.
Publicado: (2024)
Fast and Robust Phase Retrieval via Deep Expectation-Consistent Approximation
por: Shastri, Saurav K., et al.
Publicado: (2024)
por: Shastri, Saurav K., et al.
Publicado: (2024)
Untrained Neural Nets for Snapshot Compressive Imaging: Theory and Algorithms
por: Zhao, Mengyu, et al.
Publicado: (2024)
por: Zhao, Mengyu, et al.
Publicado: (2024)
Exposing the Deception: Uncovering More Forgery Clues for Deepfake Detection
por: Ba, Zhongjie, et al.
Publicado: (2024)
por: Ba, Zhongjie, et al.
Publicado: (2024)
SAFT: Sensitivity-Aware Filtering and Transmission for Adaptive 3D Point Cloud Communication over Wireless Channels
por: Mekki, Huda Adam Sirag, et al.
Publicado: (2026)
por: Mekki, Huda Adam Sirag, et al.
Publicado: (2026)
Continual Few-shot Adaptation for Synthetic Fingerprint Detection
por: Benjamin, Joseph Geo, et al.
Publicado: (2026)
por: Benjamin, Joseph Geo, et al.
Publicado: (2026)
Scale What Counts, Mask What Matters: Evaluating Foundation Models for Zero-Shot Cross-Domain Wi-Fi Sensing
por: Jiang, Cheng, et al.
Publicado: (2025)
por: Jiang, Cheng, et al.
Publicado: (2025)
Edge Collaborative Gaussian Splatting with Integrated Rendering and Communication
por: Wan, Yujie, et al.
Publicado: (2025)
por: Wan, Yujie, et al.
Publicado: (2025)
Compression Beyond Pixels: Semantic Compression with Multimodal Foundation Models
por: Shen, Ruiqi, et al.
Publicado: (2025)
por: Shen, Ruiqi, et al.
Publicado: (2025)
Rate-Distortion Limits for Multimodal Retrieval: Theory, Optimal Codes, and Finite-Sample Guarantees
por: Chen, Thomas Y.
Publicado: (2025)
por: Chen, Thomas Y.
Publicado: (2025)
Sampling Strategies for Efficient Training of Deep Learning Object Detection Algorithms
por: Shen, Gefei, et al.
Publicado: (2025)
por: Shen, Gefei, et al.
Publicado: (2025)
Mixture of Balanced Information Bottlenecks for Long-Tailed Visual Recognition
por: Lan, Yifan, et al.
Publicado: (2025)
por: Lan, Yifan, et al.
Publicado: (2025)
Information Entropy-Based Framework for Quantifying Tortuosity in Meibomian Gland Uneven Atrophy
por: Wang, Kesheng, et al.
Publicado: (2025)
por: Wang, Kesheng, et al.
Publicado: (2025)
Evolving Token Communication with Parametric Memory Network
por: Chen, Weixuan, et al.
Publicado: (2026)
por: Chen, Weixuan, et al.
Publicado: (2026)
Deep Learning for Climate Action: Computer Vision Analysis of Visual Narratives on X
por: Prasse, Katharina, et al.
Publicado: (2025)
por: Prasse, Katharina, et al.
Publicado: (2025)
Scaling Training Data with Lossy Image Compression
por: Mentzer, Katherine L., et al.
Publicado: (2024)
por: Mentzer, Katherine L., et al.
Publicado: (2024)
Theoretical Guarantees of Data Augmented Last Layer Retraining Methods
por: Welfert, Monica, et al.
Publicado: (2024)
por: Welfert, Monica, et al.
Publicado: (2024)
Generalizations of the Normalized Radon Cumulative Distribution Transform for Limited Data Recognition
por: Beckmann, Matthias, et al.
Publicado: (2025)
por: Beckmann, Matthias, et al.
Publicado: (2025)
Successive Interference Cancellation-aided Diffusion Models for Joint Channel Estimation and Data Detection in Low Rank Channel Scenarios
por: Bhattacharya, Sagnik, et al.
Publicado: (2025)
por: Bhattacharya, Sagnik, et al.
Publicado: (2025)
Ejemplares similares
-
A Hybrid Ensemble Learning Framework for Image-Based Solar Panel Classification
por: Tetarwal, Vivek, et al.
Publicado: (2025) -
Leveraging Perceptual Scores for Dataset Pruning in Computer Vision Tasks
por: Singh, Raghavendra
Publicado: (2024) -
Histogram Driven Amplitude Embedding for Qubit Efficient Quantum Image Compression
por: Tomar, Sahil, et al.
Publicado: (2025) -
Task-aware Distributed Source Coding under Dynamic Bandwidth
por: Li, Po-han, et al.
Publicado: (2023) -
TokenCom: Vision-Language Model for Multimodal and Multitask Token Communications
por: Jiang, Feibo, et al.
Publicado: (2026)