Leveraging Perceptual Scores for Dataset Pruning in Computer Vision Tasks
Fuente:
arXiv
Gespeichert in:
| 1. Verfasser: | Singh, Raghavendra |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
A Comprehensive Review on Computer Vision Analysis of Aerial Data
von: Tetarwal, Vivek, et al.
Veröffentlicht: (2024)
von: Tetarwal, Vivek, et al.
Veröffentlicht: (2024)
Perceptual Scales Predicted by Fisher Information Metrics
von: Vacher, Jonathan, et al.
Veröffentlicht: (2023)
von: Vacher, Jonathan, et al.
Veröffentlicht: (2023)
Task-aware Distributed Source Coding under Dynamic Bandwidth
von: Li, Po-han, et al.
Veröffentlicht: (2023)
von: Li, Po-han, et al.
Veröffentlicht: (2023)
TokenCom: Vision-Language Model for Multimodal and Multitask Token Communications
von: Jiang, Feibo, et al.
Veröffentlicht: (2026)
von: Jiang, Feibo, et al.
Veröffentlicht: (2026)
Recursive Vision Transformer with Dynamic Depth and Width Adjustment for Resource-Efficient Image Semantic Communication
von: Zhang, Zhilong, et al.
Veröffentlicht: (2026)
von: Zhang, Zhilong, et al.
Veröffentlicht: (2026)
How Much Is a Dataset Worth? Scaling Laws, the Vendi Score, and Matrix Spectral Functions
von: Bilmes, Jeff A., et al.
Veröffentlicht: (2026)
von: Bilmes, Jeff A., et al.
Veröffentlicht: (2026)
Task-Oriented Co-Design of Communication, Computing, and Control for Edge-Enabled Industrial Cyber-Physical Systems
von: Diao, Yufeng, et al.
Veröffentlicht: (2025)
von: Diao, Yufeng, et al.
Veröffentlicht: (2025)
SpectraIrisPAD: Leveraging Vision Foundation Models for Spectrally Conditioned Multispectral Iris Presentation Attack Detection
von: Ramachandra, Raghavendra, et al.
Veröffentlicht: (2025)
von: Ramachandra, Raghavendra, et al.
Veröffentlicht: (2025)
Enhancing 3D Robotic Vision Robustness by Minimizing Adversarial Mutual Information through a Curriculum Training Approach
von: Darabi, Nastaran, et al.
Veröffentlicht: (2024)
von: Darabi, Nastaran, et al.
Veröffentlicht: (2024)
LVM4CSI: Enabling Direct Application of Pre-Trained Large Vision Models for Wireless Channel Tasks
von: Guo, Jiajia, et al.
Veröffentlicht: (2025)
von: Guo, Jiajia, et al.
Veröffentlicht: (2025)
Vision Transformer-based Semantic Communications With Importance-Aware Quantization
von: Park, Joohyuk, et al.
Veröffentlicht: (2024)
von: Park, Joohyuk, et al.
Veröffentlicht: (2024)
Exploiting Information Redundancy in Attention Maps for Extreme Quantization of Vision Transformers
von: Maisonnave, Lucas, et al.
Veröffentlicht: (2025)
von: Maisonnave, Lucas, et al.
Veröffentlicht: (2025)
Leveraging Human-Machine Interactions for Computer Vision Dataset Quality Enhancement
von: Anzaku, Esla Timothy, et al.
Veröffentlicht: (2024)
von: Anzaku, Esla Timothy, et al.
Veröffentlicht: (2024)
Towards Efficient VLMs: Information-Theoretic Driven Compression via Adaptive Structural Pruning
von: Xu, Zhaoqi, et al.
Veröffentlicht: (2025)
von: Xu, Zhaoqi, et al.
Veröffentlicht: (2025)
Synonymous Variational Inference for Perceptual Image Compression
von: Liang, Zijian, et al.
Veröffentlicht: (2025)
von: Liang, Zijian, et al.
Veröffentlicht: (2025)
TwinTURBO: Semi-Supervised Fine-Tuning of Foundation Models via Mutual Information Decompositions for Downstream Task and Latent Spaces
von: Quétant, Guillaume, et al.
Veröffentlicht: (2025)
von: Quétant, Guillaume, et al.
Veröffentlicht: (2025)
High Perceptual Quality Wireless Image Delivery with Denoising Diffusion Models
von: Yilmaz, Selim F., et al.
Veröffentlicht: (2023)
von: Yilmaz, Selim F., et al.
Veröffentlicht: (2023)
Directional Confusions Reveal Divergent Inductive Biases Through Rate-Distortion Geometry in Human and Machine Vision
von: Caglar, Leyla Roksan, et al.
Veröffentlicht: (2026)
von: Caglar, Leyla Roksan, et al.
Veröffentlicht: (2026)
Aligning Task- and Reconstruction-Oriented Communications for Edge Intelligence
von: Diao, Yufeng, et al.
Veröffentlicht: (2025)
von: Diao, Yufeng, et al.
Veröffentlicht: (2025)
Alternating Minimization Schemes for Computing Rate-Distortion-Perception Functions with $f$-Divergence Perception Constraints
von: Serra, Giuseppe, et al.
Veröffentlicht: (2024)
von: Serra, Giuseppe, et al.
Veröffentlicht: (2024)
HawkRover: An Autonomous mmWave Vehicular Communication Testbed with Multi-sensor Fusion and Deep Learning
von: Zhu, Ethan, et al.
Veröffentlicht: (2024)
von: Zhu, Ethan, et al.
Veröffentlicht: (2024)
Spatial Channel State Information Prediction with Generative AI: Towards Holographic Communication and Digital Radio Twin
von: Zhang, Lihao, et al.
Veröffentlicht: (2024)
von: Zhang, Lihao, et al.
Veröffentlicht: (2024)
Shot Segmentation Based on Von Neumann Entropy for Key Frame Extraction
von: Zhang, Xueqing, et al.
Veröffentlicht: (2024)
von: Zhang, Xueqing, et al.
Veröffentlicht: (2024)
Track Initialization and Re-Identification for~3D Multi-View Multi-Object Tracking
von: Van Ma, Linh, et al.
Veröffentlicht: (2024)
von: Van Ma, Linh, et al.
Veröffentlicht: (2024)
Fast and Robust Phase Retrieval via Deep Expectation-Consistent Approximation
von: Shastri, Saurav K., et al.
Veröffentlicht: (2024)
von: Shastri, Saurav K., et al.
Veröffentlicht: (2024)
Untrained Neural Nets for Snapshot Compressive Imaging: Theory and Algorithms
von: Zhao, Mengyu, et al.
Veröffentlicht: (2024)
von: Zhao, Mengyu, et al.
Veröffentlicht: (2024)
Exposing the Deception: Uncovering More Forgery Clues for Deepfake Detection
von: Ba, Zhongjie, et al.
Veröffentlicht: (2024)
von: Ba, Zhongjie, et al.
Veröffentlicht: (2024)
SAFT: Sensitivity-Aware Filtering and Transmission for Adaptive 3D Point Cloud Communication over Wireless Channels
von: Mekki, Huda Adam Sirag, et al.
Veröffentlicht: (2026)
von: Mekki, Huda Adam Sirag, et al.
Veröffentlicht: (2026)
Continual Few-shot Adaptation for Synthetic Fingerprint Detection
von: Benjamin, Joseph Geo, et al.
Veröffentlicht: (2026)
von: Benjamin, Joseph Geo, et al.
Veröffentlicht: (2026)
A Novel Image Similarity Metric for Scene Composition Structure
von: Haque, Md Redwanul, et al.
Veröffentlicht: (2025)
von: Haque, Md Redwanul, et al.
Veröffentlicht: (2025)
Scale What Counts, Mask What Matters: Evaluating Foundation Models for Zero-Shot Cross-Domain Wi-Fi Sensing
von: Jiang, Cheng, et al.
Veröffentlicht: (2025)
von: Jiang, Cheng, et al.
Veröffentlicht: (2025)
Edge Collaborative Gaussian Splatting with Integrated Rendering and Communication
von: Wan, Yujie, et al.
Veröffentlicht: (2025)
von: Wan, Yujie, et al.
Veröffentlicht: (2025)
Challenges and Solutions in Selecting Optimal Lossless Data Compression Algorithms
von: Rahman, Md. Atiqur, et al.
Veröffentlicht: (2025)
von: Rahman, Md. Atiqur, et al.
Veröffentlicht: (2025)
A Hybrid Ensemble Learning Framework for Image-Based Solar Panel Classification
von: Tetarwal, Vivek, et al.
Veröffentlicht: (2025)
von: Tetarwal, Vivek, et al.
Veröffentlicht: (2025)
Compression Beyond Pixels: Semantic Compression with Multimodal Foundation Models
von: Shen, Ruiqi, et al.
Veröffentlicht: (2025)
von: Shen, Ruiqi, et al.
Veröffentlicht: (2025)
Rate-Distortion Limits for Multimodal Retrieval: Theory, Optimal Codes, and Finite-Sample Guarantees
von: Chen, Thomas Y.
Veröffentlicht: (2025)
von: Chen, Thomas Y.
Veröffentlicht: (2025)
Robust Sparse Signal Recovery with Outliers: A Hard Thresholding Pursuit Approach Based on LAD
von: Xu, Jiao, et al.
Veröffentlicht: (2026)
von: Xu, Jiao, et al.
Veröffentlicht: (2026)
Sampling Strategies for Efficient Training of Deep Learning Object Detection Algorithms
von: Shen, Gefei, et al.
Veröffentlicht: (2025)
von: Shen, Gefei, et al.
Veröffentlicht: (2025)
Mixture of Balanced Information Bottlenecks for Long-Tailed Visual Recognition
von: Lan, Yifan, et al.
Veröffentlicht: (2025)
von: Lan, Yifan, et al.
Veröffentlicht: (2025)
Information Entropy-Based Framework for Quantifying Tortuosity in Meibomian Gland Uneven Atrophy
von: Wang, Kesheng, et al.
Veröffentlicht: (2025)
von: Wang, Kesheng, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
A Comprehensive Review on Computer Vision Analysis of Aerial Data
von: Tetarwal, Vivek, et al.
Veröffentlicht: (2024) -
Perceptual Scales Predicted by Fisher Information Metrics
von: Vacher, Jonathan, et al.
Veröffentlicht: (2023) -
Task-aware Distributed Source Coding under Dynamic Bandwidth
von: Li, Po-han, et al.
Veröffentlicht: (2023) -
TokenCom: Vision-Language Model for Multimodal and Multitask Token Communications
von: Jiang, Feibo, et al.
Veröffentlicht: (2026) -
Recursive Vision Transformer with Dynamic Depth and Width Adjustment for Resource-Efficient Image Semantic Communication
von: Zhang, Zhilong, et al.
Veröffentlicht: (2026)