Characterizing Disparity Between Edge Models and High-Accuracy Base Models for Vision Tasks
Fuente:
arXiv
Salvato in:
| Autori principali: | Wang, Zhenyu, Nirjon, Shahriar |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
mmJoints: Expanding Joint Representations Beyond (x,y,z) in mmWave-Based 3D Pose Estimation
di: Wang, Zhenyu, et al.
Pubblicazione: (2025)
di: Wang, Zhenyu, et al.
Pubblicazione: (2025)
mmWEAVER: Environment-Specific mmWave Signal Synthesis from a Photo and Activity Description
di: Monjur, Mahathir, et al.
Pubblicazione: (2025)
di: Monjur, Mahathir, et al.
Pubblicazione: (2025)
Collaborative Edge-to-Server Inference for Vision-Language Models
di: Song, Soochang, et al.
Pubblicazione: (2025)
di: Song, Soochang, et al.
Pubblicazione: (2025)
Evolving Prompt Adaptation for Vision-Language Models
di: Zhang, Enming, et al.
Pubblicazione: (2026)
di: Zhang, Enming, et al.
Pubblicazione: (2026)
NuWa: Deriving Lightweight Task-Specific Vision Transformers for Edge Devices
di: Wei, Ziteng, et al.
Pubblicazione: (2025)
di: Wei, Ziteng, et al.
Pubblicazione: (2025)
More Thought, Less Accuracy? On the Dual Nature of Reasoning in Vision-Language Models
di: Tian, Xinyu, et al.
Pubblicazione: (2025)
di: Tian, Xinyu, et al.
Pubblicazione: (2025)
LVLM-Aided Alignment of Task-Specific Vision Models
di: Koebler, Alexander, et al.
Pubblicazione: (2025)
di: Koebler, Alexander, et al.
Pubblicazione: (2025)
A Unified Debiasing Approach for Vision-Language Models across Modalities and Tasks
di: Jung, Hoin, et al.
Pubblicazione: (2024)
di: Jung, Hoin, et al.
Pubblicazione: (2024)
Aligned Vector Quantization for Edge-Cloud Collabrative Vision-Language Models
di: Liu, Xiao, et al.
Pubblicazione: (2024)
di: Liu, Xiao, et al.
Pubblicazione: (2024)
Federated Learning of Low-Rank One-Shot Image Detection Models in Edge Devices with Scalable Accuracy and Compute Complexity
di: Hannaan, Abdul, et al.
Pubblicazione: (2025)
di: Hannaan, Abdul, et al.
Pubblicazione: (2025)
Can Vision Models Truly Forget? Mirage: Representation-Level Certification of Visual Unlearning
di: Yu, Zhenyu, et al.
Pubblicazione: (2026)
di: Yu, Zhenyu, et al.
Pubblicazione: (2026)
Vision-Language Models for Edge Networks: A Comprehensive Survey
di: Sharshar, Ahmed, et al.
Pubblicazione: (2025)
di: Sharshar, Ahmed, et al.
Pubblicazione: (2025)
Refusal as Silence: Gendered Disparities in Vision-Language Model Responses
di: Luo, Sha, et al.
Pubblicazione: (2024)
di: Luo, Sha, et al.
Pubblicazione: (2024)
Vision Language Model-Empowered Contract Theory for AIGC Task Allocation in Teleoperation
di: Zhan, Zijun, et al.
Pubblicazione: (2024)
di: Zhan, Zijun, et al.
Pubblicazione: (2024)
GazeVLM: A Vision-Language Model for Multi-Task Gaze Understanding
di: Mathew, Athul M., et al.
Pubblicazione: (2025)
di: Mathew, Athul M., et al.
Pubblicazione: (2025)
CombatVLA: An Efficient Vision-Language-Action Model for Combat Tasks in 3D Action Role-Playing Games
di: Chen, Peng, et al.
Pubblicazione: (2025)
di: Chen, Peng, et al.
Pubblicazione: (2025)
Recurrent Reasoning with Vision-Language Models for Estimating Long-Horizon Embodied Task Progress
di: Zhang, Yuelin, et al.
Pubblicazione: (2026)
di: Zhang, Yuelin, et al.
Pubblicazione: (2026)
Jack of All Tasks, Master of Many: Designing General-purpose Coarse-to-Fine Vision-Language Model
di: Pramanick, Shraman, et al.
Pubblicazione: (2023)
di: Pramanick, Shraman, et al.
Pubblicazione: (2023)
GeoRSMLLM: A Multimodal Large Language Model for Vision-Language Tasks in Geoscience and Remote Sensing
di: Zhang, Zilun, et al.
Pubblicazione: (2025)
di: Zhang, Zilun, et al.
Pubblicazione: (2025)
mmCounter: Static People Counting in Dense Indoor Scenarios Using mmWave Radar
di: Toha, Tarik Reza, et al.
Pubblicazione: (2025)
di: Toha, Tarik Reza, et al.
Pubblicazione: (2025)
Enhancing Vehicle Make and Model Recognition with 3D Attention Modules
di: Semiromizadeh, Narges, et al.
Pubblicazione: (2025)
di: Semiromizadeh, Narges, et al.
Pubblicazione: (2025)
Exploring Disparity-Accuracy Trade-offs in Face Recognition Systems: The Role of Datasets, Architectures, and Loss Functions
di: Jaiswal, Siddharth D, et al.
Pubblicazione: (2025)
di: Jaiswal, Siddharth D, et al.
Pubblicazione: (2025)
EdgeSync: Accelerating Edge-Model Updates for Data Drift through Adaptive Continuous Learning
di: Donga, Runchu, et al.
Pubblicazione: (2025)
di: Donga, Runchu, et al.
Pubblicazione: (2025)
Edge Reliability Gap in Vision-Language Models: Quantifying Failure Modes of Compressed VLMs Under Visual Corruption
di: Erol, Mehmet Kaan
Pubblicazione: (2026)
di: Erol, Mehmet Kaan
Pubblicazione: (2026)
Enhancing Vision-Language Models for Autonomous Driving through Task-Specific Prompting and Spatial Reasoning
di: Wu, Aodi, et al.
Pubblicazione: (2025)
di: Wu, Aodi, et al.
Pubblicazione: (2025)
Towards Accurate UAV Image Perception: Guiding Vision-Language Models with Stronger Task Prompts
di: Guo, Mingning, et al.
Pubblicazione: (2025)
di: Guo, Mingning, et al.
Pubblicazione: (2025)
Cognition-Inspired Dual-Stream Semantic Enhancement for Vision-Based Dynamic Emotion Modeling
di: Wang, Huanzhen, et al.
Pubblicazione: (2026)
di: Wang, Huanzhen, et al.
Pubblicazione: (2026)
TAP-SLF: Parameter-Efficient Adaptation of Vision Foundation Models for Multi-Task Ultrasound Image Analysis
di: Wan, Hui, et al.
Pubblicazione: (2026)
di: Wan, Hui, et al.
Pubblicazione: (2026)
Beyond Generation: Multi-Hop Reasoning for Factual Accuracy in Vision-Language Models
di: Hossain, Shamima
Pubblicazione: (2025)
di: Hossain, Shamima
Pubblicazione: (2025)
Fast ODE-based Sampling for Diffusion Models in Around 5 Steps
di: Zhou, Zhenyu, et al.
Pubblicazione: (2023)
di: Zhou, Zhenyu, et al.
Pubblicazione: (2023)
Dynamic Universal Approximation Theory: The Basic Theory for Deep Learning-Based Computer Vision Models
di: Wang, Wei, et al.
Pubblicazione: (2024)
di: Wang, Wei, et al.
Pubblicazione: (2024)
Edge-AI for Agriculture: Lightweight Vision Models for Disease Detection in Resource-Limited Settings
di: Joshi, Harsh
Pubblicazione: (2024)
di: Joshi, Harsh
Pubblicazione: (2024)
Dynamic Weight Adjustment for Knowledge Distillation: Leveraging Vision Transformer for High-Accuracy Lung Cancer Detection and Real-Time Deployment
di: Khan, Saif Ur Rehman, et al.
Pubblicazione: (2025)
di: Khan, Saif Ur Rehman, et al.
Pubblicazione: (2025)
VLM4VLA: Revisiting Vision-Language-Models in Vision-Language-Action Models
di: Zhang, Jianke, et al.
Pubblicazione: (2026)
di: Zhang, Jianke, et al.
Pubblicazione: (2026)
Object-Centric Vision Token Pruning for Vision Language Models
di: Li, Guangyuan, et al.
Pubblicazione: (2025)
di: Li, Guangyuan, et al.
Pubblicazione: (2025)
CGEarthEye:A High-Resolution Remote Sensing Vision Foundation Model Based on the Jilin-1 Satellite Constellation
di: Yi, Zhiwei, et al.
Pubblicazione: (2025)
di: Yi, Zhiwei, et al.
Pubblicazione: (2025)
Beyond Attention Scores: SVD-Based Vision Token Pruning for Efficient Vision-Language Models
di: Apedo, Yvon, et al.
Pubblicazione: (2026)
di: Apedo, Yvon, et al.
Pubblicazione: (2026)
Egocentric Bias in Vision-Language Models
di: Wang, Maijunxian, et al.
Pubblicazione: (2026)
di: Wang, Maijunxian, et al.
Pubblicazione: (2026)
Improved Belief-Attention in Vision Task
di: Zhang, Guoqiang
Pubblicazione: (2026)
di: Zhang, Guoqiang
Pubblicazione: (2026)
INSIGHT: Enhancing Autonomous Driving Safety through Vision-Language Models on Context-Aware Hazard Detection and Edge Case Evaluation
di: Chen, Dianwei, et al.
Pubblicazione: (2025)
di: Chen, Dianwei, et al.
Pubblicazione: (2025)
Documenti analoghi
-
mmJoints: Expanding Joint Representations Beyond (x,y,z) in mmWave-Based 3D Pose Estimation
di: Wang, Zhenyu, et al.
Pubblicazione: (2025) -
mmWEAVER: Environment-Specific mmWave Signal Synthesis from a Photo and Activity Description
di: Monjur, Mahathir, et al.
Pubblicazione: (2025) -
Collaborative Edge-to-Server Inference for Vision-Language Models
di: Song, Soochang, et al.
Pubblicazione: (2025) -
Evolving Prompt Adaptation for Vision-Language Models
di: Zhang, Enming, et al.
Pubblicazione: (2026) -
NuWa: Deriving Lightweight Task-Specific Vision Transformers for Edge Devices
di: Wei, Ziteng, et al.
Pubblicazione: (2025)