edgeVLM: Cloud-edge Collaborative Real-time VLM based on Context Transfer
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Qian, Chen, Yu, Xinran, Huang, Zewen, Li, Danyang, Ma, Qiang, Dang, Fan, Ding, Xuan, Shang, Guangyong, Yang, Zheng |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
SwiftVLM: Efficient Vision-Language Model Inference via Cross-Layer Token Bypass
von: Qian, Chen, et al.
Veröffentlicht: (2026)
von: Qian, Chen, et al.
Veröffentlicht: (2026)
OpenMoCap: Rethinking Optical Motion Capture under Real-world Occlusion
von: Qian, Chen, et al.
Veröffentlicht: (2025)
von: Qian, Chen, et al.
Veröffentlicht: (2025)
Spa-VLM: Stealthy Poisoning Attacks on RAG-based VLM
von: Yu, Lei, et al.
Veröffentlicht: (2025)
von: Yu, Lei, et al.
Veröffentlicht: (2025)
VLM-CAD: VLM-Optimized Collaborative Agent Design Workflow for Analog Circuit Sizing
von: Pan, Guanyuan, et al.
Veröffentlicht: (2026)
von: Pan, Guanyuan, et al.
Veröffentlicht: (2026)
Swim2Real: VLM-Guided System Identification for Sim-to-Real Transfer
von: Qiu, Kevin, et al.
Veröffentlicht: (2026)
von: Qiu, Kevin, et al.
Veröffentlicht: (2026)
OpenMap: Instruction Grounding via Open-Vocabulary Visual-Language Mapping
von: Li, Danyang, et al.
Veröffentlicht: (2025)
von: Li, Danyang, et al.
Veröffentlicht: (2025)
Critical edge statistics for deformed GinUEs
von: Liu, Dang-Zheng, et al.
Veröffentlicht: (2023)
von: Liu, Dang-Zheng, et al.
Veröffentlicht: (2023)
Fast-dVLM: Efficient Block-Diffusion VLM via Direct Conversion from Autoregressive VLM
von: Wu, Chengyue, et al.
Veröffentlicht: (2026)
von: Wu, Chengyue, et al.
Veröffentlicht: (2026)
DriveGenVLM: Real-world Video Generation for Vision Language Model based Autonomous Driving
von: Fu, Yongjie, et al.
Veröffentlicht: (2024)
von: Fu, Yongjie, et al.
Veröffentlicht: (2024)
Physics-Guided VLM Priors for All-Cloud Removal
von: Xu, Liying, et al.
Veröffentlicht: (2026)
von: Xu, Liying, et al.
Veröffentlicht: (2026)
DocVLM: Make Your VLM an Efficient Reader
von: Nacson, Mor Shpigel, et al.
Veröffentlicht: (2024)
von: Nacson, Mor Shpigel, et al.
Veröffentlicht: (2024)
Efficient Cloud-edge Collaborative Approaches to SPARQL Queries over Large RDF graphs
von: Ma, Shidan, et al.
Veröffentlicht: (2026)
von: Ma, Shidan, et al.
Veröffentlicht: (2026)
VLM-UDMC: VLM-Enhanced Unified Decision-Making and Motion Control for Urban Autonomous Driving
von: Liu, Haichao, et al.
Veröffentlicht: (2025)
von: Liu, Haichao, et al.
Veröffentlicht: (2025)
Small-Large Collaboration: Training-efficient Concept Personalization for Large VLM using a Meta Personalized Small VLM
von: Yang, Sihan, et al.
Veröffentlicht: (2025)
von: Yang, Sihan, et al.
Veröffentlicht: (2025)
Root Cause Localization for Microservice Systems in Cloud-edge Collaborative Environments
von: Zhu, Yuhan, et al.
Veröffentlicht: (2024)
von: Zhu, Yuhan, et al.
Veröffentlicht: (2024)
GKT: A Novel Guidance-Based Knowledge Transfer Framework For Efficient Cloud-edge Collaboration LLM Deployment
von: Yao, Yao, et al.
Veröffentlicht: (2024)
von: Yao, Yao, et al.
Veröffentlicht: (2024)
VLM6D: VLM based 6Dof Pose Estimation based on RGB-D Images
von: Sarowar, Md Selim, et al.
Veröffentlicht: (2025)
von: Sarowar, Md Selim, et al.
Veröffentlicht: (2025)
VLM-Pruner: Buffering for Spatial Sparsity in an Efficient VLM Centrifugal Token Pruning Paradigm
von: Wu, Zhenkai, et al.
Veröffentlicht: (2025)
von: Wu, Zhenkai, et al.
Veröffentlicht: (2025)
TRANSPORTER: Transferring Visual Semantics from VLM Manifolds
von: Stergiou, Alexandros
Veröffentlicht: (2025)
von: Stergiou, Alexandros
Veröffentlicht: (2025)
CoDriveVLM: VLM-Enhanced Urban Cooperative Dispatching and Motion Planning for Future Autonomous Mobility on Demand Systems
von: Liu, Haichao, et al.
Veröffentlicht: (2025)
von: Liu, Haichao, et al.
Veröffentlicht: (2025)
Multilingual VLM Training: Adapting an English-Trained VLM to French
von: Lahmi, Jules, et al.
Veröffentlicht: (2025)
von: Lahmi, Jules, et al.
Veröffentlicht: (2025)
VLM-Grounder: A VLM Agent for Zero-Shot 3D Visual Grounding
von: Xu, Runsen, et al.
Veröffentlicht: (2024)
von: Xu, Runsen, et al.
Veröffentlicht: (2024)
Capturing Gaze Shifts for Guidance: Cross-Modal Fusion Enhancement for VLM Hallucination Mitigation
von: Qi, Zheng, et al.
Veröffentlicht: (2025)
von: Qi, Zheng, et al.
Veröffentlicht: (2025)
VLM-SFD: VLM-Assisted Siamese Flow Diffusion Framework for Dual-Arm Cooperative Manipulation
von: Chen, Jiaming, et al.
Veröffentlicht: (2025)
von: Chen, Jiaming, et al.
Veröffentlicht: (2025)
Rethinking Intermediate Representation for VLM-based Robot Manipulation
von: Tang, Weiliang, et al.
Veröffentlicht: (2025)
von: Tang, Weiliang, et al.
Veröffentlicht: (2025)
EMAC+: Embodied Multimodal Agent for Collaborative Planning with VLM+LLM
von: Ao, Shuang, et al.
Veröffentlicht: (2025)
von: Ao, Shuang, et al.
Veröffentlicht: (2025)
VLM-TDP: VLM-guided Trajectory-conditioned Diffusion Policy for Robust Long-Horizon Manipulation
von: Huang, Kefeng, et al.
Veröffentlicht: (2025)
von: Huang, Kefeng, et al.
Veröffentlicht: (2025)
REO-VLM: Transforming VLM to Meet Regression Challenges in Earth Observation
von: Xue, Xizhe, et al.
Veröffentlicht: (2024)
von: Xue, Xizhe, et al.
Veröffentlicht: (2024)
EO-VLM: VLM-Guided Energy Overload Attacks on Vision Models
von: Seo, Minjae, et al.
Veröffentlicht: (2025)
von: Seo, Minjae, et al.
Veröffentlicht: (2025)
LensVLM: Selective Context Expansion for Compressed Visual Representation of Text
von: Xie, Roy, et al.
Veröffentlicht: (2026)
von: Xie, Roy, et al.
Veröffentlicht: (2026)
Nüwa: Mending the Spatial Integrity Torn by VLM Token Pruning
von: Huang, Yihong, et al.
Veröffentlicht: (2026)
von: Huang, Yihong, et al.
Veröffentlicht: (2026)
Evaluating Visual and Cultural Interpretation: The K-Viscuit Benchmark with Human-VLM Collaboration
von: Park, ChaeHun, et al.
Veröffentlicht: (2024)
von: Park, ChaeHun, et al.
Veröffentlicht: (2024)
SAIL: Test-Time Scaling for In-Context Imitation Learning with VLM
von: Sato, Makoto, et al.
Veröffentlicht: (2026)
von: Sato, Makoto, et al.
Veröffentlicht: (2026)
VLM-driven Behavior Tree for Context-aware Task Planning
von: Wake, Naoki, et al.
Veröffentlicht: (2025)
von: Wake, Naoki, et al.
Veröffentlicht: (2025)
MindPower: Enabling Theory-of-Mind Reasoning in VLM-based Embodied Agents
von: Zhang, Ruoxuan, et al.
Veröffentlicht: (2025)
von: Zhang, Ruoxuan, et al.
Veröffentlicht: (2025)
VLM-KD: Knowledge Distillation from VLM for Long-Tail Visual Recognition
von: Zhang, Zaiwei, et al.
Veröffentlicht: (2024)
von: Zhang, Zaiwei, et al.
Veröffentlicht: (2024)
MGFFD-VLM: Multi-Granularity Prompt Learning for Face Forgery Detection with VLM
von: Chen, Tao, et al.
Veröffentlicht: (2025)
von: Chen, Tao, et al.
Veröffentlicht: (2025)
Critic-V: VLM Critics Help Catch VLM Errors in Multimodal Reasoning
von: Zhang, Di, et al.
Veröffentlicht: (2024)
von: Zhang, Di, et al.
Veröffentlicht: (2024)
HapticVLM: VLM-Driven Texture Recognition Aimed at Intelligent Haptic Interaction
von: Khan, Muhammad Haris, et al.
Veröffentlicht: (2025)
von: Khan, Muhammad Haris, et al.
Veröffentlicht: (2025)
GoalVLM: VLM-driven Object Goal Navigation for Multi-Agent System
von: James, MoniJesu, et al.
Veröffentlicht: (2026)
von: James, MoniJesu, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
SwiftVLM: Efficient Vision-Language Model Inference via Cross-Layer Token Bypass
von: Qian, Chen, et al.
Veröffentlicht: (2026) -
OpenMoCap: Rethinking Optical Motion Capture under Real-world Occlusion
von: Qian, Chen, et al.
Veröffentlicht: (2025) -
Spa-VLM: Stealthy Poisoning Attacks on RAG-based VLM
von: Yu, Lei, et al.
Veröffentlicht: (2025) -
VLM-CAD: VLM-Optimized Collaborative Agent Design Workflow for Analog Circuit Sizing
von: Pan, Guanyuan, et al.
Veröffentlicht: (2026) -
Swim2Real: VLM-Guided System Identification for Sim-to-Real Transfer
von: Qiu, Kevin, et al.
Veröffentlicht: (2026)