Hybrid Visual Telemetry for Bandwidth-Constrained Robotic Vision: A Pilot Study with HEVC Base Video and JPEG ROI Stills
Fuente:
arXiv
Saved in:
| Main Authors: | Trukhina, Natalia, Vashkelis, Vadim |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Mobile Traffic Camera Calibration from Road Geometry for UAV-Based Traffic Surveillance
by: Popov, Alexey, et al.
Published: (2026)
by: Popov, Alexey, et al.
Published: (2026)
SemanticZip: A Pilot Framework for Lossy Text Compression with LLMs as Semantic Decompressors
by: Trukhina, Natalia, et al.
Published: (2026)
by: Trukhina, Natalia, et al.
Published: (2026)
HI-MoE: Hierarchical Instance-Conditioned Mixture-of-Experts for Object Detection
by: Vashkelis, Vadim, et al.
Published: (2026)
by: Vashkelis, Vadim, et al.
Published: (2026)
Compress the Context, Keep the Commitments: A Formal Framework for Verifiable LLM Context Compression
by: Trukhina, Natalia, et al.
Published: (2026)
by: Trukhina, Natalia, et al.
Published: (2026)
Learning Visually Interpretable Oscillator Networks for Soft Continuum Robots from Video
by: Krauss, Henrik, et al.
Published: (2025)
by: Krauss, Henrik, et al.
Published: (2025)
EffiComm: Bandwidth Efficient Multi Agent Communication
by: Yazgan, Melih, et al.
Published: (2025)
by: Yazgan, Melih, et al.
Published: (2025)
AnyTeleop: A General Vision-Based Dexterous Robot Arm-Hand Teleoperation System
by: Qin, Yuzhe, et al.
Published: (2023)
by: Qin, Yuzhe, et al.
Published: (2023)
SoloParkour: Constrained Reinforcement Learning for Visual Locomotion from Privileged Experience
by: Chane-Sane, Elliot, et al.
Published: (2024)
by: Chane-Sane, Elliot, et al.
Published: (2024)
Video2Reward: Generating Reward Function from Videos for Legged Robot Behavior Learning
by: Zeng, Runhao, et al.
Published: (2024)
by: Zeng, Runhao, et al.
Published: (2024)
Benchmarking Vision, Language, & Action Models on Robotic Learning Tasks
by: Guruprasad, Pranav, et al.
Published: (2024)
by: Guruprasad, Pranav, et al.
Published: (2024)
LLARVA: Vision-Action Instruction Tuning Enhances Robot Learning
by: Niu, Dantong, et al.
Published: (2024)
by: Niu, Dantong, et al.
Published: (2024)
AMPLIFY: Actionless Motion Priors for Robot Learning from Videos
by: Collins, Jeremy A., et al.
Published: (2025)
by: Collins, Jeremy A., et al.
Published: (2025)
Squint: Fast Visual Reinforcement Learning for Sim-to-Real Robotics
by: Almuzairee, Abdulaziz, et al.
Published: (2026)
by: Almuzairee, Abdulaziz, et al.
Published: (2026)
Merging and Disentangling Views in Visual Reinforcement Learning for Robotic Manipulation
by: Almuzairee, Abdulaziz, et al.
Published: (2025)
by: Almuzairee, Abdulaziz, et al.
Published: (2025)
Point Cloud Models Improve Visual Robustness in Robotic Learners
by: Peri, Skand, et al.
Published: (2024)
by: Peri, Skand, et al.
Published: (2024)
VLM See, Robot Do: Human Demo Video to Robot Action Plan via Vision Language Model
by: Wang, Beichen, et al.
Published: (2024)
by: Wang, Beichen, et al.
Published: (2024)
PVI: Plug-in Visual Injection for Vision-Language-Action Models
by: Zhang, Zezhou, et al.
Published: (2026)
by: Zhang, Zezhou, et al.
Published: (2026)
ZeroMimic: Distilling Robotic Manipulation Skills from Web Videos
by: Shi, Junyao, et al.
Published: (2025)
by: Shi, Junyao, et al.
Published: (2025)
Visual Perception Engine: Fast and Flexible Multi-Head Inference for Robotic Vision Tasks
by: Łucki, Jakub, et al.
Published: (2025)
by: Łucki, Jakub, et al.
Published: (2025)
MARVL: Multi-Stage Guidance for Robotic Manipulation via Vision-Language Models
by: Zhou, Xunlan, et al.
Published: (2026)
by: Zhou, Xunlan, et al.
Published: (2026)
Test-Time Training for Visual Foresight Vision-Language-Action Models
by: Park, Sangwu, et al.
Published: (2026)
by: Park, Sangwu, et al.
Published: (2026)
Vision-Based Safety System for Barrierless Human-Robot Collaboration
by: Amaya-Mejía, Lina María, et al.
Published: (2022)
by: Amaya-Mejía, Lina María, et al.
Published: (2022)
Compressor-VLA: Instruction-Guided Visual Token Compression for Efficient Robotic Manipulation
by: Gao, Juntao, et al.
Published: (2025)
by: Gao, Juntao, et al.
Published: (2025)
VO-DP: Semantic-Geometric Adaptive Diffusion Policy for Vision-Only Robotic Manipulation
by: Ni, Zehao, et al.
Published: (2025)
by: Ni, Zehao, et al.
Published: (2025)
ChatVLA: Unified Multimodal Understanding and Robot Control with Vision-Language-Action Model
by: Zhou, Zhongyi, et al.
Published: (2025)
by: Zhou, Zhongyi, et al.
Published: (2025)
AVA-VLA: Improving Vision-Language-Action models with Active Visual Attention
by: Xiao, Lei, et al.
Published: (2025)
by: Xiao, Lei, et al.
Published: (2025)
Out of Sight, Still in Mind: Reasoning and Planning about Unobserved Objects with Video Tracking Enabled Memory Models
by: Huang, Yixuan, et al.
Published: (2023)
by: Huang, Yixuan, et al.
Published: (2023)
ViViDex: Learning Vision-based Dexterous Manipulation from Human Videos
by: Chen, Zerui, et al.
Published: (2024)
by: Chen, Zerui, et al.
Published: (2024)
Vision-based Manipulation from Single Human Video with Open-World Object Graphs
by: Zhu, Yifeng, et al.
Published: (2024)
by: Zhu, Yifeng, et al.
Published: (2024)
Scalable Vision-Language-Action Model Pretraining for Robotic Manipulation with Real-Life Human Activity Videos
by: Li, Qixiu, et al.
Published: (2025)
by: Li, Qixiu, et al.
Published: (2025)
Being-H0: Vision-Language-Action Pretraining from Large-Scale Human Videos
by: Luo, Hao, et al.
Published: (2025)
by: Luo, Hao, et al.
Published: (2025)
A Vision-Based Shared-Control Teleoperation Scheme for Controlling the Robotic Arm of a Four-Legged Robot
by: da Silva, Murilo Vinicius, et al.
Published: (2025)
by: da Silva, Murilo Vinicius, et al.
Published: (2025)
GR-2: A Generative Video-Language-Action Model with Web-Scale Knowledge for Robot Manipulation
by: Cheang, Chi-Lam, et al.
Published: (2024)
by: Cheang, Chi-Lam, et al.
Published: (2024)
Nothing Stands Still: A Spatiotemporal Benchmark on 3D Point Cloud Registration Under Large Geometric and Temporal Change
by: Sun, Tao, et al.
Published: (2023)
by: Sun, Tao, et al.
Published: (2023)
Beyond Domain Randomization: Event-Inspired Perception for Visually Robust Adversarial Imitation from Videos
by: Ramazzina, Andrea, et al.
Published: (2025)
by: Ramazzina, Andrea, et al.
Published: (2025)
Modeling Robotics Dataset Construction as an Artifact-Based Build Process
by: Pohl, Leon, et al.
Published: (2026)
by: Pohl, Leon, et al.
Published: (2026)
Continuous Object State Recognition for Cooking Robots Using Pre-Trained Vision-Language Models and Black-box Optimization
by: Kawaharazuka, Kento, et al.
Published: (2024)
by: Kawaharazuka, Kento, et al.
Published: (2024)
Bootstrapping Reinforcement Learning with Imitation for Vision-Based Agile Flight
by: Xing, Jiaxu, et al.
Published: (2024)
by: Xing, Jiaxu, et al.
Published: (2024)
ReasonDrive: Efficient Visual Question Answering for Autonomous Vehicles with Reasoning-Enhanced Small Vision-Language Models
by: Chahe, Amirhosein, et al.
Published: (2025)
by: Chahe, Amirhosein, et al.
Published: (2025)
Reinforcement Learning of Dolly-In Filming Using a Ground-Based Robot
by: Lorimer, Philip, et al.
Published: (2025)
by: Lorimer, Philip, et al.
Published: (2025)
Similar Items
-
Mobile Traffic Camera Calibration from Road Geometry for UAV-Based Traffic Surveillance
by: Popov, Alexey, et al.
Published: (2026) -
SemanticZip: A Pilot Framework for Lossy Text Compression with LLMs as Semantic Decompressors
by: Trukhina, Natalia, et al.
Published: (2026) -
HI-MoE: Hierarchical Instance-Conditioned Mixture-of-Experts for Object Detection
by: Vashkelis, Vadim, et al.
Published: (2026) -
Compress the Context, Keep the Commitments: A Formal Framework for Verifiable LLM Context Compression
by: Trukhina, Natalia, et al.
Published: (2026) -
Learning Visually Interpretable Oscillator Networks for Soft Continuum Robots from Video
by: Krauss, Henrik, et al.
Published: (2025)