Mondrian: On-Device High-Performance Video Analytics with Compressive Packed Inference
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Jeon, Changmin, Kim, Seonjun, Yi, Juheon, Lee, Youngki |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
VLM in a flash: I/O-Efficient Sparsification of Vision-Language Model via Neuron Chunking
von: Yang, Kichang, et al.
Veröffentlicht: (2025)
von: Yang, Kichang, et al.
Veröffentlicht: (2025)
Enabling Cross-Camera Collaboration for Video Analytics on Distributed Smart Cameras
von: Min, Chulhong, et al.
Veröffentlicht: (2024)
von: Min, Chulhong, et al.
Veröffentlicht: (2024)
Attention-Propagation Network for Egocentric Heatmap to 3D Pose Lifting
von: Kang, Taeho, et al.
Veröffentlicht: (2024)
von: Kang, Taeho, et al.
Veröffentlicht: (2024)
MPMAvatar: Learning 3D Gaussian Avatars with Accurate and Robust Physics-Based Dynamics
von: Lee, Changmin, et al.
Veröffentlicht: (2025)
von: Lee, Changmin, et al.
Veröffentlicht: (2025)
Clustered Error Correction with Grouped 4D Gaussian Splatting
von: Kang, Taeho, et al.
Veröffentlicht: (2025)
von: Kang, Taeho, et al.
Veröffentlicht: (2025)
Towards Real-time Video Compressive Sensing on Mobile Devices
von: Cao, Miao, et al.
Veröffentlicht: (2024)
von: Cao, Miao, et al.
Veröffentlicht: (2024)
Revisiting Cross-Domain Problem for LiDAR-based 3D Object Detection
von: Zhang, Ruixiao, et al.
Veröffentlicht: (2024)
von: Zhang, Ruixiao, et al.
Veröffentlicht: (2024)
PackUV: Packed Gaussian UV Maps for 4D Volumetric Video
von: Rai, Aashish, et al.
Veröffentlicht: (2026)
von: Rai, Aashish, et al.
Veröffentlicht: (2026)
WorldPack: Compressed Memory Improves Spatial Consistency in Video World Modeling
von: Oshima, Yuta, et al.
Veröffentlicht: (2025)
von: Oshima, Yuta, et al.
Veröffentlicht: (2025)
GaussianVideo: Efficient Video Representation and Compression by Gaussian Splatting
von: Lee, Inseo, et al.
Veröffentlicht: (2025)
von: Lee, Inseo, et al.
Veröffentlicht: (2025)
Conditional Video Generation for High-Efficiency Video Compression
von: Yi, Fangqiu, et al.
Veröffentlicht: (2025)
von: Yi, Fangqiu, et al.
Veröffentlicht: (2025)
ROI-Packing: Efficient Region-Based Compression for Machine Vision
von: Eimon, Md Eimran Hossain, et al.
Veröffentlicht: (2025)
von: Eimon, Md Eimran Hossain, et al.
Veröffentlicht: (2025)
PackForcing: Short Video Training Suffices for Long Video Sampling and Long Context Inference
von: Mao, Xiaofeng, et al.
Veröffentlicht: (2026)
von: Mao, Xiaofeng, et al.
Veröffentlicht: (2026)
PhysHanDI: Physics-Based Reconstruction of Hand-Deformable Object Interactions
von: Lee, Jihyun, et al.
Veröffentlicht: (2026)
von: Lee, Jihyun, et al.
Veröffentlicht: (2026)
Video Compression Commander: Plug-and-Play Inference Acceleration for Video Large Language Models
von: Liu, Xuyang, et al.
Veröffentlicht: (2025)
von: Liu, Xuyang, et al.
Veröffentlicht: (2025)
On-device Sora: Enabling Training-Free Diffusion-based Text-to-Video Generation for Mobile Devices
von: Kim, Bosung, et al.
Veröffentlicht: (2025)
von: Kim, Bosung, et al.
Veröffentlicht: (2025)
On-device Sora: Enabling Training-Free Diffusion-based Text-to-Video Generation for Mobile Devices
von: Kim, Bosung, et al.
Veröffentlicht: (2025)
von: Kim, Bosung, et al.
Veröffentlicht: (2025)
Revisiting Weakly-Supervised Video Scene Graph Generation via Pair Affinity Learning
von: Kang, Minseok, et al.
Veröffentlicht: (2026)
von: Kang, Minseok, et al.
Veröffentlicht: (2026)
D-Cube: Exploiting Hyper-Features of Diffusion Model for Robust Medical Classification
von: Jang, Minhee, et al.
Veröffentlicht: (2024)
von: Jang, Minhee, et al.
Veröffentlicht: (2024)
Gait Recognition from Highly Compressed Videos
von: Niculae, Andrei, et al.
Veröffentlicht: (2024)
von: Niculae, Andrei, et al.
Veröffentlicht: (2024)
Detect Closer Surfaces that can be Seen: New Modeling and Evaluation in Cross-domain 3D Object Detection
von: Zhang, Ruixiao, et al.
Veröffentlicht: (2024)
von: Zhang, Ruixiao, et al.
Veröffentlicht: (2024)
CompSplat: Compression-aware 3D Gaussian Splatting for Real-world Video
von: Song, Hojun, et al.
Veröffentlicht: (2026)
von: Song, Hojun, et al.
Veröffentlicht: (2026)
PM-VIS+: High-Performance Video Instance Segmentation without Video Annotation
von: Yang, Zhangjing, et al.
Veröffentlicht: (2024)
von: Yang, Zhangjing, et al.
Veröffentlicht: (2024)
Simplifying Two-Stage Detectors for On-Device Inference in Remote Sensing
von: Kang, Jaemin, et al.
Veröffentlicht: (2024)
von: Kang, Jaemin, et al.
Veröffentlicht: (2024)
SIn-NeRF2NeRF: Editing 3D Scenes with Instructions through Segmentation and Inpainting
von: Hong, Jiseung, et al.
Veröffentlicht: (2024)
von: Hong, Jiseung, et al.
Veröffentlicht: (2024)
GranAlign: Granularity-Aware Alignment Framework for Zero-Shot Video Moment Retrieval
von: Jeon, Mingyu, et al.
Veröffentlicht: (2026)
von: Jeon, Mingyu, et al.
Veröffentlicht: (2026)
Sali4Vid: Saliency-Aware Video Reweighting and Adaptive Caption Retrieval for Dense Video Captioning
von: Jeon, MinJu, et al.
Veröffentlicht: (2025)
von: Jeon, MinJu, et al.
Veröffentlicht: (2025)
OTT-Vid: Optimal Transport Temporal Token Compression for Video Large Language Models
von: Kang, Minseok, et al.
Veröffentlicht: (2026)
von: Kang, Minseok, et al.
Veröffentlicht: (2026)
MobiFuse: A High-Precision On-device Depth Perception System with Multi-Data Fusion
von: Zhang, Jinrui, et al.
Veröffentlicht: (2024)
von: Zhang, Jinrui, et al.
Veröffentlicht: (2024)
Neural varifolds: an aggregate representation for quantifying the geometry of point clouds
von: Lee, Juheon, et al.
Veröffentlicht: (2024)
von: Lee, Juheon, et al.
Veröffentlicht: (2024)
OmniVLM: A Token-Compressed, Sub-Billion-Parameter Vision-Language Model for Efficient On-Device Inference
von: Chen, Wei, et al.
Veröffentlicht: (2024)
von: Chen, Wei, et al.
Veröffentlicht: (2024)
Deep Forcing: Training-Free Long Video Generation with Deep Sink and Participative Compression
von: Yi, Jung, et al.
Veröffentlicht: (2025)
von: Yi, Jung, et al.
Veröffentlicht: (2025)
Follow the Saliency: Supervised Saliency for Retrieval-augmented Dense Video Captioning
von: Choi, Seung hee, et al.
Veröffentlicht: (2026)
von: Choi, Seung hee, et al.
Veröffentlicht: (2026)
PM-VIS: High-Performance Box-Supervised Video Instance Segmentation
von: Yang, Zhangjing, et al.
Veröffentlicht: (2024)
von: Yang, Zhangjing, et al.
Veröffentlicht: (2024)
Pack and Detect: Fast Object Detection in Videos Using Region-of-Interest Packing
von: Kumar, Athindran Ramesh, et al.
Veröffentlicht: (2018)
von: Kumar, Athindran Ramesh, et al.
Veröffentlicht: (2018)
Learning Multi-View Spatial Reasoning from Cross-View Relations
von: Jeong, Suchae, et al.
Veröffentlicht: (2026)
von: Jeong, Suchae, et al.
Veröffentlicht: (2026)
EMCompress: Video-LLMs with Endomorphic Multimodal Compression
von: Fan, Zheyu, et al.
Veröffentlicht: (2025)
von: Fan, Zheyu, et al.
Veröffentlicht: (2025)
MORDA: A Synthetic Dataset to Facilitate Adaptation of Object Detectors to Unseen Real-target Domain While Preserving Performance on Real-source Domain
von: Lim, Hojun, et al.
Veröffentlicht: (2025)
von: Lim, Hojun, et al.
Veröffentlicht: (2025)
Perception-Oriented Latent Coding for High-Performance Compressed Domain Semantic Inference
von: Zhang, Xu, et al.
Veröffentlicht: (2025)
von: Zhang, Xu, et al.
Veröffentlicht: (2025)
Hierarchically Structured Neural Bones for Reconstructing Animatable Objects from Casual Videos
von: Jeon, Subin, et al.
Veröffentlicht: (2024)
von: Jeon, Subin, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
VLM in a flash: I/O-Efficient Sparsification of Vision-Language Model via Neuron Chunking
von: Yang, Kichang, et al.
Veröffentlicht: (2025) -
Enabling Cross-Camera Collaboration for Video Analytics on Distributed Smart Cameras
von: Min, Chulhong, et al.
Veröffentlicht: (2024) -
Attention-Propagation Network for Egocentric Heatmap to 3D Pose Lifting
von: Kang, Taeho, et al.
Veröffentlicht: (2024) -
MPMAvatar: Learning 3D Gaussian Avatars with Accurate and Robust Physics-Based Dynamics
von: Lee, Changmin, et al.
Veröffentlicht: (2025) -
Clustered Error Correction with Grouped 4D Gaussian Splatting
von: Kang, Taeho, et al.
Veröffentlicht: (2025)