EffiPerception: an Efficient Framework for Various Perception Tasks
Fuente:
arXiv
Guardado en:
| Autores principales: | Xiang, Xinhao, Dräger, Simon, Zhang, Jiawei |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Are AI-Generated Driving Videos Ready for Autonomous Driving? A Diagnostic Evaluation Framework
por: Xiang, Xinhao, et al.
Publicado: (2025)
por: Xiang, Xinhao, et al.
Publicado: (2025)
PTQAT: A Hybrid Parameter-Efficient Quantization Algorithm for 3D Perception Tasks
por: Wang, Xinhao, et al.
Publicado: (2025)
por: Wang, Xinhao, et al.
Publicado: (2025)
TTSA3R: Training-Free Temporal-Spatial Adaptive Persistent State for Streaming 3D Reconstruction
por: Zheng, Zhijie, et al.
Publicado: (2026)
por: Zheng, Zhijie, et al.
Publicado: (2026)
EffiVED:Efficient Video Editing via Text-instruction Diffusion Models
por: Zhang, Zhenghao, et al.
Publicado: (2024)
por: Zhang, Zhenghao, et al.
Publicado: (2024)
EffiMiniVLM: A Compact Dual-Encoder Regression Framework
por: Khor, Yin-Loon, et al.
Publicado: (2026)
por: Khor, Yin-Loon, et al.
Publicado: (2026)
An Efficient Adaptive Compression Method for Human Perception and Machine Vision Tasks
por: Liu, Lei, et al.
Publicado: (2025)
por: Liu, Lei, et al.
Publicado: (2025)
CoT4Det: A Chain-of-Thought Framework for Perception-Oriented Vision-Language Tasks
por: Qi, Yu, et al.
Publicado: (2025)
por: Qi, Yu, et al.
Publicado: (2025)
A Point-Based Approach to Efficient LiDAR Multi-Task Perception
por: Lang, Christopher, et al.
Publicado: (2024)
por: Lang, Christopher, et al.
Publicado: (2024)
EffiComm: Bandwidth Efficient Multi Agent Communication
por: Yazgan, Melih, et al.
Publicado: (2025)
por: Yazgan, Melih, et al.
Publicado: (2025)
All-in-One Transferring Image Compression from Human Perception to Multi-Machine Perception
por: Zhao, Jiancheng, et al.
Publicado: (2025)
por: Zhao, Jiancheng, et al.
Publicado: (2025)
AIGVE-Tool: AI-Generated Video Evaluation Toolkit with Multifaceted Benchmark
por: Xiang, Xinhao, et al.
Publicado: (2025)
por: Xiang, Xinhao, et al.
Publicado: (2025)
Efficient Universal Perception Encoder
por: Zhu, Chenchen, et al.
Publicado: (2026)
por: Zhu, Chenchen, et al.
Publicado: (2026)
Generation is Required for Data-Efficient Perception
por: Brady, Jack, et al.
Publicado: (2025)
por: Brady, Jack, et al.
Publicado: (2025)
Task-Aware Image Signal Processor for Advanced Visual Perception
por: Chen, Kai, et al.
Publicado: (2025)
por: Chen, Kai, et al.
Publicado: (2025)
TurboTrain: Towards Efficient and Balanced Multi-Task Learning for Multi-Agent Perception and Prediction
por: Zhou, Zewei, et al.
Publicado: (2025)
por: Zhou, Zewei, et al.
Publicado: (2025)
GeoFocus: Blending Efficient Global-to-Local Perception for Multimodal Geometry Problem-Solving
por: Deng, Linger, et al.
Publicado: (2026)
por: Deng, Linger, et al.
Publicado: (2026)
Resource Efficient Perception for Vision Systems
por: Subramanyam, A V, et al.
Publicado: (2024)
por: Subramanyam, A V, et al.
Publicado: (2024)
BEVCon: Advancing Bird's Eye View Perception with Contrastive Learning
por: Leng, Ziyang, et al.
Publicado: (2025)
por: Leng, Ziyang, et al.
Publicado: (2025)
CORP: A Multi-Modal Dataset for Campus-Oriented Roadside Perception Tasks
por: Wang, Beibei, et al.
Publicado: (2024)
por: Wang, Beibei, et al.
Publicado: (2024)
MMTL-UniAD: A Unified Framework for Multimodal and Multi-Task Learning in Assistive Driving Perception
por: Liu, Wenzhuo, et al.
Publicado: (2025)
por: Liu, Wenzhuo, et al.
Publicado: (2025)
Perception in Plan: Coupled Perception and Planning for End-to-End Autonomous Driving
por: Zhang, Bozhou, et al.
Publicado: (2025)
por: Zhang, Bozhou, et al.
Publicado: (2025)
A Prediction-as-Perception Framework for 3D Object Detection
por: Zhang, Song, et al.
Publicado: (2026)
por: Zhang, Song, et al.
Publicado: (2026)
CoSense3D: an Agent-based Efficient Learning Framework for Collective Perception
por: Yuan, Yunshuang, et al.
Publicado: (2024)
por: Yuan, Yunshuang, et al.
Publicado: (2024)
Which2comm: An Efficient Collaborative Perception Framework for 3D Object Detection
por: Yu, Duanrui, et al.
Publicado: (2025)
por: Yu, Duanrui, et al.
Publicado: (2025)
Towards Open-Vocabulary Multimodal 3D Object Detection with Attributes
por: Xiang, Xinhao, et al.
Publicado: (2025)
por: Xiang, Xinhao, et al.
Publicado: (2025)
OccFace: Unified Occlusion-Aware Facial Landmark Detection with Per-Point Visibility
por: Xiang, Xinhao, et al.
Publicado: (2026)
por: Xiang, Xinhao, et al.
Publicado: (2026)
Efficient Dual-domain Image Dehazing with Haze Prior Perception
por: Zheng, Lirong, et al.
Publicado: (2025)
por: Zheng, Lirong, et al.
Publicado: (2025)
Decoupling Perception and Calibration: Label-Efficient Image Quality Assessment Framework
por: Li, Xinyue, et al.
Publicado: (2026)
por: Li, Xinyue, et al.
Publicado: (2026)
Efficiently Disentangling CLIP for Multi-Object Perception
por: Rawlekar, Samyak, et al.
Publicado: (2025)
por: Rawlekar, Samyak, et al.
Publicado: (2025)
UV-M3TL: A Unified and Versatile Multimodal Multi-Task Learning Framework for Assistive Driving Perception
por: Liu, Wenzhuo, et al.
Publicado: (2026)
por: Liu, Wenzhuo, et al.
Publicado: (2026)
An Extensible Framework for Open Heterogeneous Collaborative Perception
por: Lu, Yifan, et al.
Publicado: (2024)
por: Lu, Yifan, et al.
Publicado: (2024)
TinyBEV: Cross Modal Knowledge Distillation for Efficient Multi Task Bird's Eye View Perception and Planning
por: Khan, Reeshad, et al.
Publicado: (2025)
por: Khan, Reeshad, et al.
Publicado: (2025)
Solution for Point Tracking Task of ECCV 2nd Perception Test Challenge 2024
por: Zhang, Yuxuan, et al.
Publicado: (2024)
por: Zhang, Yuxuan, et al.
Publicado: (2024)
Multi-Task Learning for Robot Perception with Imbalanced Data
por: Erkent, Ozgur
Publicado: (2026)
por: Erkent, Ozgur
Publicado: (2026)
Perception in Reflection
por: Wei, Yana, et al.
Publicado: (2025)
por: Wei, Yana, et al.
Publicado: (2025)
The Solution for Single Object Tracking Task of Perception Test Challenge 2024
por: Zhong, Zhiqiang, et al.
Publicado: (2024)
por: Zhong, Zhiqiang, et al.
Publicado: (2024)
TriLiteNet: Lightweight Model for Multi-Task Visual Perception
por: Che, Quang-Huy, et al.
Publicado: (2025)
por: Che, Quang-Huy, et al.
Publicado: (2025)
CoLC: Communication-Efficient Collaborative Perception with LiDAR Completion
por: Han, Yushan, et al.
Publicado: (2026)
por: Han, Yushan, et al.
Publicado: (2026)
ESAM++: Efficient Online 3D Perception on the Edge
por: Liu, Qin, et al.
Publicado: (2026)
por: Liu, Qin, et al.
Publicado: (2026)
On the Federated Learning Framework for Cooperative Perception
por: Zhang, Zhenrong, et al.
Publicado: (2024)
por: Zhang, Zhenrong, et al.
Publicado: (2024)
Ejemplares similares
-
Are AI-Generated Driving Videos Ready for Autonomous Driving? A Diagnostic Evaluation Framework
por: Xiang, Xinhao, et al.
Publicado: (2025) -
PTQAT: A Hybrid Parameter-Efficient Quantization Algorithm for 3D Perception Tasks
por: Wang, Xinhao, et al.
Publicado: (2025) -
TTSA3R: Training-Free Temporal-Spatial Adaptive Persistent State for Streaming 3D Reconstruction
por: Zheng, Zhijie, et al.
Publicado: (2026) -
EffiVED:Efficient Video Editing via Text-instruction Diffusion Models
por: Zhang, Zhenghao, et al.
Publicado: (2024) -
EffiMiniVLM: A Compact Dual-Encoder Regression Framework
por: Khor, Yin-Loon, et al.
Publicado: (2026)