Small, Versatile and Mighty: A Range-View Perception Framework
Fuente:
arXiv
Saved in:
| Main Authors: | Meng, Qiang, Wang, Xiao, Wang, JiaBao, Yan, Liujiang, Wang, Ke |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Towards Stable 3D Object Detection
by: Wang, Jiabao, et al.
Published: (2024)
by: Wang, Jiabao, et al.
Published: (2024)
Small but Mighty: Enhancing 3D Point Clouds Semantic Segmentation with U-Next Framework
by: Zeng, Ziyin, et al.
Published: (2023)
by: Zeng, Ziyin, et al.
Published: (2023)
OPUS: Occupancy Prediction Using a Sparse Set
by: Wang, Jiabao, et al.
Published: (2024)
by: Wang, Jiabao, et al.
Published: (2024)
Small but Mighty: Dynamic Wavelet Expert-Guided Fine-Tuning of Large-Scale Models for Optical Remote Sensing Object Segmentation
by: Sun, Yanguang, et al.
Published: (2026)
by: Sun, Yanguang, et al.
Published: (2026)
UV-M3TL: A Unified and Versatile Multimodal Multi-Task Learning Framework for Assistive Driving Perception
by: Liu, Wenzhuo, et al.
Published: (2026)
by: Liu, Wenzhuo, et al.
Published: (2026)
Spatio-Temporal Correlation Guided Geometric Partitioning for Versatile Video Coding
by: Meng, Xuewei, et al.
Published: (2026)
by: Meng, Xuewei, et al.
Published: (2026)
Vision-Driven 2D Supervised Fine-Tuning Framework for Bird's Eye View Perception
by: He, Lei, et al.
Published: (2024)
by: He, Lei, et al.
Published: (2024)
SparseFusion: Efficient Sparse Multi-Modal Fusion Framework for Long-Range 3D Perception
by: Li, Yiheng, et al.
Published: (2024)
by: Li, Yiheng, et al.
Published: (2024)
WAVE: Learning Unified & Versatile Audio-Visual Embeddings with Multimodal LLM
by: Tang, Changli, et al.
Published: (2025)
by: Tang, Changli, et al.
Published: (2025)
IA-MVS: Instance-Focused Adaptive Depth Sampling for Multi-View Stereo
by: Wang, Yinzhe, et al.
Published: (2025)
by: Wang, Yinzhe, et al.
Published: (2025)
A Versatile Framework for Multi-scene Person Re-identification
by: Zheng, Wei-Shi, et al.
Published: (2024)
by: Zheng, Wei-Shi, et al.
Published: (2024)
Long-SCOPE: Fully Sparse Long-Range Cooperative 3D Perception
by: Wang, Jiahao, et al.
Published: (2026)
by: Wang, Jiahao, et al.
Published: (2026)
A Dual-way Enhanced Framework from Text Matching Point of View for Multimodal Entity Linking
by: Song, Shezheng, et al.
Published: (2023)
by: Song, Shezheng, et al.
Published: (2023)
A Lightweight Multi-Scale Attention Framework for Real-Time Spinal Endoscopic Instance Segmentation
by: Lai, Qi, et al.
Published: (2025)
by: Lai, Qi, et al.
Published: (2025)
Real-IAD: A Real-World Multi-View Dataset for Benchmarking Versatile Industrial Anomaly Detection
by: Wang, Chengjie, et al.
Published: (2024)
by: Wang, Chengjie, et al.
Published: (2024)
PICD: Versatile Perceptual Image Compression with Diffusion Rendering
by: Xu, Tongda, et al.
Published: (2025)
by: Xu, Tongda, et al.
Published: (2025)
RopeBEV: A Multi-Camera Roadside Perception Network in Bird's-Eye-View
by: Jia, Jinrang, et al.
Published: (2024)
by: Jia, Jinrang, et al.
Published: (2024)
SCAFusion: A Multimodal 3D Detection Framework for Small Object Detection in Lunar Surface Exploration
by: Chen, Xin, et al.
Published: (2025)
by: Chen, Xin, et al.
Published: (2025)
Robust Drone-View Geo-Localization via Content-Viewpoint Disentanglement
by: Li, Ke, et al.
Published: (2025)
by: Li, Ke, et al.
Published: (2025)
PairDropGS: Paired Dropout-Induced Consistency Regularization for Sparse-View Gaussian Splatting
by: Li, Hantang, et al.
Published: (2026)
by: Li, Hantang, et al.
Published: (2026)
ActiView: Evaluating Active Perception Ability for Multimodal Large Language Models
by: Wang, Ziyue, et al.
Published: (2024)
by: Wang, Ziyue, et al.
Published: (2024)
Dynamic View Synthesis from Small Camera Motion Videos
by: Sun, Huiqiang, et al.
Published: (2025)
by: Sun, Huiqiang, et al.
Published: (2025)
An Extensible Framework for Open Heterogeneous Collaborative Perception
by: Lu, Yifan, et al.
Published: (2024)
by: Lu, Yifan, et al.
Published: (2024)
ViewFormer: Exploring Spatiotemporal Modeling for Multi-View 3D Occupancy Perception via View-Guided Transformers
by: Li, Jinke, et al.
Published: (2024)
by: Li, Jinke, et al.
Published: (2024)
A Prediction-as-Perception Framework for 3D Object Detection
by: Zhang, Song, et al.
Published: (2026)
by: Zhang, Song, et al.
Published: (2026)
SVIA: A Street View Image Anonymization Framework for Self-Driving Applications
by: Liu, Dongyu, et al.
Published: (2025)
by: Liu, Dongyu, et al.
Published: (2025)
Versatile Recompression-Aware Perceptual Image Super-Resolution
by: He, Mingwei, et al.
Published: (2025)
by: He, Mingwei, et al.
Published: (2025)
CoopDETR: A Unified Cooperative Perception Framework for 3D Detection via Object Query
by: Wang, Zhe, et al.
Published: (2025)
by: Wang, Zhe, et al.
Published: (2025)
Which2comm: An Efficient Collaborative Perception Framework for 3D Object Detection
by: Yu, Duanrui, et al.
Published: (2025)
by: Yu, Duanrui, et al.
Published: (2025)
A Global Depth-Range-Free Multi-View Stereo Transformer Network with Pose Embedding
by: Dong, Yitong, et al.
Published: (2024)
by: Dong, Yitong, et al.
Published: (2024)
VIP: Versatile Image Outpainting Empowered by Multimodal Large Language Model
by: Yang, Jinze, et al.
Published: (2024)
by: Yang, Jinze, et al.
Published: (2024)
Antidote: A Unified Framework for Mitigating LVLM Hallucinations in Counterfactual Presupposition and Object Perception
by: Wu, Yuanchen, et al.
Published: (2025)
by: Wu, Yuanchen, et al.
Published: (2025)
MultiEgo: A Multi-View Egocentric Video Dataset for 4D Scene Reconstruction
by: Li, Bate, et al.
Published: (2025)
by: Li, Bate, et al.
Published: (2025)
Enhancing Cross-View Geo-Localization Generalization via Global-Local Consistency and Geometric Equivariance
by: Wang, Xiaowei, et al.
Published: (2025)
by: Wang, Xiaowei, et al.
Published: (2025)
Empowering Vector Graphics with Consistently Arbitrary Viewing and View-dependent Visibility
by: Li, Yidi, et al.
Published: (2025)
by: Li, Yidi, et al.
Published: (2025)
CoT4Det: A Chain-of-Thought Framework for Perception-Oriented Vision-Language Tasks
by: Qi, Yu, et al.
Published: (2025)
by: Qi, Yu, et al.
Published: (2025)
WUTDet: A 100K-Scale Ship Detection Dataset and Benchmarks with Dense Small Objects
by: Liang, Junxiong, et al.
Published: (2026)
by: Liang, Junxiong, et al.
Published: (2026)
Urban Safety Perception Assessments via Integrating Multimodal Large Language Models with Street View Images
by: Zhang, Jiaxin, et al.
Published: (2024)
by: Zhang, Jiaxin, et al.
Published: (2024)
FrameNeRF: A Simple and Efficient Framework for Few-shot Novel View Synthesis
by: Xing, Yan, et al.
Published: (2024)
by: Xing, Yan, et al.
Published: (2024)
JointSplat: Probabilistic Joint Flow-Depth Optimization for Sparse-View Gaussian Splatting
by: Xiao, Yang, et al.
Published: (2025)
by: Xiao, Yang, et al.
Published: (2025)
Similar Items
-
Towards Stable 3D Object Detection
by: Wang, Jiabao, et al.
Published: (2024) -
Small but Mighty: Enhancing 3D Point Clouds Semantic Segmentation with U-Next Framework
by: Zeng, Ziyin, et al.
Published: (2023) -
OPUS: Occupancy Prediction Using a Sparse Set
by: Wang, Jiabao, et al.
Published: (2024) -
Small but Mighty: Dynamic Wavelet Expert-Guided Fine-Tuning of Large-Scale Models for Optical Remote Sensing Object Segmentation
by: Sun, Yanguang, et al.
Published: (2026) -
UV-M3TL: A Unified and Versatile Multimodal Multi-Task Learning Framework for Assistive Driving Perception
by: Liu, Wenzhuo, et al.
Published: (2026)