Dense360: Dense Understanding from Omnidirectional Panoramas
Fuente:
arXiv
Saved in:
| Main Authors: | Zhou, Yikang, Zhang, Tao, Zhang, Dizhe, Ji, Shunping, Li, Xiangtai, Qi, Lu |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
DVIS-DAQ: Improving Video Segmentation via Dynamic Anchor Queries
by: Zhou, Yikang, et al.
Published: (2024)
by: Zhou, Yikang, et al.
Published: (2024)
DiT360: High-Fidelity Panoramic Image Generation via Hybrid Training
by: Feng, Haoran, et al.
Published: (2025)
by: Feng, Haoran, et al.
Published: (2025)
Sa2VA: Marrying SAM2 with LLaVA for Dense Grounded Understanding of Images and Videos
by: Yuan, Haobo, et al.
Published: (2025)
by: Yuan, Haobo, et al.
Published: (2025)
DenseWorld-1M: Towards Detailed Dense Grounded Caption in the Real World
by: Li, Xiangtai, et al.
Published: (2025)
by: Li, Xiangtai, et al.
Published: (2025)
Are They the Same? Exploring Visual Correspondence Shortcomings of Multimodal LLMs
by: Zhou, Yikang, et al.
Published: (2025)
by: Zhou, Yikang, et al.
Published: (2025)
The 1st Solution for 7th LSVOS RVOS Track: SaSaSa2VA
by: Niu, Quanzhu, et al.
Published: (2025)
by: Niu, Quanzhu, et al.
Published: (2025)
Beyond Appearance: Geometric Cues for Robust Video Instance Segmentation
by: Niu, Quanzhu, et al.
Published: (2025)
by: Niu, Quanzhu, et al.
Published: (2025)
Unified Dense Prediction of Video Diffusion
by: Yang, Lehan, et al.
Published: (2025)
by: Yang, Lehan, et al.
Published: (2025)
Point Cloud Mamba: Point Cloud Learning via State Space Model
by: Zhang, Tao, et al.
Published: (2024)
by: Zhang, Tao, et al.
Published: (2024)
P2PFormer: A Primitive-to-polygon Method for Regular Building Contour Extraction from Remote Sensing Images
by: Zhang, Tao, et al.
Published: (2024)
by: Zhang, Tao, et al.
Published: (2024)
D$^2$GS: Depth-and-Density Guided Gaussian Splatting for Stable and Accurate Sparse-View Reconstruction
by: Song, Meixi, et al.
Published: (2025)
by: Song, Meixi, et al.
Published: (2025)
Pixel-SAIL: Single Transformer For Pixel-Grounded Understanding
by: Zhang, Tao, et al.
Published: (2025)
by: Zhang, Tao, et al.
Published: (2025)
OMG-LLaVA: Bridging Image-level, Object-level, Pixel-level Reasoning and Understanding
by: Zhang, Tao, et al.
Published: (2024)
by: Zhang, Tao, et al.
Published: (2024)
DenseScan: Advancing 3D Scene Understanding with 2D Dense Annotation
by: Wang, Zirui, et al.
Published: (2025)
by: Wang, Zirui, et al.
Published: (2025)
SaSaSaSa2VA: 2nd Place of the 5th PVUW MeViS-Text Track
by: Gong, Dengxian, et al.
Published: (2026)
by: Gong, Dengxian, et al.
Published: (2026)
Omnidirectional Spatial Modeling from Correlated Panoramas
by: Zhang, Xinshen, et al.
Published: (2025)
by: Zhang, Xinshen, et al.
Published: (2025)
Depth Any Panoramas: A Foundation Model for Panoramic Depth Estimation
by: Lin, Xin, et al.
Published: (2025)
by: Lin, Xin, et al.
Published: (2025)
CLIPSelf: Vision Transformer Distills Itself for Open-Vocabulary Dense Prediction
by: Wu, Size, et al.
Published: (2023)
by: Wu, Size, et al.
Published: (2023)
BiDense: Binarization for Dense Prediction
by: Yin, Rui, et al.
Published: (2024)
by: Yin, Rui, et al.
Published: (2024)
One Flight Over the Gap: A Survey from Perspective to Panoramic Vision
by: Lin, Xin, et al.
Published: (2025)
by: Lin, Xin, et al.
Published: (2025)
SAMTok: Representing Any Mask with Two Words
by: Zhou, Yikang, et al.
Published: (2026)
by: Zhou, Yikang, et al.
Published: (2026)
360DVD: Controllable Panorama Video Generation with 360-Degree Video Diffusion Model
by: Wang, Qian, et al.
Published: (2024)
by: Wang, Qian, et al.
Published: (2024)
AirSim360: A Panoramic Simulation Platform within Drone View
by: Ge, Xian, et al.
Published: (2025)
by: Ge, Xian, et al.
Published: (2025)
Dense Multimodal Alignment for Open-Vocabulary 3D Scene Understanding
by: Li, Ruihuang, et al.
Published: (2024)
by: Li, Ruihuang, et al.
Published: (2024)
MovieChat: From Dense Token to Sparse Memory for Long Video Understanding
by: Song, Enxin, et al.
Published: (2023)
by: Song, Enxin, et al.
Published: (2023)
TUNA: Comprehensive Fine-grained Temporal Understanding Evaluation on Dense Dynamic Videos
by: Kong, Fanheng, et al.
Published: (2025)
by: Kong, Fanheng, et al.
Published: (2025)
Robust and Flexible Omnidirectional Depth Estimation with Multiple 360-degree Cameras
by: Li, Ming, et al.
Published: (2024)
by: Li, Ming, et al.
Published: (2024)
SE360: Semantic Edit in 360$^\circ$ Panoramas via Hierarchical Data Construction
by: Zhong, Haoyi, et al.
Published: (2025)
by: Zhong, Haoyi, et al.
Published: (2025)
Pose-Free Omnidirectional Gaussian Splatting for 360-Degree Videos with Consistent Depth Priors
by: Zhuang, Chuanqing, et al.
Published: (2026)
by: Zhuang, Chuanqing, et al.
Published: (2026)
Dense Matchers for Dense Tracking
by: Jelínek, Tomáš, et al.
Published: (2024)
by: Jelínek, Tomáš, et al.
Published: (2024)
What Makes for Text to 360-degree Panorama Generation with Stable Diffusion?
by: Ni, Jinhong, et al.
Published: (2025)
by: Ni, Jinhong, et al.
Published: (2025)
Vision Transformers: From Semantic Segmentation to Dense Prediction
by: Zhang, Li, et al.
Published: (2022)
by: Zhang, Li, et al.
Published: (2022)
DPBridge: Latent Diffusion Bridge for Dense Prediction
by: Ji, Haorui, et al.
Published: (2024)
by: Ji, Haorui, et al.
Published: (2024)
Towards Omnidirectional Reasoning with 360-R1: A Dataset, Benchmark, and GRPO-based Method
by: Zhang, Xinshen, et al.
Published: (2025)
by: Zhang, Xinshen, et al.
Published: (2025)
360VOTS: Visual Object Tracking and Segmentation in Omnidirectional Videos
by: Xu, Yinzhe, et al.
Published: (2024)
by: Xu, Yinzhe, et al.
Published: (2024)
DeH4R: A Decoupled and Hybrid Method for Road Network Graph Extraction
by: Gong, Dengxian, et al.
Published: (2025)
by: Gong, Dengxian, et al.
Published: (2025)
Parameter Aware Mamba Model for Multi-task Dense Prediction
by: Yu, Xinzhuo, et al.
Published: (2025)
by: Yu, Xinzhuo, et al.
Published: (2025)
Seam360GS: Seamless 360° Gaussian Splatting from Real-World Omnidirectional Images
by: Shin, Changha, et al.
Published: (2025)
by: Shin, Changha, et al.
Published: (2025)
Dense Vision Transformer Compression with Few Samples
by: Zhang, Hanxiao, et al.
Published: (2024)
by: Zhang, Hanxiao, et al.
Published: (2024)
HomoMatcher: Dense Feature Matching Results with Semi-Dense Efficiency by Homography Estimation
by: Wang, Xiaolong, et al.
Published: (2024)
by: Wang, Xiaolong, et al.
Published: (2024)
Similar Items
-
DVIS-DAQ: Improving Video Segmentation via Dynamic Anchor Queries
by: Zhou, Yikang, et al.
Published: (2024) -
DiT360: High-Fidelity Panoramic Image Generation via Hybrid Training
by: Feng, Haoran, et al.
Published: (2025) -
Sa2VA: Marrying SAM2 with LLaVA for Dense Grounded Understanding of Images and Videos
by: Yuan, Haobo, et al.
Published: (2025) -
DenseWorld-1M: Towards Detailed Dense Grounded Caption in the Real World
by: Li, Xiangtai, et al.
Published: (2025) -
Are They the Same? Exploring Visual Correspondence Shortcomings of Multimodal LLMs
by: Zhou, Yikang, et al.
Published: (2025)