MTPano: Multi-Task Panoramic Scene Understanding via Label-Free Integration of Dense Prediction Priors
Fuente:
arXiv
Saved in:
| Main Authors: | Zhang, Jingdong, Zhan, Xiaohang, Zhang, Lingzhi, Wang, Yizhou, Yu, Zhengming, Wang, Jionghao, Wang, Wenping, Li, Xin |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Multi-Task Label Discovery via Hierarchical Task Tokens for Partially Annotated Dense Predictions
by: Zhang, Jingdong, et al.
Published: (2024)
by: Zhang, Jingdong, et al.
Published: (2024)
Beyond Thinking: Imagining in 360$^\circ$ for Humanoid Visual Search
by: Zhang, Jingdong, et al.
Published: (2026)
by: Zhang, Jingdong, et al.
Published: (2026)
Learning What Matters: Adaptive Information-Theoretic Objectives for Robot Exploration
by: Yu, Youwei, et al.
Published: (2026)
by: Yu, Youwei, et al.
Published: (2026)
SPGen: Spherical Projection as Consistent and Flexible Representation for Single Image 3D Shape Generation
by: Zhang, Jingdong, et al.
Published: (2025)
by: Zhang, Jingdong, et al.
Published: (2025)
FreeQ-Graph: Free-form Querying with Semantic Consistent Scene Graph for 3D Scene Understanding
by: Zhan, Chenlu, et al.
Published: (2025)
by: Zhan, Chenlu, et al.
Published: (2025)
UniSER: A Foundation Model for Unified Soft Effects Removal
by: Zhang, Jingdong, et al.
Published: (2025)
by: Zhang, Jingdong, et al.
Published: (2025)
On Multi-Step Theorem Prediction via Non-Parametric Structural Priors
by: Zhao, Junbo, et al.
Published: (2026)
by: Zhao, Junbo, et al.
Published: (2026)
MTMamba: Enhancing Multi-Task Dense Scene Understanding by Mamba-Based Decoders
by: Lin, Baijiong, et al.
Published: (2024)
by: Lin, Baijiong, et al.
Published: (2024)
DenseScan: Advancing 3D Scene Understanding with 2D Dense Annotation
by: Wang, Zirui, et al.
Published: (2025)
by: Wang, Zirui, et al.
Published: (2025)
Spatial As Deep: Spatial CNN for Traffic Scene Understanding
by: Pan, Xingang, et al.
Published: (2017)
by: Pan, Xingang, et al.
Published: (2017)
Disentangled Clothed Avatar Generation from Text Descriptions
by: Wang, Jionghao, et al.
Published: (2023)
by: Wang, Jionghao, et al.
Published: (2023)
3D-Aware Multi-Task Learning with Cross-View Correlations for Dense Scene Understanding
by: Wang, Xiaoye, et al.
Published: (2025)
by: Wang, Xiaoye, et al.
Published: (2025)
MTMamba++: Enhancing Multi-Task Dense Scene Understanding via Mamba-Based Decoders
by: Lin, Baijiong, et al.
Published: (2024)
by: Lin, Baijiong, et al.
Published: (2024)
LaRender: Training-Free Occlusion Control in Image Generation via Latent Rendering
by: Zhan, Xiaohang, et al.
Published: (2025)
by: Zhan, Xiaohang, et al.
Published: (2025)
SolidGS: Consolidating Gaussian Surfel Splatting for Sparse-View Surface Reconstruction
by: Shen, Zhuowen, et al.
Published: (2024)
by: Shen, Zhuowen, et al.
Published: (2024)
3R-GS: Best Practice in Optimizing Camera Poses Along with 3DGS
by: Huang, Zhisheng, et al.
Published: (2025)
by: Huang, Zhisheng, et al.
Published: (2025)
FastScene: Text-Driven Fast 3D Indoor Scene Generation via Panoramic Gaussian Splatting
by: Ma, Yikun, et al.
Published: (2024)
by: Ma, Yikun, et al.
Published: (2024)
BridgeNet: Comprehensive and Effective Feature Interactions via Bridge Feature for Multi-task Dense Predictions
by: Zhang, Jingdong, et al.
Published: (2023)
by: Zhang, Jingdong, et al.
Published: (2023)
From Sparse to Dense: Multi-View GRPO for Flow Models via Augmented Condition Space
by: Bu, Jiazi, et al.
Published: (2026)
by: Bu, Jiazi, et al.
Published: (2026)
Multi-Task Dense Prediction via Mixture of Low-Rank Experts
by: Yang, Yuqi, et al.
Published: (2024)
by: Yang, Yuqi, et al.
Published: (2024)
Indoor Scene Reconstruction with Fine-Grained Details Using Hybrid Representation and Normal Prior Enhancement
by: Ye, Sheng, et al.
Published: (2023)
by: Ye, Sheng, et al.
Published: (2023)
PRIMA: Boosting Animal Mesh Recovery with Biological Priors and Test-Time Adaptation
by: Yu, Xiaohang, et al.
Published: (2026)
by: Yu, Xiaohang, et al.
Published: (2026)
HomoMatcher: Dense Feature Matching Results with Semi-Dense Efficiency by Homography Estimation
by: Wang, Xiaolong, et al.
Published: (2024)
by: Wang, Xiaolong, et al.
Published: (2024)
Cross-Task Affinity Learning for Multitask Dense Scene Predictions
by: Sinodinos, Dimitrios, et al.
Published: (2024)
by: Sinodinos, Dimitrios, et al.
Published: (2024)
R3DS: Reality-linked 3D Scenes for Panoramic Scene Understanding
by: Wu, Qirui, et al.
Published: (2024)
by: Wu, Qirui, et al.
Published: (2024)
PanopticNeRF-360: Panoramic 3D-to-2D Label Transfer in Urban Scenes
by: Fu, Xiao, et al.
Published: (2023)
by: Fu, Xiao, et al.
Published: (2023)
SP-SLAM: Neural Real-Time Dense SLAM With Scene Priors
by: Hong, Zhen, et al.
Published: (2025)
by: Hong, Zhen, et al.
Published: (2025)
MagicGeo: Training-Free Text-Guided Geometric Diagram Generation
by: Wang, Junxiao, et al.
Published: (2025)
by: Wang, Junxiao, et al.
Published: (2025)
Can Class-Priors Help Single-Positive Multi-Label Learning?
by: Liu, Biao, et al.
Published: (2023)
by: Liu, Biao, et al.
Published: (2023)
Controllable 3D Outdoor Scene Generation via Scene Graphs
by: Liu, Yuheng, et al.
Published: (2025)
by: Liu, Yuheng, et al.
Published: (2025)
A Vanilla Multi-Task Framework for Dense Visual Prediction Solution to 1st VCL Challenge -- Multi-Task Robustness Track
by: Chen, Zehui, et al.
Published: (2024)
by: Chen, Zehui, et al.
Published: (2024)
Dense Connector for MLLMs
by: Yao, Huanjin, et al.
Published: (2024)
by: Yao, Huanjin, et al.
Published: (2024)
Panoramic Affordance Prediction
by: Zhang, Zixin, et al.
Published: (2026)
by: Zhang, Zixin, et al.
Published: (2026)
Enhancing Mamba Decoder with Bidirectional Interaction in Multi-Task Dense Prediction
by: Cao, Mang, et al.
Published: (2025)
by: Cao, Mang, et al.
Published: (2025)
EdgeCrafter: Compact ViTs for Edge Dense Prediction via Task-Specialized Distillation
by: Liu, Longfei, et al.
Published: (2026)
by: Liu, Longfei, et al.
Published: (2026)
MMGen: Unified Multi-modal Image Generation and Understanding in One Go
by: Wang, Jiepeng, et al.
Published: (2025)
by: Wang, Jiepeng, et al.
Published: (2025)
OmniX: From Unified Panoramic Generation and Perception to Graphics-Ready 3D Scenes
by: Huang, Yukun, et al.
Published: (2025)
by: Huang, Yukun, et al.
Published: (2025)
Exploring Partial Multi-Label Learning via Integrating Semantic Co-occurrence Knowledge
by: Wu, Xin, et al.
Published: (2025)
by: Wu, Xin, et al.
Published: (2025)
Quantum Probabilistic Label Refining: Enhancing Label Quality for Robust Image Classification
by: Qi, Fang, et al.
Published: (2025)
by: Qi, Fang, et al.
Published: (2025)
Language-Augmented Symbolic Planner for Open-World Task Planning
by: Chen, Guanqi, et al.
Published: (2024)
by: Chen, Guanqi, et al.
Published: (2024)
Similar Items
-
Multi-Task Label Discovery via Hierarchical Task Tokens for Partially Annotated Dense Predictions
by: Zhang, Jingdong, et al.
Published: (2024) -
Beyond Thinking: Imagining in 360$^\circ$ for Humanoid Visual Search
by: Zhang, Jingdong, et al.
Published: (2026) -
Learning What Matters: Adaptive Information-Theoretic Objectives for Robot Exploration
by: Yu, Youwei, et al.
Published: (2026) -
SPGen: Spherical Projection as Consistent and Flexible Representation for Single Image 3D Shape Generation
by: Zhang, Jingdong, et al.
Published: (2025) -
FreeQ-Graph: Free-form Querying with Semantic Consistent Scene Graph for 3D Scene Understanding
by: Zhan, Chenlu, et al.
Published: (2025)