SceneParser: Hierarchical Scene Parsing for Visual Semantics Understanding
Fuente:
arXiv
Saved in:
| Main Authors: | Xu, Pengxin, Lin, Xincheng, Xiao, Luping, Jiang, Qing, Zhang, Meishan, Fei, Hao, Zhang, Shanghang, Chen, Xingyu |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
TB-HSU: Hierarchical 3D Scene Understanding with Contextual Affordances
by: Xu, Wenting, et al.
Published: (2024)
by: Xu, Wenting, et al.
Published: (2024)
Hierarchical Context Transformer for Multi-level Semantic Scene Understanding
by: Hao, Luoying, et al.
Published: (2025)
by: Hao, Luoying, et al.
Published: (2025)
Adaptive Visual Scene Understanding: Incremental Scene Graph Generation
by: Khandelwal, Naitik, et al.
Published: (2023)
by: Khandelwal, Naitik, et al.
Published: (2023)
RoomPilot: Controllable Indoor Scene Synthesis via Multimodal Semantic Parsing
by: Chen, Wentang, et al.
Published: (2025)
by: Chen, Wentang, et al.
Published: (2025)
Unsupervised Collaborative Domain Adaptation for Driving Scene Parsing
by: Fan, Jiahe, et al.
Published: (2026)
by: Fan, Jiahe, et al.
Published: (2026)
Text-Scene: A Scene-to-Language Parsing Framework for 3D Scene Understanding
by: Li, Haoyuan, et al.
Published: (2025)
by: Li, Haoyuan, et al.
Published: (2025)
NAUTILUS: A Large Multimodal Model for Underwater Scene Understanding
by: Xu, Wei, et al.
Published: (2025)
by: Xu, Wei, et al.
Published: (2025)
UniGround: Universal 3D Visual Grounding via Training-Free Scene Parsing
by: Zhang, Jiaxi, et al.
Published: (2026)
by: Zhang, Jiaxi, et al.
Published: (2026)
RoadFormer: Duplex Transformer for RGB-Normal Semantic Road Scene Parsing
by: Li, Jiahang, et al.
Published: (2023)
by: Li, Jiahang, et al.
Published: (2023)
CreatiParser: Generative Image Parsing of Raster Graphic Designs into Editable Layers
by: Chen, Weidong, et al.
Published: (2026)
by: Chen, Weidong, et al.
Published: (2026)
Towards Real-World Document Parsing via Realistic Scene Synthesis and Document-Aware Training
by: Li, Gengluo, et al.
Published: (2026)
by: Li, Gengluo, et al.
Published: (2026)
TextVidBench: A Benchmark for Long Video Scene Text Understanding
by: Zhong, Yangyang, et al.
Published: (2025)
by: Zhong, Yangyang, et al.
Published: (2025)
Learning 4D Panoptic Scene Graph Generation from Rich 2D Visual Scene
by: Wu, Shengqiong, et al.
Published: (2025)
by: Wu, Shengqiong, et al.
Published: (2025)
Traffic Scene Parsing through the TSP6K Dataset
by: Jiang, Peng-Tao, et al.
Published: (2023)
by: Jiang, Peng-Tao, et al.
Published: (2023)
FreeQ-Graph: Free-form Querying with Semantic Consistent Scene Graph for 3D Scene Understanding
by: Zhan, Chenlu, et al.
Published: (2025)
by: Zhan, Chenlu, et al.
Published: (2025)
DC-Scene: Data-Centric Learning for 3D Scene Understanding
by: Huang, Ting, et al.
Published: (2025)
by: Huang, Ting, et al.
Published: (2025)
AVS-Net: Point Sampling with Adaptive Voxel Size for 3D Scene Understanding
by: Yang, Hongcheng, et al.
Published: (2024)
by: Yang, Hongcheng, et al.
Published: (2024)
Semantic Gaussians: Open-Vocabulary Scene Understanding with 3D Gaussian Splatting
by: Guo, Jun, et al.
Published: (2024)
by: Guo, Jun, et al.
Published: (2024)
Hierarchical Temporal Context Learning for Camera-based Semantic Scene Completion
by: Li, Bohan, et al.
Published: (2024)
by: Li, Bohan, et al.
Published: (2024)
Open Vocabulary Semantic Scene Sketch Understanding
by: Bourouis, Ahmed, et al.
Published: (2023)
by: Bourouis, Ahmed, et al.
Published: (2023)
Training-Free Hierarchical Scene Understanding for Gaussian Splatting with Superpoint Graphs
by: Dai, Shaohui, et al.
Published: (2025)
by: Dai, Shaohui, et al.
Published: (2025)
Parser-Oriented Structural Refinement for a Stable Layout Interface in Document Parsing
by: Liu, Fuyuan, et al.
Published: (2026)
by: Liu, Fuyuan, et al.
Published: (2026)
LET-US: Long Event-Text Understanding of Scenes
by: Chen, Rui, et al.
Published: (2025)
by: Chen, Rui, et al.
Published: (2025)
Scene Understanding Enabled Semantic Communication with Open Channel Coding
by: Xiang, Zhe, et al.
Published: (2025)
by: Xiang, Zhe, et al.
Published: (2025)
UniParser: Multi-Human Parsing with Unified Correlation Representation Learning
by: Chu, Jiaming, et al.
Published: (2023)
by: Chu, Jiaming, et al.
Published: (2023)
S2GS: Streaming Semantic Gaussian Splatting for Online Scene Understanding and Reconstruction
by: Zhang, Renhe, et al.
Published: (2026)
by: Zhang, Renhe, et al.
Published: (2026)
Semantically-aware Neural Radiance Fields for Visual Scene Understanding: A Comprehensive Review
by: Nguyen, Thang-Anh-Quan, et al.
Published: (2024)
by: Nguyen, Thang-Anh-Quan, et al.
Published: (2024)
Neural Scene Designer: Self-Styled Semantic Image Manipulation
by: Lin, Jianman, et al.
Published: (2025)
by: Lin, Jianman, et al.
Published: (2025)
Unified Semantic Transformer for 3D Scene Understanding
by: Koch, Sebastian, et al.
Published: (2025)
by: Koch, Sebastian, et al.
Published: (2025)
SceneDecorator: Towards Scene-Oriented Story Generation with Scene Planning and Scene Consistency
by: Song, Quanjian, et al.
Published: (2025)
by: Song, Quanjian, et al.
Published: (2025)
PIG: Prompt Images Guidance for Night-Time Scene Parsing
by: Xie, Zhifeng, et al.
Published: (2024)
by: Xie, Zhifeng, et al.
Published: (2024)
HOSIG: Full-Body Human-Object-Scene Interaction Generation with Hierarchical Scene Perception
by: Yao, Wei, et al.
Published: (2025)
by: Yao, Wei, et al.
Published: (2025)
VEOcc: Voxel-Centric Online Semantic Occupancy Prediction For Embodied Scene Understanding
by: Wang, Ruoyu, et al.
Published: (2026)
by: Wang, Ruoyu, et al.
Published: (2026)
SceneDesigner: Controllable Multi-Object Image Generation with 9-DoF Pose Manipulation
by: Qin, Zhenyuan, et al.
Published: (2025)
by: Qin, Zhenyuan, et al.
Published: (2025)
HexPlane Representation for 3D Semantic Scene Understanding
by: Chen, Zeren, et al.
Published: (2025)
by: Chen, Zeren, et al.
Published: (2025)
Chat-Scene: Bridging 3D Scene and Large Language Models with Object Identifiers
by: Huang, Haifeng, et al.
Published: (2023)
by: Huang, Haifeng, et al.
Published: (2023)
InstructScene: Instruction-Driven 3D Indoor Scene Synthesis with Semantic Graph Prior
by: Lin, Chenguo, et al.
Published: (2024)
by: Lin, Chenguo, et al.
Published: (2024)
Sparse3DPR: Training-Free 3D Hierarchical Scene Parsing and Task-Adaptive Subgraph Reasoning from Sparse RGB Views
by: Feng, Haida, et al.
Published: (2025)
by: Feng, Haida, et al.
Published: (2025)
Universal Scene Graph Generation
by: Wu, Shengqiong, et al.
Published: (2025)
by: Wu, Shengqiong, et al.
Published: (2025)
SemanticNVS: Improving Semantic Scene Understanding in Generative Novel View Synthesis
by: Chen, Xinya, et al.
Published: (2026)
by: Chen, Xinya, et al.
Published: (2026)
Similar Items
-
TB-HSU: Hierarchical 3D Scene Understanding with Contextual Affordances
by: Xu, Wenting, et al.
Published: (2024) -
Hierarchical Context Transformer for Multi-level Semantic Scene Understanding
by: Hao, Luoying, et al.
Published: (2025) -
Adaptive Visual Scene Understanding: Incremental Scene Graph Generation
by: Khandelwal, Naitik, et al.
Published: (2023) -
RoomPilot: Controllable Indoor Scene Synthesis via Multimodal Semantic Parsing
by: Chen, Wentang, et al.
Published: (2025) -
Unsupervised Collaborative Domain Adaptation for Driving Scene Parsing
by: Fan, Jiahe, et al.
Published: (2026)