Open Panoramic Segmentation
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Zheng, Junwei, Liu, Ruiping, Chen, Yufan, Peng, Kunyu, Wu, Chengzhi, Yang, Kailun, Zhang, Jiaming, Stiefelhagen, Rainer |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
EgoExoMem: Cross-View Memory Reasoning over Synchronized Egocentric and Exocentric Videos
von: Liu, Ruiping, et al.
Veröffentlicht: (2026)
von: Liu, Ruiping, et al.
Veröffentlicht: (2026)
What if? Emulative Simulation with World Models for Situated Reasoning
von: Liu, Ruiping, et al.
Veröffentlicht: (2026)
von: Liu, Ruiping, et al.
Veröffentlicht: (2026)
Situat3DChange: Situated 3D Change Understanding Dataset for Multimodal Large Language Model
von: Liu, Ruiping, et al.
Veröffentlicht: (2025)
von: Liu, Ruiping, et al.
Veröffentlicht: (2025)
MateRobot: Material Recognition in Wearable Robotics for People with Visual Impairments
von: Zheng, Junwei, et al.
Veröffentlicht: (2023)
von: Zheng, Junwei, et al.
Veröffentlicht: (2023)
RoDLA: Benchmarking the Robustness of Document Layout Analysis Models
von: Chen, Yufan, et al.
Veröffentlicht: (2024)
von: Chen, Yufan, et al.
Veröffentlicht: (2024)
HybriDLA: Hybrid Generation for Document Layout Analysis
von: Chen, Yufan, et al.
Veröffentlicht: (2025)
von: Chen, Yufan, et al.
Veröffentlicht: (2025)
Graph-based Document Structure Analysis
von: Chen, Yufan, et al.
Veröffentlicht: (2025)
von: Chen, Yufan, et al.
Veröffentlicht: (2025)
Scene-agnostic Pose Regression for Visual Localization
von: Zheng, Junwei, et al.
Veröffentlicht: (2025)
von: Zheng, Junwei, et al.
Veröffentlicht: (2025)
Fourier Prompt Tuning for Modality-Incomplete Scene Segmentation
von: Liu, Ruiping, et al.
Veröffentlicht: (2024)
von: Liu, Ruiping, et al.
Veröffentlicht: (2024)
RHO: Robust Holistic OSM-Based Metric Cross-View Geo-Localization
von: Zheng, Junwei, et al.
Veröffentlicht: (2026)
von: Zheng, Junwei, et al.
Veröffentlicht: (2026)
OneBEV: Using One Panoramic Image for Bird's-Eye-View Semantic Mapping
von: Wei, Jiale, et al.
Veröffentlicht: (2024)
von: Wei, Jiale, et al.
Veröffentlicht: (2024)
Elevating Skeleton-Based Action Recognition with Efficient Multi-Modality Self-Supervision
von: Wei, Yiping, et al.
Veröffentlicht: (2023)
von: Wei, Yiping, et al.
Veröffentlicht: (2023)
Skeleton-Based Human Action Recognition with Noisy Labels
von: Xu, Yi, et al.
Veröffentlicht: (2024)
von: Xu, Yi, et al.
Veröffentlicht: (2024)
Exploring Self-supervised Skeleton-based Action Recognition in Occluded Environments
von: Chen, Yifei, et al.
Veröffentlicht: (2023)
von: Chen, Yifei, et al.
Veröffentlicht: (2023)
Towards Multi-Source Domain Generalization for Sleep Staging with Noisy Labels
von: Wang, Kening, et al.
Veröffentlicht: (2026)
von: Wang, Kening, et al.
Veröffentlicht: (2026)
TransKD: Transformer Knowledge Distillation for Efficient Semantic Segmentation
von: Liu, Ruiping, et al.
Veröffentlicht: (2022)
von: Liu, Ruiping, et al.
Veröffentlicht: (2022)
MICA: Multi-Agent Industrial Coordination Assistant
von: Wen, Di, et al.
Veröffentlicht: (2025)
von: Wen, Di, et al.
Veröffentlicht: (2025)
DriveXQA: Cross-modal Visual Question Answering for Adverse Driving Scene Understanding
von: Tao, Mingzhe, et al.
Veröffentlicht: (2026)
von: Tao, Mingzhe, et al.
Veröffentlicht: (2026)
SGR3 Model: Scene Graph Retrieval-Reasoning Model in 3D
von: Wang, Zirui, et al.
Veröffentlicht: (2026)
von: Wang, Zirui, et al.
Veröffentlicht: (2026)
Comb, Prune, Distill: Towards Unified Pruning for Vision Model Compression
von: Schmitt, Jonas, et al.
Veröffentlicht: (2024)
von: Schmitt, Jonas, et al.
Veröffentlicht: (2024)
Exploring Video-Based Driver Activity Recognition under Noisy Labels
von: Fan, Linjuan, et al.
Veröffentlicht: (2025)
von: Fan, Linjuan, et al.
Veröffentlicht: (2025)
OAFuser: Towards Omni-Aperture Fusion for Light Field Semantic Segmentation
von: Teng, Fei, et al.
Veröffentlicht: (2023)
von: Teng, Fei, et al.
Veröffentlicht: (2023)
Referring Atomic Video Action Recognition
von: Peng, Kunyu, et al.
Veröffentlicht: (2024)
von: Peng, Kunyu, et al.
Veröffentlicht: (2024)
ObjectFinder: An Open-Vocabulary Assistive System for Interactive Object Search by Blind People
von: Liu, Ruiping, et al.
Veröffentlicht: (2024)
von: Liu, Ruiping, et al.
Veröffentlicht: (2024)
Seeing Beyond: Extrapolative Domain Adaptive Panoramic Segmentation
von: Zheng, Yuanfan, et al.
Veröffentlicht: (2026)
von: Zheng, Yuanfan, et al.
Veröffentlicht: (2026)
Behind Every Domain There is a Shift: Adapting Distortion-aware Vision Transformers for Panoramic Semantic Segmentation
von: Zhang, Jiaming, et al.
Veröffentlicht: (2022)
von: Zhang, Jiaming, et al.
Veröffentlicht: (2022)
EReLiFM: Evidential Reliability-Aware Residual Flow Meta-Learning for Open-Set Domain Generalization under Noisy Labels
von: Peng, Kunyu, et al.
Veröffentlicht: (2025)
von: Peng, Kunyu, et al.
Veröffentlicht: (2025)
Occlusion-Aware Seamless Segmentation
von: Cao, Yihong, et al.
Veröffentlicht: (2024)
von: Cao, Yihong, et al.
Veröffentlicht: (2024)
Deformable Mamba for Wide Field of View Segmentation
von: Hu, Jie, et al.
Veröffentlicht: (2024)
von: Hu, Jie, et al.
Veröffentlicht: (2024)
RoHOI: Robustness Benchmark for Human-Object Interaction Detection
von: Wen, Di, et al.
Veröffentlicht: (2025)
von: Wen, Di, et al.
Veröffentlicht: (2025)
Mitigating Label Noise using Prompt-Based Hyperbolic Meta-Learning in Open-Set Domain Generalization
von: Peng, Kunyu, et al.
Veröffentlicht: (2024)
von: Peng, Kunyu, et al.
Veröffentlicht: (2024)
@Bench: Benchmarking Vision-Language Models for Human-centered Assistive Technology
von: Jiang, Xin, et al.
Veröffentlicht: (2024)
von: Jiang, Xin, et al.
Veröffentlicht: (2024)
HopaDIFF: Holistic-Partial Aware Fourier Conditioned Diffusion for Referring Human Action Segmentation in Multi-Person Scenarios
von: Peng, Kunyu, et al.
Veröffentlicht: (2025)
von: Peng, Kunyu, et al.
Veröffentlicht: (2025)
mmWalk: Towards Multi-modal Multi-view Walking Assistance
von: Ying, Kedi, et al.
Veröffentlicht: (2025)
von: Ying, Kedi, et al.
Veröffentlicht: (2025)
More than the Sum: Panorama-Language Models for Adverse Omni-Scenes
von: Fan, Weijia, et al.
Veröffentlicht: (2026)
von: Fan, Weijia, et al.
Veröffentlicht: (2026)
Advancing Open-Set Domain Generalization Using Evidential Bi-Level Hardest Domain Scheduler
von: Peng, Kunyu, et al.
Veröffentlicht: (2024)
von: Peng, Kunyu, et al.
Veröffentlicht: (2024)
$M^2$-Occ: Resilient 3D Semantic Occupancy Prediction for Autonomous Driving with Incomplete Camera Inputs
von: Lin, Kaixin, et al.
Veröffentlicht: (2026)
von: Lin, Kaixin, et al.
Veröffentlicht: (2026)
CHAOS: Chart Analysis with Outlier Samples
von: Moured, Omar, et al.
Veröffentlicht: (2025)
von: Moured, Omar, et al.
Veröffentlicht: (2025)
Go Beyond Earth: Understanding Human Actions and Scenes in Microgravity Environments
von: Wen, Di, et al.
Veröffentlicht: (2025)
von: Wen, Di, et al.
Veröffentlicht: (2025)
RefAtomNet++: Advancing Referring Atomic Video Action Recognition using Semantic Retrieval based Multi-Trajectory Mamba
von: Peng, Kunyu, et al.
Veröffentlicht: (2025)
von: Peng, Kunyu, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
EgoExoMem: Cross-View Memory Reasoning over Synchronized Egocentric and Exocentric Videos
von: Liu, Ruiping, et al.
Veröffentlicht: (2026) -
What if? Emulative Simulation with World Models for Situated Reasoning
von: Liu, Ruiping, et al.
Veröffentlicht: (2026) -
Situat3DChange: Situated 3D Change Understanding Dataset for Multimodal Large Language Model
von: Liu, Ruiping, et al.
Veröffentlicht: (2025) -
MateRobot: Material Recognition in Wearable Robotics for People with Visual Impairments
von: Zheng, Junwei, et al.
Veröffentlicht: (2023) -
RoDLA: Benchmarking the Robustness of Document Layout Analysis Models
von: Chen, Yufan, et al.
Veröffentlicht: (2024)