RoadFormer+: Delivering RGB-X Scene Parsing through Scale-Aware Information Decoupling and Advanced Heterogeneous Feature Fusion
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Huang, Jianxin, Li, Jiahang, Jia, Ning, Sun, Yuxiang, Liu, Chengju, Chen, Qijun, Fan, Rui |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
RoadFormer: Duplex Transformer for RGB-Normal Semantic Road Scene Parsing
von: Li, Jiahang, et al.
Veröffentlicht: (2023)
von: Li, Jiahang, et al.
Veröffentlicht: (2023)
HAPNet: Toward Superior RGB-Thermal Scene Parsing via Hybrid, Asymmetric, and Progressive Heterogeneous Feature Fusion
von: Li, Jiahang, et al.
Veröffentlicht: (2024)
von: Li, Jiahang, et al.
Veröffentlicht: (2024)
DepthMatch: Semi-Supervised RGB-D Scene Parsing through Depth-Guided Regularization
von: Huang, Jianxin, et al.
Veröffentlicht: (2025)
von: Huang, Jianxin, et al.
Veröffentlicht: (2025)
RoadFormer : Local-Global Feature Fusion for Road Surface Classification in Autonomous Driving
von: Wang, Tianze, et al.
Veröffentlicht: (2025)
von: Wang, Tianze, et al.
Veröffentlicht: (2025)
SNE-RoadSegV2: Advancing Heterogeneous Feature Fusion and Fallibility Awareness for Freespace Detection
von: Feng, Yi, et al.
Veröffentlicht: (2024)
von: Feng, Yi, et al.
Veröffentlicht: (2024)
Fully Exploiting Vision Foundation Model's Profound Prior Knowledge for Generalizable RGB-Depth Driving Scene Parsing
von: Guo, Sicen, et al.
Veröffentlicht: (2025)
von: Guo, Sicen, et al.
Veröffentlicht: (2025)
NavComposer: Composing Language Instructions for Navigation Trajectories through Action-Scene-Object Modularization
von: He, Zongtao, et al.
Veröffentlicht: (2025)
von: He, Zongtao, et al.
Veröffentlicht: (2025)
A Saliency Enhanced Feature Fusion based multiscale RGB-D Salient Object Detection Network
von: Huang, Rui, et al.
Veröffentlicht: (2024)
von: Huang, Rui, et al.
Veröffentlicht: (2024)
TransPose: 6D Object Pose Estimation with Geometry-Aware Transformer
von: Lin, Xiao, et al.
Veröffentlicht: (2023)
von: Lin, Xiao, et al.
Veröffentlicht: (2023)
Kinematics-Aware Multi-Policy Reinforcement Learning for Force-Capable Humanoid Loco-Manipulation
von: Xiao, Kaiyan, et al.
Veröffentlicht: (2025)
von: Xiao, Kaiyan, et al.
Veröffentlicht: (2025)
DiffPixelFormer: Differential Pixel-Aware Transformer for RGB-D Indoor Scene Segmentation
von: Gong, Yan, et al.
Veröffentlicht: (2025)
von: Gong, Yan, et al.
Veröffentlicht: (2025)
Text-Scene: A Scene-to-Language Parsing Framework for 3D Scene Understanding
von: Li, Haoyuan, et al.
Veröffentlicht: (2025)
von: Li, Haoyuan, et al.
Veröffentlicht: (2025)
Towards Balanced RGB-TSDF Fusion for Consistent Semantic Scene Completion by 3D RGB Feature Completion and a Classwise Entropy Loss Function
von: Ding, Laiyan, et al.
Veröffentlicht: (2024)
von: Ding, Laiyan, et al.
Veröffentlicht: (2024)
RTFDNet: Fusion-Decoupling for Robust RGB-T Segmentation
von: Tan, Kunyu, et al.
Veröffentlicht: (2026)
von: Tan, Kunyu, et al.
Veröffentlicht: (2026)
Unsupervised Collaborative Domain Adaptation for Driving Scene Parsing
von: Fan, Jiahe, et al.
Veröffentlicht: (2026)
von: Fan, Jiahe, et al.
Veröffentlicht: (2026)
MSRFormer: Road Network Representation Learning using Multi-scale Feature Fusion of Heterogeneous Spatial Interactions
von: Yang, Jian, et al.
Veröffentlicht: (2025)
von: Yang, Jian, et al.
Veröffentlicht: (2025)
UDTIRI: An Online Open-Source Intelligent Road Inspection Benchmark Suite
von: Guo, Sicen, et al.
Veröffentlicht: (2023)
von: Guo, Sicen, et al.
Veröffentlicht: (2023)
PASTS: Progress-Aware Spatio-Temporal Transformer Speaker For Vision-and-Language Navigation
von: Wang, Liuyi, et al.
Veröffentlicht: (2023)
von: Wang, Liuyi, et al.
Veröffentlicht: (2023)
A Dual Semantic-Aware Recurrent Global-Adaptive Network For Vision-and-Language Navigation
von: Wang, Liuyi, et al.
Veröffentlicht: (2023)
von: Wang, Liuyi, et al.
Veröffentlicht: (2023)
IRFusionFormer: Enhancing Pavement Crack Segmentation with RGB-T Fusion and Topological-Based Loss
von: Xiao, Ruiqiang, et al.
Veröffentlicht: (2024)
von: Xiao, Ruiqiang, et al.
Veröffentlicht: (2024)
Modality-Decoupled RGB-Thermal Object Detector via Query Fusion
von: Tian, Chao, et al.
Veröffentlicht: (2026)
von: Tian, Chao, et al.
Veröffentlicht: (2026)
Unleashing the Power of Motion and Depth: A Selective Fusion Strategy for RGB-D Video Salient Object Detection
von: He, Jiahao, et al.
Veröffentlicht: (2025)
von: He, Jiahao, et al.
Veröffentlicht: (2025)
Adaptive Denoising-Enhanced LiDAR Odometry for Degeneration Resilience in Diverse Terrains
von: Ji, Mazeyu, et al.
Veröffentlicht: (2023)
von: Ji, Mazeyu, et al.
Veröffentlicht: (2023)
Fine-Grained Spatiotemporal Motion Alignment for Contrastive Video Representation Learning
von: Zhu, Minghao, et al.
Veröffentlicht: (2023)
von: Zhu, Minghao, et al.
Veröffentlicht: (2023)
WoundFormer: Multi-Scale Spatial Feature Fusion for Multi-Class Wound Tissue Segmentation
von: Kabir, Muhammad Ashad, et al.
Veröffentlicht: (2026)
von: Kabir, Muhammad Ashad, et al.
Veröffentlicht: (2026)
A Comprehensive Survey on Video Scene Parsing:Advances, Challenges, and Prospects
von: Xie, Guohuan, et al.
Veröffentlicht: (2025)
von: Xie, Guohuan, et al.
Veröffentlicht: (2025)
Traffic Scene Parsing through the TSP6K Dataset
von: Jiang, Peng-Tao, et al.
Veröffentlicht: (2023)
von: Jiang, Peng-Tao, et al.
Veröffentlicht: (2023)
Applying Unsupervised Semantic Segmentation to High-Resolution UAV Imagery for Enhanced Road Scene Parsing
von: Ma, Zihan, et al.
Veröffentlicht: (2024)
von: Ma, Zihan, et al.
Veröffentlicht: (2024)
Two-Stream Interactive Joint Learning of Scene Parsing and Geometric Vision Tasks
von: Tang, Guanfeng, et al.
Veröffentlicht: (2026)
von: Tang, Guanfeng, et al.
Veröffentlicht: (2026)
These Maps Are Made by Propagation: Adapting Deep Stereo Networks to Road Scenarios with Decisive Disparity Diffusion
von: Liu, Chuang-Wei, et al.
Veröffentlicht: (2024)
von: Liu, Chuang-Wei, et al.
Veröffentlicht: (2024)
CSFNet: A Cosine Similarity Fusion Network for Real-Time RGB-X Semantic Segmentation of Driving Scenes
von: Qashqai, Danial, et al.
Veröffentlicht: (2024)
von: Qashqai, Danial, et al.
Veröffentlicht: (2024)
DeFusion: An Effective Decoupling Fusion Network for Multi-Modal Pregnancy Prediction
von: Ouyang, Xueqiang, et al.
Veröffentlicht: (2025)
von: Ouyang, Xueqiang, et al.
Veröffentlicht: (2025)
PIG: Prompt Images Guidance for Night-Time Scene Parsing
von: Xie, Zhifeng, et al.
Veröffentlicht: (2024)
von: Xie, Zhifeng, et al.
Veröffentlicht: (2024)
Exploring Modality-Aware Fusion and Decoupled Temporal Propagation for Multi-Modal Object Tracking
von: Wang, Shilei, et al.
Veröffentlicht: (2026)
von: Wang, Shilei, et al.
Veröffentlicht: (2026)
Establishing Reality-Virtuality Interconnections in Urban Digital Twins for Superior Intelligent Road Inspection and Simulation
von: Zhang, Yikang, et al.
Veröffentlicht: (2024)
von: Zhang, Yikang, et al.
Veröffentlicht: (2024)
TFDet: Target-Aware Fusion for RGB-T Pedestrian Detection
von: Zhang, Xue, et al.
Veröffentlicht: (2023)
von: Zhang, Xue, et al.
Veröffentlicht: (2023)
DepthGait: Multi-Scale Cross-Level Feature Fusion of RGB-Derived Depth and Silhouette Sequences for Robust Gait Recognition
von: Li, Xinzhu, et al.
Veröffentlicht: (2025)
von: Li, Xinzhu, et al.
Veröffentlicht: (2025)
SongFormer: Scaling Music Structure Analysis with Heterogeneous Supervision
von: Hao, Chunbo, et al.
Veröffentlicht: (2025)
von: Hao, Chunbo, et al.
Veröffentlicht: (2025)
DecoupledGaussian: Object-Scene Decoupling for Physics-Based Interaction
von: Wang, Miaowei, et al.
Veröffentlicht: (2025)
von: Wang, Miaowei, et al.
Veröffentlicht: (2025)
Feature-Aware Noise Contrastive Learning for Unsupervised Red Panda Re-Identification
von: Zhang, Jincheng, et al.
Veröffentlicht: (2024)
von: Zhang, Jincheng, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
RoadFormer: Duplex Transformer for RGB-Normal Semantic Road Scene Parsing
von: Li, Jiahang, et al.
Veröffentlicht: (2023) -
HAPNet: Toward Superior RGB-Thermal Scene Parsing via Hybrid, Asymmetric, and Progressive Heterogeneous Feature Fusion
von: Li, Jiahang, et al.
Veröffentlicht: (2024) -
DepthMatch: Semi-Supervised RGB-D Scene Parsing through Depth-Guided Regularization
von: Huang, Jianxin, et al.
Veröffentlicht: (2025) -
RoadFormer : Local-Global Feature Fusion for Road Surface Classification in Autonomous Driving
von: Wang, Tianze, et al.
Veröffentlicht: (2025) -
SNE-RoadSegV2: Advancing Heterogeneous Feature Fusion and Fallibility Awareness for Freespace Detection
von: Feng, Yi, et al.
Veröffentlicht: (2024)