HAPNet: Toward Superior RGB-Thermal Scene Parsing via Hybrid, Asymmetric, and Progressive Heterogeneous Feature Fusion
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Li, Jiahang, Yun, Peng, Xu, Yang, Zhang, Ye, Sun, Mingjian, Chen, Qijun, Alexander, Ilin, Fan, Rui |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
RoadFormer+: Delivering RGB-X Scene Parsing through Scale-Aware Information Decoupling and Advanced Heterogeneous Feature Fusion
von: Huang, Jianxin, et al.
Veröffentlicht: (2024)
von: Huang, Jianxin, et al.
Veröffentlicht: (2024)
RoadFormer: Duplex Transformer for RGB-Normal Semantic Road Scene Parsing
von: Li, Jiahang, et al.
Veröffentlicht: (2023)
von: Li, Jiahang, et al.
Veröffentlicht: (2023)
DepthMatch: Semi-Supervised RGB-D Scene Parsing through Depth-Guided Regularization
von: Huang, Jianxin, et al.
Veröffentlicht: (2025)
von: Huang, Jianxin, et al.
Veröffentlicht: (2025)
Fully Exploiting Vision Foundation Model's Profound Prior Knowledge for Generalizable RGB-Depth Driving Scene Parsing
von: Guo, Sicen, et al.
Veröffentlicht: (2025)
von: Guo, Sicen, et al.
Veröffentlicht: (2025)
Unsupervised Collaborative Domain Adaptation for Driving Scene Parsing
von: Fan, Jiahe, et al.
Veröffentlicht: (2026)
von: Fan, Jiahe, et al.
Veröffentlicht: (2026)
RTFDNet: Fusion-Decoupling for Robust RGB-T Segmentation
von: Tan, Kunyu, et al.
Veröffentlicht: (2026)
von: Tan, Kunyu, et al.
Veröffentlicht: (2026)
SNE-RoadSegV2: Advancing Heterogeneous Feature Fusion and Fallibility Awareness for Freespace Detection
von: Feng, Yi, et al.
Veröffentlicht: (2024)
von: Feng, Yi, et al.
Veröffentlicht: (2024)
Towards Balanced RGB-TSDF Fusion for Consistent Semantic Scene Completion by 3D RGB Feature Completion and a Classwise Entropy Loss Function
von: Ding, Laiyan, et al.
Veröffentlicht: (2024)
von: Ding, Laiyan, et al.
Veröffentlicht: (2024)
Text-Scene: A Scene-to-Language Parsing Framework for 3D Scene Understanding
von: Li, Haoyuan, et al.
Veröffentlicht: (2025)
von: Li, Haoyuan, et al.
Veröffentlicht: (2025)
Hybrid Cross-Device Localization via Neural Metric Learning and Feature Fusion
von: Lin, Meixia, et al.
Veröffentlicht: (2026)
von: Lin, Meixia, et al.
Veröffentlicht: (2026)
Transformer-based RGB-T Tracking with Channel and Spatial Feature Fusion
von: Li, Yunfeng, et al.
Veröffentlicht: (2024)
von: Li, Yunfeng, et al.
Veröffentlicht: (2024)
Two-Stream Interactive Joint Learning of Scene Parsing and Geometric Vision Tasks
von: Tang, Guanfeng, et al.
Veröffentlicht: (2026)
von: Tang, Guanfeng, et al.
Veröffentlicht: (2026)
RGB-D Indiscernible Object Counting in Underwater Scenes
von: Sun, Guolei, et al.
Veröffentlicht: (2023)
von: Sun, Guolei, et al.
Veröffentlicht: (2023)
Progressive Feature Fusion Network for Enhancing Image Quality Assessment
von: Wu, Kaiqun, et al.
Veröffentlicht: (2024)
von: Wu, Kaiqun, et al.
Veröffentlicht: (2024)
Modality-Decoupled RGB-Thermal Object Detector via Query Fusion
von: Tian, Chao, et al.
Veröffentlicht: (2026)
von: Tian, Chao, et al.
Veröffentlicht: (2026)
Unleashing the Power of Motion and Depth: A Selective Fusion Strategy for RGB-D Video Salient Object Detection
von: He, Jiahao, et al.
Veröffentlicht: (2025)
von: He, Jiahao, et al.
Veröffentlicht: (2025)
A Saliency Enhanced Feature Fusion based multiscale RGB-D Salient Object Detection Network
von: Huang, Rui, et al.
Veröffentlicht: (2024)
von: Huang, Rui, et al.
Veröffentlicht: (2024)
A Comprehensive Survey on Video Scene Parsing:Advances, Challenges, and Prospects
von: Xie, Guohuan, et al.
Veröffentlicht: (2025)
von: Xie, Guohuan, et al.
Veröffentlicht: (2025)
Frequency-Guided Fusion For RGB-Thermal Semantic Segmentation
von: Canıtez, İsmail Emre, et al.
Veröffentlicht: (2026)
von: Canıtez, İsmail Emre, et al.
Veröffentlicht: (2026)
Efficient RGB-D Scene Understanding via Multi-task Adaptive Learning and Cross-dimensional Feature Guidance
von: Sun, Guodong, et al.
Veröffentlicht: (2026)
von: Sun, Guodong, et al.
Veröffentlicht: (2026)
Asymmetric Feature Fusion for Image Retrieval
von: Wu, Hui, et al.
Veröffentlicht: (2024)
von: Wu, Hui, et al.
Veröffentlicht: (2024)
Towards Real-World Document Parsing via Realistic Scene Synthesis and Document-Aware Training
von: Li, Gengluo, et al.
Veröffentlicht: (2026)
von: Li, Gengluo, et al.
Veröffentlicht: (2026)
PIG: Prompt Images Guidance for Night-Time Scene Parsing
von: Xie, Zhifeng, et al.
Veröffentlicht: (2024)
von: Xie, Zhifeng, et al.
Veröffentlicht: (2024)
HPE:Answering Complex Questions over Text by Hybrid Question Parsing and Execution
von: Liu, Ye, et al.
Veröffentlicht: (2023)
von: Liu, Ye, et al.
Veröffentlicht: (2023)
Face Liveness Detection Using RGB and Thermal Image Fusion
von: Erşan, Merve, et al.
Veröffentlicht: (2026)
von: Erşan, Merve, et al.
Veröffentlicht: (2026)
Spectral-Aware Global Fusion for RGB-Thermal Semantic Segmentation
von: Zhang, Ce, et al.
Veröffentlicht: (2025)
von: Zhang, Ce, et al.
Veröffentlicht: (2025)
Dynamic Graph Neural Network with Adaptive Features Selection for RGB-D Based Indoor Scene Recognition
von: Liu, Qiong, et al.
Veröffentlicht: (2026)
von: Liu, Qiong, et al.
Veröffentlicht: (2026)
Surgical Scene Segmentation by Transformer With Asymmetric Feature Enhancement
von: Yuan, Cheng, et al.
Veröffentlicht: (2024)
von: Yuan, Cheng, et al.
Veröffentlicht: (2024)
Research Status and Prospects of MIG and Hybrid Welding of Titanium and Titanium Alloys
von: Xi Niu, et al.
Veröffentlicht: (2025)
von: Xi Niu, et al.
Veröffentlicht: (2025)
Traffic Scene Parsing through the TSP6K Dataset
von: Jiang, Peng-Tao, et al.
Veröffentlicht: (2023)
von: Jiang, Peng-Tao, et al.
Veröffentlicht: (2023)
SceneParser: Hierarchical Scene Parsing for Visual Semantics Understanding
von: Xu, Pengxin, et al.
Veröffentlicht: (2026)
von: Xu, Pengxin, et al.
Veröffentlicht: (2026)
Toward Cognitive AI
von: Ilin, Nikolai
Veröffentlicht: (2025)
von: Ilin, Nikolai
Veröffentlicht: (2025)
Dolphin: Document Image Parsing via Heterogeneous Anchor Prompting
von: Feng, Hao, et al.
Veröffentlicht: (2025)
von: Feng, Hao, et al.
Veröffentlicht: (2025)
RGB-Thermal Infrared Fusion for Robust Depth Estimation in Complex Environments
von: Meng, Zelin, et al.
Veröffentlicht: (2025)
von: Meng, Zelin, et al.
Veröffentlicht: (2025)
Establishing Reality-Virtuality Interconnections in Urban Digital Twins for Superior Intelligent Road Inspection and Simulation
von: Zhang, Yikang, et al.
Veröffentlicht: (2024)
von: Zhang, Yikang, et al.
Veröffentlicht: (2024)
Glass Surface Segmentation with an RGB-D Camera via Weighted Feature Fusion for Service Robots
von: Lin, Henghong, et al.
Veröffentlicht: (2025)
von: Lin, Henghong, et al.
Veröffentlicht: (2025)
Bring Event into RGB and LiDAR: Hierarchical Visual-Motion Fusion for Scene Flow
von: Zhou, Hanyu, et al.
Veröffentlicht: (2024)
von: Zhou, Hanyu, et al.
Veröffentlicht: (2024)
Online,Target-Free LiDAR-Camera Extrinsic Calibration via Cross-Modal Mask Matching
von: Huang, Zhiwei, et al.
Veröffentlicht: (2024)
von: Huang, Zhiwei, et al.
Veröffentlicht: (2024)
Temporal Propagation of Asymmetric Feature Pyramid for Surgical Scene Segmentation
von: Yuan, Cheng, et al.
Veröffentlicht: (2025)
von: Yuan, Cheng, et al.
Veröffentlicht: (2025)
MicLog: Towards Accurate and Efficient LLM-based Log Parsing via Progressive Meta In-Context Learning
von: Yu, Jianbo, et al.
Veröffentlicht: (2026)
von: Yu, Jianbo, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
RoadFormer+: Delivering RGB-X Scene Parsing through Scale-Aware Information Decoupling and Advanced Heterogeneous Feature Fusion
von: Huang, Jianxin, et al.
Veröffentlicht: (2024) -
RoadFormer: Duplex Transformer for RGB-Normal Semantic Road Scene Parsing
von: Li, Jiahang, et al.
Veröffentlicht: (2023) -
DepthMatch: Semi-Supervised RGB-D Scene Parsing through Depth-Guided Regularization
von: Huang, Jianxin, et al.
Veröffentlicht: (2025) -
Fully Exploiting Vision Foundation Model's Profound Prior Knowledge for Generalizable RGB-Depth Driving Scene Parsing
von: Guo, Sicen, et al.
Veröffentlicht: (2025) -
Unsupervised Collaborative Domain Adaptation for Driving Scene Parsing
von: Fan, Jiahe, et al.
Veröffentlicht: (2026)