Depth and Image Fusion for Road Obstacle Detection Using Stereo Camera
Fuente:
arXiv
Guardado en:
| Autores principales: | Perezyabov, Oleg, Gavrilenkov, Mikhail, Afanasyev, Ilya |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Single Image Dehazing Using Scene Depth Ordering
por: Ling, Pengyang, et al.
Publicado: (2024)
por: Ling, Pengyang, et al.
Publicado: (2024)
Disparity-based Stereo Image Compression with Aligned Cross-View Priors
por: Zhai, Yongqi, et al.
Publicado: (2022)
por: Zhai, Yongqi, et al.
Publicado: (2022)
Scalable Image Coding for Humans and Machines Using Feature Fusion Network
por: Shindo, Takahiro, et al.
Publicado: (2024)
por: Shindo, Takahiro, et al.
Publicado: (2024)
DepthGait: Multi-Scale Cross-Level Feature Fusion of RGB-Derived Depth and Silhouette Sequences for Robust Gait Recognition
por: Li, Xinzhu, et al.
Publicado: (2025)
por: Li, Xinzhu, et al.
Publicado: (2025)
Extending Depth of Field for Varifocal Multiview Images
por: Li, Zhilong, et al.
Publicado: (2024)
por: Li, Zhilong, et al.
Publicado: (2024)
TMFNet: Two-Stream Multi-Channels Fusion Networks for Color Image Operation Chain Detection
por: Niu, Yakun, et al.
Publicado: (2024)
por: Niu, Yakun, et al.
Publicado: (2024)
Interactive Spatial-Frequency Fusion Mamba for Multi-Modal Image Fusion
por: Zhu, Yixin, et al.
Publicado: (2026)
por: Zhu, Yixin, et al.
Publicado: (2026)
SFFNet: Synergistic Feature Fusion Network With Dual-Domain Edge Enhancement for UAV Image Object Detection
por: Zhang, Wenfeng, et al.
Publicado: (2026)
por: Zhang, Wenfeng, et al.
Publicado: (2026)
DanceCamera3D: 3D Camera Movement Synthesis with Music and Dance
por: Wang, Zixuan, et al.
Publicado: (2024)
por: Wang, Zixuan, et al.
Publicado: (2024)
Multi-Modal Image Fusion via Intervention-Stable Feature Learning
por: Wang, Xue, et al.
Publicado: (2026)
por: Wang, Xue, et al.
Publicado: (2026)
WaveMamba: Wavelet-Driven Mamba Fusion for RGB-Infrared Object Detection
por: Zhu, Haodong, et al.
Publicado: (2025)
por: Zhu, Haodong, et al.
Publicado: (2025)
Transcending Fusion: A Multi-Scale Alignment Method for Remote Sensing Image-Text Retrieval
por: Yang, Rui, et al.
Publicado: (2024)
por: Yang, Rui, et al.
Publicado: (2024)
ASAP: Advancing Semantic Alignment Promotes Multi-Modal Manipulation Detecting and Grounding
por: Zhang, Zhenxing, et al.
Publicado: (2024)
por: Zhang, Zhenxing, et al.
Publicado: (2024)
MRD: Multi-resolution Retrieval-Detection Fusion for High-Resolution Image Understanding
por: Yang, Fan, et al.
Publicado: (2025)
por: Yang, Fan, et al.
Publicado: (2025)
Automatic Prompt Generation and Grounding Object Detection for Zero-Shot Image Anomaly Detection
por: Cheung, Tsun-Hin, et al.
Publicado: (2024)
por: Cheung, Tsun-Hin, et al.
Publicado: (2024)
Camera Trajectory Generation: A Comprehensive Survey of Methods, Metrics, and Future Directions
por: Dehghanian, Zahra, et al.
Publicado: (2025)
por: Dehghanian, Zahra, et al.
Publicado: (2025)
Reinforcing Pre-trained Models Using Counterfactual Images
por: Li, Xiang, et al.
Publicado: (2024)
por: Li, Xiang, et al.
Publicado: (2024)
Marker-Based Extrinsic Calibration Method for Accurate Multi-Camera 3D Reconstruction
por: Garcia-D'Urso, Nahuel, et al.
Publicado: (2025)
por: Garcia-D'Urso, Nahuel, et al.
Publicado: (2025)
An Effective Image Copy-Move Forgery Detection Using Entropy Information
por: Jiang, Li, et al.
Publicado: (2023)
por: Jiang, Li, et al.
Publicado: (2023)
DanceCamAnimator: Keyframe-Based Controllable 3D Dance Camera Synthesis
por: Wang, Zixuan, et al.
Publicado: (2024)
por: Wang, Zixuan, et al.
Publicado: (2024)
Copy-Move Forgery Detection and Question Answering for Remote Sensing Image
por: Zhang, Ze, et al.
Publicado: (2024)
por: Zhang, Ze, et al.
Publicado: (2024)
Learning Efficient Unsupervised Satellite Image-based Building Damage Detection
por: Zhang, Yiyun, et al.
Publicado: (2023)
por: Zhang, Yiyun, et al.
Publicado: (2023)
A Sleep Monitoring System Based on Audio, Video and Depth Information
por: Chen, Lyn Chao-ling, et al.
Publicado: (2025)
por: Chen, Lyn Chao-ling, et al.
Publicado: (2025)
Nutrition Estimation for Dietary Management: A Transformer Approach with Depth Sensing
por: Kwan, Zhengyi, et al.
Publicado: (2024)
por: Kwan, Zhengyi, et al.
Publicado: (2024)
VAAS: Vision-Attention Anomaly Scoring for Image Manipulation Detection in Digital Forensics
por: Bamigbade, Opeyemi, et al.
Publicado: (2025)
por: Bamigbade, Opeyemi, et al.
Publicado: (2025)
GMFVAD: Using Grained Multi-modal Feature to Improve Video Anomaly Detection
por: Dai, Guangyu, et al.
Publicado: (2025)
por: Dai, Guangyu, et al.
Publicado: (2025)
"Humor, Art, or Misinformation?": A Multimodal Dataset for Intent-Aware Synthetic Image Detection
por: Skoularikis, Anastasios, et al.
Publicado: (2025)
por: Skoularikis, Anastasios, et al.
Publicado: (2025)
FakeBench: Probing Explainable Fake Image Detection via Large Multimodal Models
por: Li, Yixuan, et al.
Publicado: (2024)
por: Li, Yixuan, et al.
Publicado: (2024)
MMSD3.0: A Multi-Image Benchmark for Real-World Multimodal Sarcasm Detection
por: Zhao, Haochen, et al.
Publicado: (2025)
por: Zhao, Haochen, et al.
Publicado: (2025)
Anchoring Emotions in Text: Robust Multimodal Fusion for Mimicry Intensity Estimation
por: Zhu, Lingsi, et al.
Publicado: (2026)
por: Zhu, Lingsi, et al.
Publicado: (2026)
MoRAG -- Multi-Fusion Retrieval Augmented Generation for Human Motion
por: Kalakonda, Sai Shashank, et al.
Publicado: (2024)
por: Kalakonda, Sai Shashank, et al.
Publicado: (2024)
Tile Classification Based Viewport Prediction with Multi-modal Fusion Transformer
por: Zhang, Zhihao, et al.
Publicado: (2023)
por: Zhang, Zhihao, et al.
Publicado: (2023)
Enhancing Interactive Image Retrieval With Query Rewriting Using Large Language Models and Vision Language Models
por: Zhu, Hongyi, et al.
Publicado: (2024)
por: Zhu, Hongyi, et al.
Publicado: (2024)
Relating CNN-Transformer Fusion Network for Change Detection
por: Gao, Yuhao, et al.
Publicado: (2024)
por: Gao, Yuhao, et al.
Publicado: (2024)
KAN-Based Fusion of Dual-Domain for Audio-Driven Facial Landmarks Generation
por: Vo-Thanh, Hoang-Son, et al.
Publicado: (2024)
por: Vo-Thanh, Hoang-Son, et al.
Publicado: (2024)
Cross-Attention Fusion of Visual and Geometric Features for Large Vocabulary Arabic Lipreading
por: Daou, Samar, et al.
Publicado: (2024)
por: Daou, Samar, et al.
Publicado: (2024)
MTFusion: Reconstructing Any 3D Object from Single Image Using Multi-word Textual Inversion
por: Liu, Yu, et al.
Publicado: (2024)
por: Liu, Yu, et al.
Publicado: (2024)
FeatDistill: A Feature Distillation Enhanced Multi-Expert Ensemble Framework for Robust AI-generated Image Detection
por: Tu, Zhilin, et al.
Publicado: (2026)
por: Tu, Zhilin, et al.
Publicado: (2026)
DreamCinema: Cinematic Transfer with Free Camera and 3D Character
por: Chen, Weiliang, et al.
Publicado: (2024)
por: Chen, Weiliang, et al.
Publicado: (2024)
GS-ProCams: Gaussian Splatting-based Projector-Camera Systems
por: Deng, Qingyue, et al.
Publicado: (2024)
por: Deng, Qingyue, et al.
Publicado: (2024)
Ejemplares similares
-
Single Image Dehazing Using Scene Depth Ordering
por: Ling, Pengyang, et al.
Publicado: (2024) -
Disparity-based Stereo Image Compression with Aligned Cross-View Priors
por: Zhai, Yongqi, et al.
Publicado: (2022) -
Scalable Image Coding for Humans and Machines Using Feature Fusion Network
por: Shindo, Takahiro, et al.
Publicado: (2024) -
DepthGait: Multi-Scale Cross-Level Feature Fusion of RGB-Derived Depth and Silhouette Sequences for Robust Gait Recognition
por: Li, Xinzhu, et al.
Publicado: (2025) -
Extending Depth of Field for Varifocal Multiview Images
por: Li, Zhilong, et al.
Publicado: (2024)