MSMVD: Exploiting Multi-scale Image Features via Multi-scale BEV Features for Multi-view Pedestrian Detection
Fuente:
arXiv
Saved in:
| Main Authors: | Yamane, Taiga, Suzuki, Satoshi, Masumura, Ryo, Orihashi, Shota, Tanaka, Tomohiro, Ihori, Mana, Makishima, Naoki, Kawata, Naotaka |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Few-shot Personalization via In-Context Learning for Speech Emotion Recognition based on Speech-Language Model
by: Ihori, Mana, et al.
Published: (2025)
by: Ihori, Mana, et al.
Published: (2025)
Joint Modeling of Big Five and HEXACO for Multimodal Apparent Personality-trait Recognition
by: Masumura, Ryo, et al.
Published: (2025)
by: Masumura, Ryo, et al.
Published: (2025)
Difference Vector Equalization for Robust Fine-tuning of Vision-Language Models
by: Suzuki, Satoshi, et al.
Published: (2025)
by: Suzuki, Satoshi, et al.
Published: (2025)
MVTrajecter: Multi-View Pedestrian Tracking with Trajectory Motion Cost and Trajectory Appearance Cost
by: Yamane, Taiga, et al.
Published: (2025)
by: Yamane, Taiga, et al.
Published: (2025)
MVAFormer: RGB-based Multi-View Spatio-Temporal Action Recognition with Transformer
by: Yamane, Taiga, et al.
Published: (2025)
by: Yamane, Taiga, et al.
Published: (2025)
GraphBEV: Towards Robust BEV Feature Alignment for Multi-Modal 3D Object Detection
by: Song, Ziying, et al.
Published: (2024)
by: Song, Ziying, et al.
Published: (2024)
GeoBEV: Learning Geometric BEV Representation for Multi-view 3D Object Detection
by: Zhang, Jinqing, et al.
Published: (2024)
by: Zhang, Jinqing, et al.
Published: (2024)
Alignment-Free Training for Transducer-based Multi-Talker ASR
by: Moriya, Takafumi, et al.
Published: (2024)
by: Moriya, Takafumi, et al.
Published: (2024)
MVUDA: Unsupervised Domain Adaptation for Multi-view Pedestrian Detection
by: Brorsson, Erik, et al.
Published: (2024)
by: Brorsson, Erik, et al.
Published: (2024)
Factor-Conditioned Speaking-Style Captioning
by: Ando, Atsushi, et al.
Published: (2024)
by: Ando, Atsushi, et al.
Published: (2024)
Multi-scale Semantic Prior Features Guided Deep Neural Network for Urban Street-view Image
by: Zeng, Jianshun, et al.
Published: (2024)
by: Zeng, Jianshun, et al.
Published: (2024)
Multi-scale Feature Fusion with Point Pyramid for 3D Object Detection
by: Lu, Weihao, et al.
Published: (2024)
by: Lu, Weihao, et al.
Published: (2024)
EMDFNet: Efficient Multi-scale and Diverse Feature Network for Traffic Sign Detection
by: Li, Pengyu, et al.
Published: (2024)
by: Li, Pengyu, et al.
Published: (2024)
Multi-scale Feature Enhancement in Multi-task Learning for Medical Image Analysis
by: Bui, Phuoc-Nguyen, et al.
Published: (2024)
by: Bui, Phuoc-Nguyen, et al.
Published: (2024)
Integrating Multi-scale and Multi-filtration Topological Features for Medical Image Classification
by: Gu, Pengfei, et al.
Published: (2025)
by: Gu, Pengfei, et al.
Published: (2025)
Exploiting Modality-Specific Features For Multi-Modal Manipulation Detection And Grounding
by: Wang, Jiazhen, et al.
Published: (2023)
by: Wang, Jiazhen, et al.
Published: (2023)
HV-BEV: Decoupling Horizontal and Vertical Feature Sampling for Multi-View 3D Object Detection
by: Wu, Di, et al.
Published: (2024)
by: Wu, Di, et al.
Published: (2024)
Learning Multi-scale Spatial-frequency Features for Image Denoising
by: Zhao, Xu, et al.
Published: (2025)
by: Zhao, Xu, et al.
Published: (2025)
RMFAT: Recurrent Multi-scale Feature Atmospheric Turbulence Mitigator
by: Liu, Zhiming, et al.
Published: (2025)
by: Liu, Zhiming, et al.
Published: (2025)
FSD-BEV: Foreground Self-Distillation for Multi-view 3D Object Detection
by: Jiang, Zheng, et al.
Published: (2024)
by: Jiang, Zheng, et al.
Published: (2024)
Transferable Class Statistics and Multi-scale Feature Approximation for 3D Object Detection
by: Peng, Hao, et al.
Published: (2025)
by: Peng, Hao, et al.
Published: (2025)
MGDFIS: Multi-scale Global-detail Feature Integration Strategy for Small Object Detection
by: Wang, Yuxiang, et al.
Published: (2025)
by: Wang, Yuxiang, et al.
Published: (2025)
Sparse BEV Fusion with Self-View Consistency for Multi-View Detection and Tracking
by: Toida, Keisuke, et al.
Published: (2025)
by: Toida, Keisuke, et al.
Published: (2025)
TiGDistill-BEV: Multi-view BEV 3D Object Detection via Target Inner-Geometry Learning Distillation
by: Xu, Shaoqing, et al.
Published: (2024)
by: Xu, Shaoqing, et al.
Published: (2024)
ShibuyaSocial: Multi-scale Model of Pedestrian Flows in Scramble Crossing
by: Sakurai, Akihiro, et al.
Published: (2025)
by: Sakurai, Akihiro, et al.
Published: (2025)
All-in-One ASR: Unifying Encoder-Decoder Models of CTC, Attention, and Transducer in Dual-Mode ASR
by: Moriya, Takafumi, et al.
Published: (2025)
by: Moriya, Takafumi, et al.
Published: (2025)
TriBand-BEV: Real-Time LiDAR-Only 3D Pedestrian Detection via Height-Aware BEV and High-Resolution Feature Fusion
by: Khoshkdahan, Mohammad, et al.
Published: (2026)
by: Khoshkdahan, Mohammad, et al.
Published: (2026)
Network Anomaly Traffic Detection via Multi-view Feature Fusion
by: Hao, Song, et al.
Published: (2024)
by: Hao, Song, et al.
Published: (2024)
Improving Speaker Representations Using Contrastive Losses on Multi-scale Features
by: Dixit, Satvik, et al.
Published: (2024)
by: Dixit, Satvik, et al.
Published: (2024)
Consolidating Attention Features for Multi-view Image Editing
by: Patashnik, Or, et al.
Published: (2024)
by: Patashnik, Or, et al.
Published: (2024)
MapFusion: A Novel BEV Feature Fusion Network for Multi-modal Map Construction
by: Hao, Xiaoshuai, et al.
Published: (2025)
by: Hao, Xiaoshuai, et al.
Published: (2025)
Large-scale Multi-objective Feature Selection: A Multi-phase Search Space Shrinking Approach
by: Bidgoli, Azam Asilian, et al.
Published: (2024)
by: Bidgoli, Azam Asilian, et al.
Published: (2024)
BEV-VAE: Multi-view Image Generation with Spatial Consistency for Autonomous Driving
by: Chen, Zeming, et al.
Published: (2025)
by: Chen, Zeming, et al.
Published: (2025)
MsFIN: Multi-scale Feature Interaction Network for Traffic Accident Anticipation
by: Wu, Tongshuai, et al.
Published: (2025)
by: Wu, Tongshuai, et al.
Published: (2025)
Interpretable QSPR Modeling using Recursive Feature Machines and Multi-scale Fingerprints
by: Shen, Jiaxuan, et al.
Published: (2024)
by: Shen, Jiaxuan, et al.
Published: (2024)
EVC-MF: End-to-end Video Captioning Network with Multi-scale Features
by: Niu, Tian-Zi, et al.
Published: (2024)
by: Niu, Tian-Zi, et al.
Published: (2024)
ContrastAlign: Toward Robust BEV Feature Alignment via Contrastive Learning for Multi-Modal 3D Object Detection
by: Song, Ziying, et al.
Published: (2024)
by: Song, Ziying, et al.
Published: (2024)
Multi-view Feature Augmentation with Adaptive Class Activation Mapping
by: Gao, Xiang, et al.
Published: (2022)
by: Gao, Xiang, et al.
Published: (2022)
Feature Disentanglement in GANs for Photorealistic Multi‐view Hair Transfer
by: Jiayi Xu, et al.
Published: (2025)
by: Jiayi Xu, et al.
Published: (2025)
EMIFF: Enhanced Multi-scale Image Feature Fusion for Vehicle-Infrastructure Cooperative 3D Object Detection
by: Wang, Zhe, et al.
Published: (2024)
by: Wang, Zhe, et al.
Published: (2024)
Similar Items
-
Few-shot Personalization via In-Context Learning for Speech Emotion Recognition based on Speech-Language Model
by: Ihori, Mana, et al.
Published: (2025) -
Joint Modeling of Big Five and HEXACO for Multimodal Apparent Personality-trait Recognition
by: Masumura, Ryo, et al.
Published: (2025) -
Difference Vector Equalization for Robust Fine-tuning of Vision-Language Models
by: Suzuki, Satoshi, et al.
Published: (2025) -
MVTrajecter: Multi-View Pedestrian Tracking with Trajectory Motion Cost and Trajectory Appearance Cost
by: Yamane, Taiga, et al.
Published: (2025) -
MVAFormer: RGB-based Multi-View Spatio-Temporal Action Recognition with Transformer
by: Yamane, Taiga, et al.
Published: (2025)