Gespeichert in:
| Hauptverfasser: | Gordon, Lucia, Lang, Nico, Ressijac, Catherine, Davies, Andrew |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | https://arxiv.org/abs/2410.04833 |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Find Rhinos without Finding Rhinos: Active Learning with Multimodal Imagery of South African Rhino Habitats
von: Gordon, Lucia, et al.
Veröffentlicht: (2024)
von: Gordon, Lucia, et al.
Veröffentlicht: (2024)
Contrasting local and global modeling with machine learning and satellite data: A case study estimating tree canopy height in African savannas
von: Rolf, Esther, et al.
Veröffentlicht: (2024)
von: Rolf, Esther, et al.
Veröffentlicht: (2024)
CSA: Data-efficient Mapping of Unimodal Features to Multimodal Features
von: Li, Po-han, et al.
Veröffentlicht: (2024)
von: Li, Po-han, et al.
Veröffentlicht: (2024)
NavMapFusion: Diffusion-based Fusion of Navigation Maps for Online Vectorized HD Map Construction
von: Monninger, Thomas, et al.
Veröffentlicht: (2025)
von: Monninger, Thomas, et al.
Veröffentlicht: (2025)
MemFusionMap: Working Memory Fusion for Online Vectorized HD Map Construction
von: Song, Jingyu, et al.
Veröffentlicht: (2024)
von: Song, Jingyu, et al.
Veröffentlicht: (2024)
A Multimodal Fusion Model Leveraging MLP Mixer and Handcrafted Features-based Deep Learning Networks for Facial Palsy Detection
von: Oo, Heng Yim Nicole, et al.
Veröffentlicht: (2025)
von: Oo, Heng Yim Nicole, et al.
Veröffentlicht: (2025)
Evaluating GPT-5 as a Multimodal Clinical Reasoner: A Landscape Commentary
von: Florea, Alexandru, et al.
Veröffentlicht: (2026)
von: Florea, Alexandru, et al.
Veröffentlicht: (2026)
MMEarth: Exploring Multi-Modal Pretext Tasks For Geospatial Representation Learning
von: Nedungadi, Vishal, et al.
Veröffentlicht: (2024)
von: Nedungadi, Vishal, et al.
Veröffentlicht: (2024)
Automatic Image Annotation for Mapped Features Detection
von: Noizet, Maxime, et al.
Veröffentlicht: (2024)
von: Noizet, Maxime, et al.
Veröffentlicht: (2024)
Data-Efficient Multimodal Fusion on a Single GPU
von: Vouitsis, Noël, et al.
Veröffentlicht: (2023)
von: Vouitsis, Noël, et al.
Veröffentlicht: (2023)
Learnable Expansion of Graph Operators for Multi-Modal Feature Fusion
von: Ding, Dexuan, et al.
Veröffentlicht: (2024)
von: Ding, Dexuan, et al.
Veröffentlicht: (2024)
Unsupervised Active Learning via Natural Feature Progressive Framework
von: Liu, Yuxi, et al.
Veröffentlicht: (2025)
von: Liu, Yuxi, et al.
Veröffentlicht: (2025)
Token Activation Map to Visually Explain Multimodal LLMs
von: Li, Yi, et al.
Veröffentlicht: (2025)
von: Li, Yi, et al.
Veröffentlicht: (2025)
FedAFD: Multimodal Federated Learning via Adversarial Fusion and Distillation
von: Tan, Min, et al.
Veröffentlicht: (2026)
von: Tan, Min, et al.
Veröffentlicht: (2026)
Enhancing Out-of-Distribution Detection with Multitesting-based Layer-wise Feature Fusion
von: Li, Jiawei, et al.
Veröffentlicht: (2024)
von: Li, Jiawei, et al.
Veröffentlicht: (2024)
FreeBind: Free Lunch in Unified Multimodal Space via Knowledge Fusion
von: Wang, Zehan, et al.
Veröffentlicht: (2024)
von: Wang, Zehan, et al.
Veröffentlicht: (2024)
StitchFusion: Weaving Any Visual Modalities to Enhance Multimodal Semantic Segmentation
von: Li, Bingyu, et al.
Veröffentlicht: (2024)
von: Li, Bingyu, et al.
Veröffentlicht: (2024)
Feature Fusion for Improved Classification: Combining Dempster-Shafer Theory and Multiple CNN Architectures
von: Alzahem, Ayyub, et al.
Veröffentlicht: (2024)
von: Alzahem, Ayyub, et al.
Veröffentlicht: (2024)
VIFO: Visual Feature Empowered Multivariate Time Series Forecasting with Cross-Modal Fusion
von: Wang, Yanlong, et al.
Veröffentlicht: (2025)
von: Wang, Yanlong, et al.
Veröffentlicht: (2025)
Addressing Asynchronicity in Clinical Multimodal Fusion via Individualized Chest X-ray Generation
von: Yao, Wenfang, et al.
Veröffentlicht: (2024)
von: Yao, Wenfang, et al.
Veröffentlicht: (2024)
Exploring a Multimodal Fusion-based Deep Learning Network for Detecting Facial Palsy
von: Oo, Heng Yim Nicole, et al.
Veröffentlicht: (2024)
von: Oo, Heng Yim Nicole, et al.
Veröffentlicht: (2024)
Robult: Leveraging Redundancy and Modality Specific Features for Robust Multimodal Learning
von: Nguyen, Duy A., et al.
Veröffentlicht: (2025)
von: Nguyen, Duy A., et al.
Veröffentlicht: (2025)
Nacala-Roof-Material: Drone Imagery for Roof Detection, Classification, and Segmentation to Support Mosquito-borne Disease Risk Assessment
von: Guthula, Venkanna Babu, et al.
Veröffentlicht: (2024)
von: Guthula, Venkanna Babu, et al.
Veröffentlicht: (2024)
The Lie Derivative for Measuring Learned Equivariance
von: Gruver, Nate, et al.
Veröffentlicht: (2022)
von: Gruver, Nate, et al.
Veröffentlicht: (2022)
Visual Explanations of Image-Text Representations via Multi-Modal Information Bottleneck Attribution
von: Wang, Ying, et al.
Veröffentlicht: (2023)
von: Wang, Ying, et al.
Veröffentlicht: (2023)
MapIQ: Evaluating Multimodal Large Language Models for Map Question Answering
von: Srivastava, Varun, et al.
Veröffentlicht: (2025)
von: Srivastava, Varun, et al.
Veröffentlicht: (2025)
FusionEnsemble-Net: An Attention-Based Ensemble of Spatiotemporal Networks for Multimodal Sign Language Recognition
von: Islam, Md. Milon, et al.
Veröffentlicht: (2025)
von: Islam, Md. Milon, et al.
Veröffentlicht: (2025)
Multi-Layer Feature Fusion with Cross-Channel Attention-Based U-Net for Kidney Tumor Segmentation
von: Neha, Fnu, et al.
Veröffentlicht: (2024)
von: Neha, Fnu, et al.
Veröffentlicht: (2024)
Multi-scale Quaternion CNN and BiGRU with Cross Self-attention Feature Fusion for Fault Diagnosis of Bearing
von: Liu, Huanbai, et al.
Veröffentlicht: (2024)
von: Liu, Huanbai, et al.
Veröffentlicht: (2024)
MGDFIS: Multi-scale Global-detail Feature Integration Strategy for Small Object Detection
von: Wang, Yuxiang, et al.
Veröffentlicht: (2025)
von: Wang, Yuxiang, et al.
Veröffentlicht: (2025)
FILS: Self-Supervised Video Feature Prediction In Semantic Language Space
von: Ahmadian, Mona, et al.
Veröffentlicht: (2024)
von: Ahmadian, Mona, et al.
Veröffentlicht: (2024)
WiTUnet: A U-Shaped Architecture Integrating CNN and Transformer for Improved Feature Alignment and Local Information Fusion
von: Wang, Bin, et al.
Veröffentlicht: (2024)
von: Wang, Bin, et al.
Veröffentlicht: (2024)
VLSU: Mapping the Limits of Joint Multimodal Understanding for AI Safety
von: Palaskar, Shruti, et al.
Veröffentlicht: (2025)
von: Palaskar, Shruti, et al.
Veröffentlicht: (2025)
EDNet: Edge-Optimized Small Target Detection in UAV Imagery -- Faster Context Attention, Better Feature Fusion, and Hardware Acceleration
von: Song, Zhifan, et al.
Veröffentlicht: (2025)
von: Song, Zhifan, et al.
Veröffentlicht: (2025)
LiMTR: Time Series Motion Prediction for Diverse Road Users through Multimodal Feature Integration
von: Oerlemans, Camiel, et al.
Veröffentlicht: (2024)
von: Oerlemans, Camiel, et al.
Veröffentlicht: (2024)
MoPE: Mixture of Prompt Experts for Parameter-Efficient and Scalable Multimodal Fusion
von: Jiang, Ruixiang, et al.
Veröffentlicht: (2024)
von: Jiang, Ruixiang, et al.
Veröffentlicht: (2024)
From Data to Insights: A Covariate Analysis of the IARPA BRIAR Dataset for Multimodal Biometric Recognition Algorithms at Altitude and Range
von: Bolme, David S., et al.
Veröffentlicht: (2024)
von: Bolme, David S., et al.
Veröffentlicht: (2024)
From Radiologist Report to Image Label: Assessing Latent Dirichlet Allocation in Training Neural Networks for Orthopedic Radiograph Classification
von: Olczak, Jakub, et al.
Veröffentlicht: (2024)
von: Olczak, Jakub, et al.
Veröffentlicht: (2024)
Spatiotemporal Air Quality Mapping in Urban Areas Using Sparse Sensor Data, Satellite Imagery, Meteorological Factors, and Spatial Features
von: Ahmad, Osama, et al.
Veröffentlicht: (2025)
von: Ahmad, Osama, et al.
Veröffentlicht: (2025)
GAF-FusionNet: Multimodal ECG Analysis via Gramian Angular Fields and Split Attention
von: Qin, Jiahao, et al.
Veröffentlicht: (2024)
von: Qin, Jiahao, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Find Rhinos without Finding Rhinos: Active Learning with Multimodal Imagery of South African Rhino Habitats
von: Gordon, Lucia, et al.
Veröffentlicht: (2024) -
Contrasting local and global modeling with machine learning and satellite data: A case study estimating tree canopy height in African savannas
von: Rolf, Esther, et al.
Veröffentlicht: (2024) -
CSA: Data-efficient Mapping of Unimodal Features to Multimodal Features
von: Li, Po-han, et al.
Veröffentlicht: (2024) -
NavMapFusion: Diffusion-based Fusion of Navigation Maps for Online Vectorized HD Map Construction
von: Monninger, Thomas, et al.
Veröffentlicht: (2025) -
MemFusionMap: Working Memory Fusion for Online Vectorized HD Map Construction
von: Song, Jingyu, et al.
Veröffentlicht: (2024)