FlyAwareV2: A Multimodal Cross-Domain UAV Dataset for Urban Scene Understanding
Fuente:
arXiv
Saved in:
| Main Authors: | Barbato, Francesco, Caligiuri, Matteo, Zanuttigh, Pietro |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Federated Medical Image Classification under Class and Domain Imbalance exploiting Synthetic Sample Generation
by: Pavan, Martina, et al.
Published: (2026)
by: Pavan, Martina, et al.
Published: (2026)
When Cars meet Drones: Hyperbolic Federated Learning for Source-Free Domain Adaptation in Adverse Weather
by: Rizzoli, Giulia, et al.
Published: (2024)
by: Rizzoli, Giulia, et al.
Published: (2024)
FedPromo: Federated Lightweight Proxy Models at the Edge Bring New Domains to Foundation Models
by: Caligiuri, Matteo, et al.
Published: (2025)
by: Caligiuri, Matteo, et al.
Published: (2025)
Continual Road-Scene Semantic Segmentation via Feature-Aligned Symmetric Multi-Modal Network
by: Barbato, Francesco, et al.
Published: (2023)
by: Barbato, Francesco, et al.
Published: (2023)
NIGHT -- Non-Line-of-Sight Imaging from Indirect Time of Flight Data
by: Caligiuri, Matteo, et al.
Published: (2024)
by: Caligiuri, Matteo, et al.
Published: (2024)
Cross-Architecture Auxiliary Feature Space Translation for Efficient Few-Shot Personalized Object Detection
by: Barbato, Francesco, et al.
Published: (2024)
by: Barbato, Francesco, et al.
Published: (2024)
A Modular System for Enhanced Robustness of Multimedia Understanding Networks via Deep Parametric Estimation
by: Barbato, Francesco, et al.
Published: (2024)
by: Barbato, Francesco, et al.
Published: (2024)
Split&Splat: Zero-Shot Panoptic Segmentation via Explicit Instance Modeling and 3D Gaussian Splatting
by: Monchieri, Leonardo, et al.
Published: (2026)
by: Monchieri, Leonardo, et al.
Published: (2026)
RECALL+: Adversarial Web-based Replay for Continual Learning in Semantic Segmentation
by: Liu, Chang, et al.
Published: (2023)
by: Liu, Chang, et al.
Published: (2023)
EV-Flying: an Event-based Dataset for In-The-Wild Recognition of Flying Objects
by: Magrini, Gabriele, et al.
Published: (2025)
by: Magrini, Gabriele, et al.
Published: (2025)
MultimodalStudio: A Heterogeneous Sensor Dataset and Framework for Neural Rendering across Multiple Imaging Modalities
by: Lincetto, Federico, et al.
Published: (2025)
by: Lincetto, Federico, et al.
Published: (2025)
Learning from the Web: Language Drives Weakly-Supervised Incremental Learning for Semantic Segmentation
by: Liu, Chang, et al.
Published: (2024)
by: Liu, Chang, et al.
Published: (2024)
Syn-Mediverse: A Multimodal Synthetic Dataset for Intelligent Scene Understanding of Healthcare Facilities
by: Mohan, Rohit, et al.
Published: (2023)
by: Mohan, Rohit, et al.
Published: (2023)
SkyScenes: A Synthetic Dataset for Aerial Scene Understanding
by: Khose, Sahil, et al.
Published: (2023)
by: Khose, Sahil, et al.
Published: (2023)
What Demands Attention in Urban Street Scenes? From Scene Understanding towards Road Safety: A Survey of Vision-driven Datasets and Studies
by: Huang, Yaoqi, et al.
Published: (2025)
by: Huang, Yaoqi, et al.
Published: (2025)
TrueCity: Real and Simulated Urban Data for Cross-Domain 3D Scene Understanding
by: Nguyen, Duc, et al.
Published: (2025)
by: Nguyen, Duc, et al.
Published: (2025)
Open-Vocabulary Domain Generalization in Urban-Scene Segmentation
by: Zhao, Dong, et al.
Published: (2026)
by: Zhao, Dong, et al.
Published: (2026)
Cross-Modal and Uncertainty-Aware Agglomeration for Open-Vocabulary 3D Scene Understanding
by: Li, Jinlong, et al.
Published: (2025)
by: Li, Jinlong, et al.
Published: (2025)
SpatialFly: Geometry-Guided Representation Alignment for UAV Vision-and-Language Navigation in Urban Environments
by: Jiang, Wen, et al.
Published: (2026)
by: Jiang, Wen, et al.
Published: (2026)
3D-Aware Multi-Task Learning with Cross-View Correlations for Dense Scene Understanding
by: Wang, Xiaoye, et al.
Published: (2025)
by: Wang, Xiaoye, et al.
Published: (2025)
SANPO: A Scene Understanding, Accessibility and Human Navigation Dataset
by: Waghmare, Sagar M., et al.
Published: (2023)
by: Waghmare, Sagar M., et al.
Published: (2023)
A Large-Scale Multimodal Dataset and Benchmarks for Human Activity Scene Understanding and Reasoning
by: Jiang, Siyang, et al.
Published: (2025)
by: Jiang, Siyang, et al.
Published: (2025)
NAUTILUS: A Large Multimodal Model for Underwater Scene Understanding
by: Xu, Wei, et al.
Published: (2025)
by: Xu, Wei, et al.
Published: (2025)
VUDG: A Dataset for Video Understanding Domain Generalization
by: Wang, Ziyi, et al.
Published: (2025)
by: Wang, Ziyi, et al.
Published: (2025)
From Open-Vocabulary to Vocabulary-Free Semantic Segmentation
by: Reichard, Klara, et al.
Published: (2025)
by: Reichard, Klara, et al.
Published: (2025)
SPORTS: Simultaneous Panoptic Odometry, Rendering, Tracking and Segmentation for Urban Scenes Understanding
by: Yang, Zhiliu, et al.
Published: (2025)
by: Yang, Zhiliu, et al.
Published: (2025)
OmniScene: Attention-Augmented Multimodal 4D Scene Understanding for Autonomous Driving
by: Liu, Pei, et al.
Published: (2025)
by: Liu, Pei, et al.
Published: (2025)
HUGS: Holistic Urban 3D Scene Understanding via Gaussian Splatting
by: Zhou, Hongyu, et al.
Published: (2024)
by: Zhou, Hongyu, et al.
Published: (2024)
SurgMLLMBench: A Multimodal Large Language Model Benchmark Dataset for Surgical Scene Understanding
by: Choi, Tae-Min, et al.
Published: (2025)
by: Choi, Tae-Min, et al.
Published: (2025)
MMIS: Multimodal Dataset for Interior Scene Visual Generation and Recognition
by: Kassab, Hozaifa, et al.
Published: (2024)
by: Kassab, Hozaifa, et al.
Published: (2024)
UAV-Flow Colosseo: A Real-World Benchmark for Flying-on-a-Word UAV Imitation Learning
by: Wang, Xiangyu, et al.
Published: (2025)
by: Wang, Xiangyu, et al.
Published: (2025)
Edge-Optimized Multimodal Learning for UAV Video Understanding via BLIP-2
by: Feng, Yizhan, et al.
Published: (2026)
by: Feng, Yizhan, et al.
Published: (2026)
FireRescue: A UAV-Based Dataset and Enhanced YOLO Model for Object Detection in Fire Rescue Scenes
by: Xu, Qingyu, et al.
Published: (2025)
by: Xu, Qingyu, et al.
Published: (2025)
STRIDE-QA: Visual Question Answering Dataset for Spatiotemporal Reasoning in Urban Driving Scenes
by: Ishihara, Keishi, et al.
Published: (2025)
by: Ishihara, Keishi, et al.
Published: (2025)
RSUD20K: A Dataset for Road Scene Understanding In Autonomous Driving
by: Zunair, Hasib, et al.
Published: (2024)
by: Zunair, Hasib, et al.
Published: (2024)
Chain-of-Frames: Advancing Video Understanding in Multimodal LLMs via Frame-Aware Reasoning
by: Ghazanfari, Sara, et al.
Published: (2025)
by: Ghazanfari, Sara, et al.
Published: (2025)
FreeFly-Thinking : Aligning Chain-of-Thought Reasoning with Continuous UAV Navigation
by: Zhou, Jiaxu, et al.
Published: (2026)
by: Zhou, Jiaxu, et al.
Published: (2026)
OmniHD-Scenes: A Next-Generation Multimodal Dataset for Autonomous Driving
by: Zheng, Lianqing, et al.
Published: (2024)
by: Zheng, Lianqing, et al.
Published: (2024)
Microscopic Vehicle Trajectory Datasets from UAV-collected Video for Heterogeneous, Area-Based Urban Traffic
by: Ali, Yawar, et al.
Published: (2025)
by: Ali, Yawar, et al.
Published: (2025)
AVOID: The Adverse Visual Conditions Dataset with Obstacles for Driving Scene Understanding
by: Jeong, Jongoh, et al.
Published: (2025)
by: Jeong, Jongoh, et al.
Published: (2025)
Similar Items
-
Federated Medical Image Classification under Class and Domain Imbalance exploiting Synthetic Sample Generation
by: Pavan, Martina, et al.
Published: (2026) -
When Cars meet Drones: Hyperbolic Federated Learning for Source-Free Domain Adaptation in Adverse Weather
by: Rizzoli, Giulia, et al.
Published: (2024) -
FedPromo: Federated Lightweight Proxy Models at the Edge Bring New Domains to Foundation Models
by: Caligiuri, Matteo, et al.
Published: (2025) -
Continual Road-Scene Semantic Segmentation via Feature-Aligned Symmetric Multi-Modal Network
by: Barbato, Francesco, et al.
Published: (2023) -
NIGHT -- Non-Line-of-Sight Imaging from Indirect Time of Flight Data
by: Caligiuri, Matteo, et al.
Published: (2024)