Quantifying the synthetic and real domain gap in aerial scene understanding
Fuente:
arXiv
Saved in:
| Main Author: | Marcu, Alina |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Transferring disentangled representations: bridging the gap between synthetic and real images
by: Dapueto, Jacopo, et al.
Published: (2024)
by: Dapueto, Jacopo, et al.
Published: (2024)
Assessing the generalization performance of SAM for ureteroscopy scene understanding
by: Villagrana, Martin, et al.
Published: (2025)
by: Villagrana, Martin, et al.
Published: (2025)
Generating metamers of human scene understanding
by: Raina, Ritik, et al.
Published: (2026)
by: Raina, Ritik, et al.
Published: (2026)
Non-verbal Real-time Human-AI Interaction in Constrained Robotic Environments
by: Costea, Dragos, et al.
Published: (2026)
by: Costea, Dragos, et al.
Published: (2026)
Improving fine-grained understanding in image-text pre-training
by: Bica, Ioana, et al.
Published: (2024)
by: Bica, Ioana, et al.
Published: (2024)
VideoDirectorGPT: Consistent Multi-scene Video Generation via LLM-Guided Planning
by: Lin, Han, et al.
Published: (2023)
by: Lin, Han, et al.
Published: (2023)
Towards Understanding and Quantifying Uncertainty for Text-to-Image Generation
by: Franchi, Gianni, et al.
Published: (2024)
by: Franchi, Gianni, et al.
Published: (2024)
Quantifying Deep Learning Model Uncertainty in Conformal Prediction
by: Karimi, Hamed, et al.
Published: (2023)
by: Karimi, Hamed, et al.
Published: (2023)
Near, far: Patch-ordering enhances vision foundation models' scene understanding
by: Pariza, Valentinos, et al.
Published: (2024)
by: Pariza, Valentinos, et al.
Published: (2024)
Multi-Task Learning with Multi-Annotation Triplet Loss for Improved Object Detection
by: Zhou, Meilun, et al.
Published: (2025)
by: Zhou, Meilun, et al.
Published: (2025)
Beyond the Veil of Similarity: Quantifying Semantic Continuity in Explainable AI
by: Huang, Qi, et al.
Published: (2024)
by: Huang, Qi, et al.
Published: (2024)
The devil is in the fine-grained details: Evaluating open-vocabulary object detectors for fine-grained understanding
by: Bianchi, Lorenzo, et al.
Published: (2023)
by: Bianchi, Lorenzo, et al.
Published: (2023)
RESQUE: Quantifying Estimator to Task and Distribution Shift for Sustainable Model Reusability
by: Sangarya, Vishwesh, et al.
Published: (2024)
by: Sangarya, Vishwesh, et al.
Published: (2024)
Beyond Morphology: Quantifying the Diagnostic Power of Color Features in Cancer Classification
by: Kheiri, Farnaz, et al.
Published: (2026)
by: Kheiri, Farnaz, et al.
Published: (2026)
Technical Note: Defining and Quantifying AND-OR Interactions for Faithful and Concise Explanation of DNNs
by: Li, Mingjie, et al.
Published: (2023)
by: Li, Mingjie, et al.
Published: (2023)
Deformable ProtoPNet: An Interpretable Image Classifier Using Deformable Prototypes
by: Donnelly, Jon, et al.
Published: (2021)
by: Donnelly, Jon, et al.
Published: (2021)
Face Density as a Proxy for Data Complexity: Quantifying the Hardness of Instance Count
by: Mohammadi-Seif, Abolfazl, et al.
Published: (2026)
by: Mohammadi-Seif, Abolfazl, et al.
Published: (2026)
Do generative video models understand physical principles?
by: Motamed, Saman, et al.
Published: (2025)
by: Motamed, Saman, et al.
Published: (2025)
Automated rock joint trace mapping using a supervised learning model trained on synthetic data generated by parametric modelling
by: Chiu, Jessica Ka Yi, et al.
Published: (2026)
by: Chiu, Jessica Ka Yi, et al.
Published: (2026)
Co-domain Symmetry for Complex-Valued Deep Learning
by: Singhal, Utkarsh, et al.
Published: (2021)
by: Singhal, Utkarsh, et al.
Published: (2021)
Rethinking Multi-domain Generalization with A General Learning Objective
by: Tan, Zhaorui, et al.
Published: (2024)
by: Tan, Zhaorui, et al.
Published: (2024)
Continual learning under domain transfer with sparse synaptic bursting
by: Beaulieu, Shawn L., et al.
Published: (2021)
by: Beaulieu, Shawn L., et al.
Published: (2021)
Using Shapley interactions to understand how models use structure
by: Singhvi, Divyansh, et al.
Published: (2024)
by: Singhvi, Divyansh, et al.
Published: (2024)
FedVG: Gradient-Guided Aggregation for Enhanced Federated Learning
by: Devkota, Alina, et al.
Published: (2026)
by: Devkota, Alina, et al.
Published: (2026)
UAV traffic scene understanding: A regulation embedded multi-modal network and a unified benchmark
by: Zhang, Yu, et al.
Published: (2026)
by: Zhang, Yu, et al.
Published: (2026)
Analyzing domain shift when using additional data for the MICCAI KiTS23 Challenge
by: Stoica, George, et al.
Published: (2023)
by: Stoica, George, et al.
Published: (2023)
Quantifying In-Context Reasoning Effects and Memorization Effects in LLMs
by: Lou, Siyu, et al.
Published: (2024)
by: Lou, Siyu, et al.
Published: (2024)
Technical Report: Quantifying and Analyzing the Generalization Power of a DNN
by: He, Yuxuan, et al.
Published: (2025)
by: He, Yuxuan, et al.
Published: (2025)
D4: Text-guided diffusion model-based domain adaptive data augmentation for vineyard shoot detection
by: Hirahara, Kentaro, et al.
Published: (2024)
by: Hirahara, Kentaro, et al.
Published: (2024)
Planktonzilla: Multimodal dataset and models for understanding plankton ecosystems
by: Montanares, Alan Gerson Contreras, et al.
Published: (2026)
by: Montanares, Alan Gerson Contreras, et al.
Published: (2026)
Multi-domain performance analysis with scores tailored to user preferences
by: Piérard, Sébastien, et al.
Published: (2025)
by: Piérard, Sébastien, et al.
Published: (2025)
This Looks Better than That: Better Interpretable Models with ProtoPNeXt
by: Willard, Frank, et al.
Published: (2024)
by: Willard, Frank, et al.
Published: (2024)
Confidence intervals uncovered: Are we ready for real-world medical imaging AI?
by: Christodoulou, Evangelia, et al.
Published: (2024)
by: Christodoulou, Evangelia, et al.
Published: (2024)
Extracting real estate values of rental apartment floor plans using graph convolutional networks
by: Takizawa, Atsushi
Published: (2023)
by: Takizawa, Atsushi
Published: (2023)
Bridging vision language model (VLM) evaluation gaps with a framework for scalable and cost-effective benchmark generation
by: Rädsch, Tim, et al.
Published: (2025)
by: Rädsch, Tim, et al.
Published: (2025)
Exploring the Adversarial Frontier: Quantifying Robustness via Adversarial Hypervolume
by: Guo, Ping, et al.
Published: (2024)
by: Guo, Ping, et al.
Published: (2024)
Semantic Approach to Quantifying the Consistency of Diffusion Model Image Generation
by: Bent, Brinnae
Published: (2024)
by: Bent, Brinnae
Published: (2024)
DEF-oriCORN: efficient 3D scene understanding for robust language-directed manipulation without demonstrations
by: Son, Dongwon, et al.
Published: (2024)
by: Son, Dongwon, et al.
Published: (2024)
SciVid: Cross-Domain Evaluation of Video Models in Scientific Applications
by: Hasson, Yana, et al.
Published: (2025)
by: Hasson, Yana, et al.
Published: (2025)
Towards Predicting the Success of Transfer-based Attacks by Quantifying Shared Feature Representations
by: Dale, Ashley S., et al.
Published: (2024)
by: Dale, Ashley S., et al.
Published: (2024)
Similar Items
-
Transferring disentangled representations: bridging the gap between synthetic and real images
by: Dapueto, Jacopo, et al.
Published: (2024) -
Assessing the generalization performance of SAM for ureteroscopy scene understanding
by: Villagrana, Martin, et al.
Published: (2025) -
Generating metamers of human scene understanding
by: Raina, Ritik, et al.
Published: (2026) -
Non-verbal Real-time Human-AI Interaction in Constrained Robotic Environments
by: Costea, Dragos, et al.
Published: (2026) -
Improving fine-grained understanding in image-text pre-training
by: Bica, Ioana, et al.
Published: (2024)