Region Mixup
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Saha, Saptarshi, Garain, Utpal |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Deep Learning for automated multi-scale functional field boundaries extraction using multi-date Sentinel-2 and PlanetScope imagery: Case Study of Netherlands and Pakistan
von: Zahid, Saba, et al.
Veröffentlicht: (2024)
von: Zahid, Saba, et al.
Veröffentlicht: (2024)
Mixture of Rationale: Multi-Modal Reasoning Mixture for Visual Question Answering
von: Li, Tao, et al.
Veröffentlicht: (2024)
von: Li, Tao, et al.
Veröffentlicht: (2024)
Emotion Recognition and Generation: A Comprehensive Review of Face, Speech, and Text Modalities
von: Mobbs, Rebecca, et al.
Veröffentlicht: (2025)
von: Mobbs, Rebecca, et al.
Veröffentlicht: (2025)
Rotation-Adaptive Point Cloud Domain Generalization via Intricate Orientation Learning
von: Liu, Bangzhen, et al.
Veröffentlicht: (2025)
von: Liu, Bangzhen, et al.
Veröffentlicht: (2025)
Neuron-based explanations of neural networks sacrifice completeness and interpretability
von: Dey, Nolan, et al.
Veröffentlicht: (2020)
von: Dey, Nolan, et al.
Veröffentlicht: (2020)
A Multi-Modal Deep Learning Based Approach for House Price Prediction
von: Hasan, Md Hasebul, et al.
Veröffentlicht: (2024)
von: Hasan, Md Hasebul, et al.
Veröffentlicht: (2024)
Representation Learning via Non-Contrastive Mutual Information
von: Guo, Zhaohan Daniel, et al.
Veröffentlicht: (2025)
von: Guo, Zhaohan Daniel, et al.
Veröffentlicht: (2025)
Visual Language Models show widespread visual deficits on neuropsychological tests
von: Tangtartharakul, Gene, et al.
Veröffentlicht: (2025)
von: Tangtartharakul, Gene, et al.
Veröffentlicht: (2025)
Reasoning is a Modality
von: Liu, Zhiguang, et al.
Veröffentlicht: (2026)
von: Liu, Zhiguang, et al.
Veröffentlicht: (2026)
CausAdv: A Causal-based Framework for Detecting Adversarial Examples
von: Debbi, Hichem
Veröffentlicht: (2024)
von: Debbi, Hichem
Veröffentlicht: (2024)
Leveraging Color Channel Independence for Improved Unsupervised Object Detection
von: Jäckl, Bastian, et al.
Veröffentlicht: (2024)
von: Jäckl, Bastian, et al.
Veröffentlicht: (2024)
Data Augmentation for Image Classification using Generative AI
von: Rahat, Fazle, et al.
Veröffentlicht: (2024)
von: Rahat, Fazle, et al.
Veröffentlicht: (2024)
Data Efficiency and Transfer Robustness in Biomedical Image Segmentation: A Study of Redundancy and Forgetting with Cellpose
von: Zhao, Shuo, et al.
Veröffentlicht: (2025)
von: Zhao, Shuo, et al.
Veröffentlicht: (2025)
Hybrid Knowledge Transfer through Attention and Logit Distillation for On-Device Vision Systems in Agricultural IoT
von: Mugisha, Stanley, et al.
Veröffentlicht: (2025)
von: Mugisha, Stanley, et al.
Veröffentlicht: (2025)
Enhancing Long-Term Re-Identification Robustness Using Synthetic Data: A Comparative Analysis
von: Pionzewski, Christian, et al.
Veröffentlicht: (2025)
von: Pionzewski, Christian, et al.
Veröffentlicht: (2025)
Rethinking Multimodal Few-Shot 3D Point Cloud Segmentation: From Fused Refinement to Decoupled Arbitration
von: Bian, Wentao, et al.
Veröffentlicht: (2026)
von: Bian, Wentao, et al.
Veröffentlicht: (2026)
Rethinking Uncertainty in Segmentation: From Estimation to Decision
von: Maganti, Saket
Veröffentlicht: (2026)
von: Maganti, Saket
Veröffentlicht: (2026)
Video-based Exercise Classification and Activated Muscle Group Prediction with Hybrid X3D-SlowFast Network
von: Pasula, Manvik, et al.
Veröffentlicht: (2024)
von: Pasula, Manvik, et al.
Veröffentlicht: (2024)
Transformers Get Stable: An End-to-End Signal Propagation Theory for Language Models
von: Kedia, Akhil, et al.
Veröffentlicht: (2024)
von: Kedia, Akhil, et al.
Veröffentlicht: (2024)
CoachMe: Decoding Sport Elements with a Reference-Based Coaching Instruction Generation Model
von: Yeh, Wei-Hsin, et al.
Veröffentlicht: (2025)
von: Yeh, Wei-Hsin, et al.
Veröffentlicht: (2025)
E = T*H/(O+B): A Dimensionless Control Parameter for Mixture-of-Experts Ecology
von: Zhang, Qingjun
Veröffentlicht: (2026)
von: Zhang, Qingjun
Veröffentlicht: (2026)
A Spitting Image: Modular Superpixel Tokenization in Vision Transformers
von: Aasan, Marius, et al.
Veröffentlicht: (2024)
von: Aasan, Marius, et al.
Veröffentlicht: (2024)
AMANet: Advancing SAR Ship Detection with Adaptive Multi-Hierarchical Attention Network
von: Ma, Xiaolin, et al.
Veröffentlicht: (2024)
von: Ma, Xiaolin, et al.
Veröffentlicht: (2024)
SAG-ViT: A Scale-Aware, High-Fidelity Patching Approach with Graph Attention for Vision Transformers
von: Venkatraman, Shravan, et al.
Veröffentlicht: (2024)
von: Venkatraman, Shravan, et al.
Veröffentlicht: (2024)
Aligning by Misaligning: Boundary-aware Curriculum Learning for Multimodal Alignment
von: Ye, Hua, et al.
Veröffentlicht: (2025)
von: Ye, Hua, et al.
Veröffentlicht: (2025)
MFTF: Mask-free Training-free Object Level Layout Control Diffusion Model
von: Yang, Shan
Veröffentlicht: (2024)
von: Yang, Shan
Veröffentlicht: (2024)
Sora as a World Model? A Complete Survey on Text-to-Video Generation
von: Puspitasari, Fachrina Dewi, et al.
Veröffentlicht: (2024)
von: Puspitasari, Fachrina Dewi, et al.
Veröffentlicht: (2024)
Disrupting Diffusion: Token-Level Attention Erasure Attack against Diffusion-based Customization
von: Liu, Yisu, et al.
Veröffentlicht: (2024)
von: Liu, Yisu, et al.
Veröffentlicht: (2024)
Unified Auto-Encoding with Masked Diffusion
von: Hansen-Estruch, Philippe, et al.
Veröffentlicht: (2024)
von: Hansen-Estruch, Philippe, et al.
Veröffentlicht: (2024)
Demo-Pose: Depth-Monocular Modality Fusion For Object Pose Estimation
von: Agarwal, Rachit, et al.
Veröffentlicht: (2026)
von: Agarwal, Rachit, et al.
Veröffentlicht: (2026)
SITUATE -- Synthetic Object Counting Dataset for VLM training
von: Peinl, René, et al.
Veröffentlicht: (2026)
von: Peinl, René, et al.
Veröffentlicht: (2026)
Think with 3D: Geometric Imagination Grounded Spatial Reasoning from Limited Views
von: Chen, Zhangquan, et al.
Veröffentlicht: (2025)
von: Chen, Zhangquan, et al.
Veröffentlicht: (2025)
Image Segmentation and Classification of E-waste for Training Robots for Waste Segregation
von: Tripathi, Prakriti
Veröffentlicht: (2025)
von: Tripathi, Prakriti
Veröffentlicht: (2025)
ProtoFlow: Interpretable and Robust Surgical Workflow Modeling with Learned Dynamic Scene Graph Prototypes
von: Holm, Felix, et al.
Veröffentlicht: (2025)
von: Holm, Felix, et al.
Veröffentlicht: (2025)
Siamese Networks for Cat Re-Identification: Exploring Neural Models for Cat Instance Recognition
von: Trein, Tobias, et al.
Veröffentlicht: (2025)
von: Trein, Tobias, et al.
Veröffentlicht: (2025)
Evaluation of Environmental Conditions on Object Detection using Oriented Bounding Boxes for AR Applications
von: Li, Vladislav, et al.
Veröffentlicht: (2023)
von: Li, Vladislav, et al.
Veröffentlicht: (2023)
Appearance-based gaze estimation enhanced with synthetic images using deep neural networks
von: Herashchenko, Dmytro, et al.
Veröffentlicht: (2023)
von: Herashchenko, Dmytro, et al.
Veröffentlicht: (2023)
From Prompt to Production:Automating Brand-Safe Marketing Imagery with Text-to-Image Models
von: Atighehchian, Parmida, et al.
Veröffentlicht: (2026)
von: Atighehchian, Parmida, et al.
Veröffentlicht: (2026)
Attentive VQ-VAE
von: Hoyos, Angello, et al.
Veröffentlicht: (2023)
von: Hoyos, Angello, et al.
Veröffentlicht: (2023)
TexTailor: Customized Text-aligned Texturing via Effective Resampling
von: Lee, Suin, et al.
Veröffentlicht: (2025)
von: Lee, Suin, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Deep Learning for automated multi-scale functional field boundaries extraction using multi-date Sentinel-2 and PlanetScope imagery: Case Study of Netherlands and Pakistan
von: Zahid, Saba, et al.
Veröffentlicht: (2024) -
Mixture of Rationale: Multi-Modal Reasoning Mixture for Visual Question Answering
von: Li, Tao, et al.
Veröffentlicht: (2024) -
Emotion Recognition and Generation: A Comprehensive Review of Face, Speech, and Text Modalities
von: Mobbs, Rebecca, et al.
Veröffentlicht: (2025) -
Rotation-Adaptive Point Cloud Domain Generalization via Intricate Orientation Learning
von: Liu, Bangzhen, et al.
Veröffentlicht: (2025) -
Neuron-based explanations of neural networks sacrifice completeness and interpretability
von: Dey, Nolan, et al.
Veröffentlicht: (2020)