Colorful Cutout: Enhancing Image Data Augmentation with Curriculum Learning
Fuente:
arXiv
Salvato in:
| Autori principali: | Choi, Juhwan, Kim, YoungBin |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
VolDoGer: LLM-assisted Datasets for Domain Generalization in Vision-Language Tasks
di: Choi, Juhwan, et al.
Pubblicazione: (2024)
di: Choi, Juhwan, et al.
Pubblicazione: (2024)
Adverb Is the Key: Simple Text Data Augmentation with Adverb Deletion
di: Choi, Juhwan, et al.
Pubblicazione: (2024)
di: Choi, Juhwan, et al.
Pubblicazione: (2024)
Learning to Detour: Shortcut Mitigating Augmentation for Weakly Supervised Semantic Segmentation
di: Kwon, JuneHyoung, et al.
Pubblicazione: (2024)
di: Kwon, JuneHyoung, et al.
Pubblicazione: (2024)
Erase Persona, Forget Lore: Benchmarking Multimodal Copyright Unlearning in Large Vision Language Models
di: Kwon, JuneHyoung, et al.
Pubblicazione: (2026)
di: Kwon, JuneHyoung, et al.
Pubblicazione: (2026)
VG-CoT: Towards Trustworthy Visual Reasoning via Grounded Chain-of-Thought
di: Lim, Byeonggeuk, et al.
Pubblicazione: (2026)
di: Lim, Byeonggeuk, et al.
Pubblicazione: (2026)
Before Forgetting, Learn to Remember: Revisiting Foundational Learning Failures in LVLM Unlearning Benchmarks
di: Kwon, JuneHyoung, et al.
Pubblicazione: (2026)
di: Kwon, JuneHyoung, et al.
Pubblicazione: (2026)
See-Saw Modality Balance: See Gradient, and Sew Impaired Vision-Language Balance to Mitigate Dominant Modality Bias
di: Kwon, JuneHyoung, et al.
Pubblicazione: (2025)
di: Kwon, JuneHyoung, et al.
Pubblicazione: (2025)
Multi-News+: Cost-efficient Dataset Cleansing via LLM-based Data Annotation
di: Choi, Juhwan, et al.
Pubblicazione: (2024)
di: Choi, Juhwan, et al.
Pubblicazione: (2024)
Diffusion Curriculum: Synthetic-to-Real Data Curriculum via Image-Guided Diffusion
di: Liang, Yijun, et al.
Pubblicazione: (2024)
di: Liang, Yijun, et al.
Pubblicazione: (2024)
A Training-Free, Task-Agnostic Framework for Enhancing MLLM Performance on High-Resolution Images
di: Lee, Jaeseong, et al.
Pubblicazione: (2025)
di: Lee, Jaeseong, et al.
Pubblicazione: (2025)
CA-Cut: Crop-Aligned Cutout for Data Augmentation to Learn More Robust Under-Canopy Navigation
di: Mamo, Robel, et al.
Pubblicazione: (2025)
di: Mamo, Robel, et al.
Pubblicazione: (2025)
AH-OCDA: Amplitude-based Curriculum Learning and Hopfield Segmentation Model for Open Compound Domain Adaptation
di: Choi, Jaehyun, et al.
Pubblicazione: (2024)
di: Choi, Jaehyun, et al.
Pubblicazione: (2024)
Inside Knowledge: Graph-based Path Generation with Explainable Data Augmentation and Curriculum Learning for Visual Indoor Navigation
di: Airinei, Daniel, et al.
Pubblicazione: (2025)
di: Airinei, Daniel, et al.
Pubblicazione: (2025)
MAESIL: Masked Autoencoder for Enhanced Self-supervised Medical Image Learning
di: Kim, Kyeonghun, et al.
Pubblicazione: (2026)
di: Kim, Kyeonghun, et al.
Pubblicazione: (2026)
Test-Time Mixup Augmentation for Data and Class-Specific Uncertainty Estimation in Deep Learning Image Classification
di: Lee, Hansang, et al.
Pubblicazione: (2022)
di: Lee, Hansang, et al.
Pubblicazione: (2022)
GPTs Are Multilingual Annotators for Sequence Generation Tasks
di: Choi, Juhwan, et al.
Pubblicazione: (2024)
di: Choi, Juhwan, et al.
Pubblicazione: (2024)
Enhancing Contrastive Learning for Retinal Imaging via Adjusted Augmentation Scales
di: Cheng, Zijie, et al.
Pubblicazione: (2025)
di: Cheng, Zijie, et al.
Pubblicazione: (2025)
SGD-Mix: Enhancing Domain-Specific Image Classification with Label-Preserving Data Augmentation
di: Dong, Yixuan, et al.
Pubblicazione: (2025)
di: Dong, Yixuan, et al.
Pubblicazione: (2025)
Gaussian Mixture Proposals with Pull-Push Learning Scheme to Capture Diverse Events for Weakly Supervised Temporal Video Grounding
di: Kim, Sunoh, et al.
Pubblicazione: (2023)
di: Kim, Sunoh, et al.
Pubblicazione: (2023)
Decoupled Data Augmentation for Improving Image Classification
di: Chen, Ruoxin, et al.
Pubblicazione: (2024)
di: Chen, Ruoxin, et al.
Pubblicazione: (2024)
Towards General Deepfake Detection with Dynamic Curriculum
di: Song, Wentang, et al.
Pubblicazione: (2024)
di: Song, Wentang, et al.
Pubblicazione: (2024)
Medal Matters: Probing LLMs' Failure Cases Through Olympic Rankings
di: Choi, Juhwan, et al.
Pubblicazione: (2024)
di: Choi, Juhwan, et al.
Pubblicazione: (2024)
Beyond Single-User Dialogue: Assessing Multi-User Dialogue State Tracking Capabilities of Large Language Models
di: Song, Sangmin, et al.
Pubblicazione: (2025)
di: Song, Sangmin, et al.
Pubblicazione: (2025)
The Effects of Mixed Sample Data Augmentation are Class Dependent
di: Lee, Haeil, et al.
Pubblicazione: (2023)
di: Lee, Haeil, et al.
Pubblicazione: (2023)
KeepOriginalAugment: Single Image-based Better Information-Preserving Data Augmentation Approach
di: Kumar, Teerath, et al.
Pubblicazione: (2024)
di: Kumar, Teerath, et al.
Pubblicazione: (2024)
Anatomy-R1: Enhancing Anatomy Reasoning in Multimodal Large Language Models via Anatomical Similarity Curriculum and Group Diversity Augmentation
di: Song, Ziyang, et al.
Pubblicazione: (2025)
di: Song, Ziyang, et al.
Pubblicazione: (2025)
Learning to Detect Multi-class Anomalies with Just One Normal Image Prompt
di: Gao, Bin-Bin
Pubblicazione: (2025)
di: Gao, Bin-Bin
Pubblicazione: (2025)
Distribution-Aware Robust Learning from Long-Tailed Data with Noisy Labels
di: Baik, Jae Soon, et al.
Pubblicazione: (2024)
di: Baik, Jae Soon, et al.
Pubblicazione: (2024)
Enhancing Generalization in Data-free Quantization via Mixup-class Prompting
di: Park, Jiwoong, et al.
Pubblicazione: (2025)
di: Park, Jiwoong, et al.
Pubblicazione: (2025)
MoST: Motion Style Transformer between Diverse Action Contents
di: Kim, Boeun, et al.
Pubblicazione: (2024)
di: Kim, Boeun, et al.
Pubblicazione: (2024)
Curriculum Prompting Foundation Models for Medical Image Segmentation
di: Zheng, Xiuqi, et al.
Pubblicazione: (2024)
di: Zheng, Xiuqi, et al.
Pubblicazione: (2024)
A Text-Image Fusion Method with Data Augmentation Capabilities for Referring Medical Image Segmentation
di: Chai, Shurong, et al.
Pubblicazione: (2025)
di: Chai, Shurong, et al.
Pubblicazione: (2025)
Genetic Learning for Designing Sim-to-Real Data Augmentations
di: Vanherle, Bram, et al.
Pubblicazione: (2024)
di: Vanherle, Bram, et al.
Pubblicazione: (2024)
Enhancing Image Quality Assessment Ability of LMMs via Retrieval-Augmented Generation
di: Fu, Kang, et al.
Pubblicazione: (2026)
di: Fu, Kang, et al.
Pubblicazione: (2026)
Enhancing Supervised Composed Image Retrieval via Reasoning-Augmented Representation Engineering
di: Li, Jun, et al.
Pubblicazione: (2025)
di: Li, Jun, et al.
Pubblicazione: (2025)
Transforming Color: A Novel Image Colorization Method
di: Shafiq, Hamza, et al.
Pubblicazione: (2024)
di: Shafiq, Hamza, et al.
Pubblicazione: (2024)
FSCM: Frequency-Enhanced Spatial-Spectral Coupled Mamba for Infrared Hyperspectral Image Colorization
di: Liu, Tingting, et al.
Pubblicazione: (2026)
di: Liu, Tingting, et al.
Pubblicazione: (2026)
RaDL: Relation-aware Disentangled Learning for Multi-Instance Text-to-Image Generation
di: Park, Geon, et al.
Pubblicazione: (2025)
di: Park, Geon, et al.
Pubblicazione: (2025)
Cross-Patient Pseudo Bags Generation and Curriculum Contrastive Learning for Imbalanced Multiclassification of Whole Slide Image
di: Wu, Yonghuang, et al.
Pubblicazione: (2024)
di: Wu, Yonghuang, et al.
Pubblicazione: (2024)
Fourier-Guided Attention Upsampling for Image Super-Resolution
di: Choi, Daejune, et al.
Pubblicazione: (2025)
di: Choi, Daejune, et al.
Pubblicazione: (2025)
Documenti analoghi
-
VolDoGer: LLM-assisted Datasets for Domain Generalization in Vision-Language Tasks
di: Choi, Juhwan, et al.
Pubblicazione: (2024) -
Adverb Is the Key: Simple Text Data Augmentation with Adverb Deletion
di: Choi, Juhwan, et al.
Pubblicazione: (2024) -
Learning to Detour: Shortcut Mitigating Augmentation for Weakly Supervised Semantic Segmentation
di: Kwon, JuneHyoung, et al.
Pubblicazione: (2024) -
Erase Persona, Forget Lore: Benchmarking Multimodal Copyright Unlearning in Large Vision Language Models
di: Kwon, JuneHyoung, et al.
Pubblicazione: (2026) -
VG-CoT: Towards Trustworthy Visual Reasoning via Grounded Chain-of-Thought
di: Lim, Byeonggeuk, et al.
Pubblicazione: (2026)