Enhancing Effectiveness and Robustness in a Low-Resource Regime via Decision-Boundary-aware Data Augmentation
Fuente:
arXiv
Saved in:
| Main Authors: | Jin, Kyohoon, Lee, Junho, Choi, Juhwan, Song, Sangmin, Kim, Youngbin |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
AutoAugment Is What You Need: Enhancing Rule-based Augmentation Methods in Low-resource Regimes
by: Choi, Juhwan, et al.
Published: (2024)
by: Choi, Juhwan, et al.
Published: (2024)
SoftEDA: Rethinking Rule-Based Data Augmentation with Soft Labels
by: Choi, Juhwan, et al.
Published: (2024)
by: Choi, Juhwan, et al.
Published: (2024)
CoBA: Counterbias Text Augmentation for Mitigating Various Spurious Correlations via Semantic Triples
by: Jin, Kyohoon, et al.
Published: (2025)
by: Jin, Kyohoon, et al.
Published: (2025)
Multi-News+: Cost-efficient Dataset Cleansing via LLM-based Data Annotation
by: Choi, Juhwan, et al.
Published: (2024)
by: Choi, Juhwan, et al.
Published: (2024)
GPTs Are Multilingual Annotators for Sequence Generation Tasks
by: Choi, Juhwan, et al.
Published: (2024)
by: Choi, Juhwan, et al.
Published: (2024)
Plug-in and Fine-tuning: Bridging the Gap between Small Language Models and Large Language Models
by: Kim, Kyeonghyun, et al.
Published: (2025)
by: Kim, Kyeonghyun, et al.
Published: (2025)
Adverb Is the Key: Simple Text Data Augmentation with Adverb Deletion
by: Choi, Juhwan, et al.
Published: (2024)
by: Choi, Juhwan, et al.
Published: (2024)
Beyond Single-User Dialogue: Assessing Multi-User Dialogue State Tracking Capabilities of Large Language Models
by: Song, Sangmin, et al.
Published: (2025)
by: Song, Sangmin, et al.
Published: (2025)
Strategic Data Ordering: Enhancing Large Language Model Performance through Curriculum Learning
by: Kim, Jisu, et al.
Published: (2024)
by: Kim, Jisu, et al.
Published: (2024)
SAGE-LD: Towards Scalable and Generalizable End-to-End Language Diarization via Simulated Data Augmentation
by: Lee, Sangmin, et al.
Published: (2025)
by: Lee, Sangmin, et al.
Published: (2025)
IM-BERT: Enhancing Robustness of BERT through the Implicit Euler Method
by: Kim, Mihyeon, et al.
Published: (2025)
by: Kim, Mihyeon, et al.
Published: (2025)
GRADE: Generating multi-hop QA and fine-gRAined Difficulty matrix for RAG Evaluation
by: Lee, Jeongsoo, et al.
Published: (2025)
by: Lee, Jeongsoo, et al.
Published: (2025)
Control Token with Dense Passage Retrieval
by: Lee, Juhwan, et al.
Published: (2024)
by: Lee, Juhwan, et al.
Published: (2024)
Colorful Cutout: Enhancing Image Data Augmentation with Curriculum Learning
by: Choi, Juhwan, et al.
Published: (2024)
by: Choi, Juhwan, et al.
Published: (2024)
FairQE: Multi-Agent Framework for Mitigating Gender Bias in Translation Quality Estimation
by: Jang, Jinhee, et al.
Published: (2026)
by: Jang, Jinhee, et al.
Published: (2026)
Focus on the Core: Efficient Attention via Pruned Token Compression for Document Classification
by: Yun, Jungmin, et al.
Published: (2024)
by: Yun, Jungmin, et al.
Published: (2024)
Aligning Extraction and Generation for Robust Retrieval-Augmented Generation
by: Song, Hwanjun, et al.
Published: (2025)
by: Song, Hwanjun, et al.
Published: (2025)
An Effective Deployment of Diffusion LM for Data Augmentation in Low-Resource Sentiment Classification
by: Chen, Zhuowei, et al.
Published: (2024)
by: Chen, Zhuowei, et al.
Published: (2024)
Enhancing Robustness of Retrieval-Augmented Language Models with In-Context Learning
by: Park, Seong-Il, et al.
Published: (2024)
by: Park, Seong-Il, et al.
Published: (2024)
The Effectiveness of Morphology-aware Segmentation in Low-Resource Neural Machine Translation
by: Sälevä, Jonne, et al.
Published: (2021)
by: Sälevä, Jonne, et al.
Published: (2021)
Evaluating the Effectiveness of Data Augmentation for Emotion Classification in Low-Resource Settings
by: Arora, Aashish, et al.
Published: (2024)
by: Arora, Aashish, et al.
Published: (2024)
Delving into Multilingual Ethical Bias: The MSQAD with Statistical Hypothesis Tests for Large Language Models
by: Yu, Seunguk, et al.
Published: (2025)
by: Yu, Seunguk, et al.
Published: (2025)
Don't be a Fool: Pooling Strategies in Offensive Language Detection from User-Intended Adversarial Attacks
by: Yu, Seunguk, et al.
Published: (2024)
by: Yu, Seunguk, et al.
Published: (2024)
SNaRe: Domain-aware Data Generation for Low-Resource Event Detection
by: Parekh, Tanmay, et al.
Published: (2025)
by: Parekh, Tanmay, et al.
Published: (2025)
Rethinking what Matters: Effective and Robust Multilingual Realignment for Low-Resource Languages
by: Nguyen, Quang Phuoc, et al.
Published: (2025)
by: Nguyen, Quang Phuoc, et al.
Published: (2025)
SALAD: Improving Robustness and Generalization through Contrastive Learning with Structure-Aware and LLM-Driven Augmented Data
by: Bae, Suyoung, et al.
Published: (2025)
by: Bae, Suyoung, et al.
Published: (2025)
UniGen: Universal Domain Generalization for Sentiment Classification via Zero-shot Dataset Generation
by: Choi, Juhwan, et al.
Published: (2024)
by: Choi, Juhwan, et al.
Published: (2024)
MELT: Materials-aware Continued Pre-training for Language Model Adaptation to Materials Science
by: Kim, Junho, et al.
Published: (2024)
by: Kim, Junho, et al.
Published: (2024)
Enhancing Low-Resource Minority Language Translation with LLMs and Retrieval-Augmented Generation for Cultural Nuances
by: Chang, Chen-Chi, et al.
Published: (2025)
by: Chang, Chen-Chi, et al.
Published: (2025)
ABEX: Data Augmentation for Low-Resource NLU via Expanding Abstract Descriptions
by: Ghosh, Sreyan, et al.
Published: (2024)
by: Ghosh, Sreyan, et al.
Published: (2024)
A Guide To Effectively Leveraging LLMs for Low-Resource Text Summarization: Data Augmentation and Semi-supervised Approaches
by: Sahu, Gaurav, et al.
Published: (2024)
by: Sahu, Gaurav, et al.
Published: (2024)
Retrieval-Augmented Data Augmentation for Low-Resource Domain Tasks
by: Seo, Minju, et al.
Published: (2024)
by: Seo, Minju, et al.
Published: (2024)
Medal Matters: Probing LLMs' Failure Cases Through Olympic Rankings
by: Choi, Juhwan, et al.
Published: (2024)
by: Choi, Juhwan, et al.
Published: (2024)
GRAM: Generative Recommendation via Semantic-aware Multi-granular Late Fusion
by: Lee, Sunkyung, et al.
Published: (2025)
by: Lee, Sunkyung, et al.
Published: (2025)
Mentor-KD: Making Small Language Models Better Multi-step Reasoners
by: Lee, Hojae, et al.
Published: (2024)
by: Lee, Hojae, et al.
Published: (2024)
Conditional Semi-Supervised Data Augmentation for Spam Message Detection with Low Resource Data
by: Nuha, Ulin, et al.
Published: (2024)
by: Nuha, Ulin, et al.
Published: (2024)
SUMMPILOT: Bridging Efficiency and Customization for Interactive Summarization System
by: Yun, JungMin, et al.
Published: (2026)
by: Yun, JungMin, et al.
Published: (2026)
Targeted Augmentation for Low-Resource Event Extraction
by: Wang, Sijia, et al.
Published: (2024)
by: Wang, Sijia, et al.
Published: (2024)
Transcending Language Boundaries: Harnessing LLMs for Low-Resource Language Translation
by: Shu, Peng, et al.
Published: (2024)
by: Shu, Peng, et al.
Published: (2024)
Rethinking the Reranker: Boundary-Aware Evidence Selection for Robust Retrieval-Augmented Generation
by: Sun, Jiashuo, et al.
Published: (2026)
by: Sun, Jiashuo, et al.
Published: (2026)
Similar Items
-
AutoAugment Is What You Need: Enhancing Rule-based Augmentation Methods in Low-resource Regimes
by: Choi, Juhwan, et al.
Published: (2024) -
SoftEDA: Rethinking Rule-Based Data Augmentation with Soft Labels
by: Choi, Juhwan, et al.
Published: (2024) -
CoBA: Counterbias Text Augmentation for Mitigating Various Spurious Correlations via Semantic Triples
by: Jin, Kyohoon, et al.
Published: (2025) -
Multi-News+: Cost-efficient Dataset Cleansing via LLM-based Data Annotation
by: Choi, Juhwan, et al.
Published: (2024) -
GPTs Are Multilingual Annotators for Sequence Generation Tasks
by: Choi, Juhwan, et al.
Published: (2024)