CoDa: Constrained Generation based Data Augmentation for Low-Resource NLP
Fuente:
arXiv
Salvato in:
| Autori principali: | Evuru, Chandra Kiran Reddy, Ghosh, Sreyan, Kumar, Sonal, S, Ramaneswaran, Tyagi, Utkarsh, Manocha, Dinesh |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
ABEX: Data Augmentation for Low-Resource NLU via Expanding Abstract Descriptions
di: Ghosh, Sreyan, et al.
Pubblicazione: (2024)
di: Ghosh, Sreyan, et al.
Pubblicazione: (2024)
ASPIRE: Language-Guided Data Augmentation for Improving Robustness Against Spurious Correlations
di: Ghosh, Sreyan, et al.
Pubblicazione: (2023)
di: Ghosh, Sreyan, et al.
Pubblicazione: (2023)
RECAP: Retrieval-Augmented Audio Captioning
di: Ghosh, Sreyan, et al.
Pubblicazione: (2023)
di: Ghosh, Sreyan, et al.
Pubblicazione: (2023)
Visual Description Grounding Reduces Hallucinations and Boosts Reasoning in LVLMs
di: Ghosh, Sreyan, et al.
Pubblicazione: (2024)
di: Ghosh, Sreyan, et al.
Pubblicazione: (2024)
CompA: Addressing the Gap in Compositional Reasoning in Audio-Language Models
di: Ghosh, Sreyan, et al.
Pubblicazione: (2023)
di: Ghosh, Sreyan, et al.
Pubblicazione: (2023)
A Closer Look at the Limitations of Instruction Tuning
di: Ghosh, Sreyan, et al.
Pubblicazione: (2024)
di: Ghosh, Sreyan, et al.
Pubblicazione: (2024)
Do Vision-Language Models Understand Compound Nouns?
di: Kumar, Sonal, et al.
Pubblicazione: (2024)
di: Kumar, Sonal, et al.
Pubblicazione: (2024)
GAMA: A Large Audio-Language Model with Advanced Audio Understanding and Complex Reasoning Abilities
di: Ghosh, Sreyan, et al.
Pubblicazione: (2024)
di: Ghosh, Sreyan, et al.
Pubblicazione: (2024)
ReCLAP: Improving Zero Shot Audio Classification by Describing Sounds
di: Ghosh, Sreyan, et al.
Pubblicazione: (2024)
di: Ghosh, Sreyan, et al.
Pubblicazione: (2024)
MultiVox: A Benchmark for Evaluating Voice Assistants for Multimodal Interactions
di: Selvakumar, Ramaneswaran, et al.
Pubblicazione: (2025)
di: Selvakumar, Ramaneswaran, et al.
Pubblicazione: (2025)
ProSE: Diffusion Priors for Speech Enhancement
di: Kumar, Sonal, et al.
Pubblicazione: (2025)
di: Kumar, Sonal, et al.
Pubblicazione: (2025)
MMAU: A Massive Multi-Task Audio Understanding and Reasoning Benchmark
di: Sakshi, S, et al.
Pubblicazione: (2024)
di: Sakshi, S, et al.
Pubblicazione: (2024)
EH-MAM: Easy-to-Hard Masked Acoustic Modeling for Self-Supervised Speech Representation Learning
di: Seth, Ashish, et al.
Pubblicazione: (2024)
di: Seth, Ashish, et al.
Pubblicazione: (2024)
LipGER: Visually-Conditioned Generative Error Correction for Robust Automatic Speech Recognition
di: Ghosh, Sreyan, et al.
Pubblicazione: (2024)
di: Ghosh, Sreyan, et al.
Pubblicazione: (2024)
PAT: Parameter-Free Audio-Text Aligner to Boost Zero-Shot Audio Classification
di: Seth, Ashish, et al.
Pubblicazione: (2024)
di: Seth, Ashish, et al.
Pubblicazione: (2024)
Synthio: Augmenting Small-Scale Audio Classification Datasets with Synthetic Data
di: Ghosh, Sreyan, et al.
Pubblicazione: (2024)
di: Ghosh, Sreyan, et al.
Pubblicazione: (2024)
EGOILLUSION: Benchmarking Hallucinations in Egocentric Video Understanding
di: Seth, Ashish, et al.
Pubblicazione: (2025)
di: Seth, Ashish, et al.
Pubblicazione: (2025)
Do Audio-Language Models Understand Linguistic Variations?
di: Selvakumar, Ramaneswaran, et al.
Pubblicazione: (2024)
di: Selvakumar, Ramaneswaran, et al.
Pubblicazione: (2024)
Failing Forward: Improving Generative Error Correction for ASR with Synthetic Data and Retrieval Augmentation
di: Ghosh, Sreyan, et al.
Pubblicazione: (2024)
di: Ghosh, Sreyan, et al.
Pubblicazione: (2024)
Audio Hallucination Attacks: Probing the Reliability of Large Audio Language Models
di: Seth, Ashish, et al.
Pubblicazione: (2026)
di: Seth, Ashish, et al.
Pubblicazione: (2026)
Do Audio-Visual Large Language Models Really See and Hear?
di: Selvakumar, Ramaneswaran, et al.
Pubblicazione: (2026)
di: Selvakumar, Ramaneswaran, et al.
Pubblicazione: (2026)
AV-RIR: Audio-Visual Room Impulse Response Estimation
di: Ratnarajah, Anton, et al.
Pubblicazione: (2023)
di: Ratnarajah, Anton, et al.
Pubblicazione: (2023)
Audio Flamingo 2: An Audio-Language Model with Long-Audio Understanding and Expert Reasoning Abilities
di: Ghosh, Sreyan, et al.
Pubblicazione: (2025)
di: Ghosh, Sreyan, et al.
Pubblicazione: (2025)
Relic: Enhancing Reward Model Generalization for Low-Resource Indic Languages with Few-Shot Examples
di: Ghosal, Soumya Suvra, et al.
Pubblicazione: (2025)
di: Ghosal, Soumya Suvra, et al.
Pubblicazione: (2025)
Video-Robin: Autoregressive Diffusion Planning for Intent-Grounded Video-to-Music Generation
di: Lokegaonkar, Vaibhavi, et al.
Pubblicazione: (2026)
di: Lokegaonkar, Vaibhavi, et al.
Pubblicazione: (2026)
Audio Flamingo 3: Advancing Audio Intelligence with Fully Open Large Audio Language Models
di: Goel, Arushi, et al.
Pubblicazione: (2025)
di: Goel, Arushi, et al.
Pubblicazione: (2025)
SILA: Signal-to-Language Augmentation for Enhanced Control in Text-to-Audio Generation
di: Kumar, Sonal, et al.
Pubblicazione: (2024)
di: Kumar, Sonal, et al.
Pubblicazione: (2024)
NagaNLP: Bootstrapping NLP for Low-Resource Nagamese Creole with Human-in-the-Loop Synthetic Data
di: Maiti, Agniva, et al.
Pubblicazione: (2025)
di: Maiti, Agniva, et al.
Pubblicazione: (2025)
PromptRefine: Enhancing Few-Shot Performance on Low-Resource Indic Languages with Example Selection from Related Example Banks
di: Ghosal, Soumya Suvra, et al.
Pubblicazione: (2024)
di: Ghosal, Soumya Suvra, et al.
Pubblicazione: (2024)
VisDoM: Multi-Document QA with Visually Rich Elements Using Multimodal Retrieval-Augmented Generation
di: Suri, Manan, et al.
Pubblicazione: (2024)
di: Suri, Manan, et al.
Pubblicazione: (2024)
Dealing with the Hard Facts of Low-Resource African NLP
di: Diarra, Yacouba, et al.
Pubblicazione: (2025)
di: Diarra, Yacouba, et al.
Pubblicazione: (2025)
Exploring NLP Benchmarks in an Extremely Low-Resource Setting
di: Nuha, Ulin, et al.
Pubblicazione: (2025)
di: Nuha, Ulin, et al.
Pubblicazione: (2025)
FairFlow: An Automated Approach to Model-based Counterfactual Data Augmentation For NLP
di: Tokpo, Ewoenam Kwaku, et al.
Pubblicazione: (2024)
di: Tokpo, Ewoenam Kwaku, et al.
Pubblicazione: (2024)
Towards Open-Ended Discovery for Low-Resource NLP
di: Dossou, Bonaventure F. P., et al.
Pubblicazione: (2025)
di: Dossou, Bonaventure F. P., et al.
Pubblicazione: (2025)
On Importance of Pruning and Distillation for Efficient Low Resource NLP
di: Mirashi, Aishwarya, et al.
Pubblicazione: (2024)
di: Mirashi, Aishwarya, et al.
Pubblicazione: (2024)
Generative-Adversarial Networks for Low-Resource Language Data Augmentation in Machine Translation
di: Zeng, Linda
Pubblicazione: (2024)
di: Zeng, Linda
Pubblicazione: (2024)
Evaluation Metrics for Text Data Augmentation in NLP
di: Amadeus, Marcellus, et al.
Pubblicazione: (2024)
di: Amadeus, Marcellus, et al.
Pubblicazione: (2024)
Frustratingly Easy Data Augmentation for Low-Resource ASR
di: Ibaraki, Katsumi, et al.
Pubblicazione: (2025)
di: Ibaraki, Katsumi, et al.
Pubblicazione: (2025)
Data Augmentation and Hyperparameter Tuning for Low-Resource MFA
di: Tosolini, Alessio, et al.
Pubblicazione: (2025)
di: Tosolini, Alessio, et al.
Pubblicazione: (2025)
Intertwining CP and NLP: The Generation of Unreasonably Constrained Sentences
di: Bonlarron, Alexandre, et al.
Pubblicazione: (2024)
di: Bonlarron, Alexandre, et al.
Pubblicazione: (2024)
Documenti analoghi
-
ABEX: Data Augmentation for Low-Resource NLU via Expanding Abstract Descriptions
di: Ghosh, Sreyan, et al.
Pubblicazione: (2024) -
ASPIRE: Language-Guided Data Augmentation for Improving Robustness Against Spurious Correlations
di: Ghosh, Sreyan, et al.
Pubblicazione: (2023) -
RECAP: Retrieval-Augmented Audio Captioning
di: Ghosh, Sreyan, et al.
Pubblicazione: (2023) -
Visual Description Grounding Reduces Hallucinations and Boosts Reasoning in LVLMs
di: Ghosh, Sreyan, et al.
Pubblicazione: (2024) -
CompA: Addressing the Gap in Compositional Reasoning in Audio-Language Models
di: Ghosh, Sreyan, et al.
Pubblicazione: (2023)