Triples-to-isiXhosa (T2X): Addressing the Challenges of Low-Resource Agglutinative Data-to-Text Generation
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Meyer, Francois, Buys, Jan |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
BabyLMs for isiXhosa: Data-Efficient Language Modelling in a Low-Resource Context
von: Matzopoulos, Alexis, et al.
Veröffentlicht: (2025)
von: Matzopoulos, Alexis, et al.
Veröffentlicht: (2025)
Automatically assessing oral narratives of Afrikaans and isiXhosa children
von: Louw, Retief, et al.
Veröffentlicht: (2025)
von: Louw, Retief, et al.
Veröffentlicht: (2025)
Text Detoxification in isiXhosa and Yorùbá: A Cross-Lingual Machine Learning Approach for Low-Resource African Languages
von: Agbeyangi, Abayomi O.
Veröffentlicht: (2026)
von: Agbeyangi, Abayomi O.
Veröffentlicht: (2026)
Feature-based analysis of oral narratives from Afrikaans and isiXhosa children
von: Sharratt, Emma, et al.
Veröffentlicht: (2025)
von: Sharratt, Emma, et al.
Veröffentlicht: (2025)
Speech Recognition for Automatically Assessing Afrikaans and isiXhosa Preschool Oral Narratives
von: Jacobs, Christiaan, et al.
Veröffentlicht: (2025)
von: Jacobs, Christiaan, et al.
Veröffentlicht: (2025)
A Systematic Analysis of Subwords and Cross-Lingual Transfer in Multilingual Translation
von: Meyer, Francois, et al.
Veröffentlicht: (2024)
von: Meyer, Francois, et al.
Veröffentlicht: (2024)
The Learning Dynamics of Subword Segmentation for Morphologically Diverse Languages
von: Meyer, Francois, et al.
Veröffentlicht: (2025)
von: Meyer, Francois, et al.
Veröffentlicht: (2025)
Tokenization Strategies for Low-Resource Agglutinative Languages in Word2Vec: Case Study on Turkish and Finnish
von: Hu, Jinfan Frank
Veröffentlicht: (2025)
von: Hu, Jinfan Frank
Veröffentlicht: (2025)
MzansiText and MzansiLM: An Open Corpus and Decoder-Only Language Model for South African Languages
von: Lombard, Anri, et al.
Veröffentlicht: (2026)
von: Lombard, Anri, et al.
Veröffentlicht: (2026)
An End-to-End Approach for Child Reading Assessment in the Xhosa Language
von: Chevtchenko, Sergio, et al.
Veröffentlicht: (2025)
von: Chevtchenko, Sergio, et al.
Veröffentlicht: (2025)
Adapting Large Language Models to a Low-Resource Agglutinative Language: A Comparative Study of LoRA and QLoRA for Bashkir
von: Arabov, Mullosharaf K., et al.
Veröffentlicht: (2026)
von: Arabov, Mullosharaf K., et al.
Veröffentlicht: (2026)
Text2Data: Low-Resource Data Generation with Textual Control
von: Wang, Shiyu, et al.
Veröffentlicht: (2024)
von: Wang, Shiyu, et al.
Veröffentlicht: (2024)
VerChol -- Grammar-First Tokenization for Agglutinative Languages
von: Raja, Prabhu
Veröffentlicht: (2026)
von: Raja, Prabhu
Veröffentlicht: (2026)
Bangla Key2Text: Text Generation from Keywords for a Low Resource Language
von: Talukder, Tonmoy, et al.
Veröffentlicht: (2026)
von: Talukder, Tonmoy, et al.
Veröffentlicht: (2026)
Marvelous Agglutinative Language Effect on Cross Lingual Transfer Learning
von: Kim, Wooyoung, et al.
Veröffentlicht: (2022)
von: Kim, Wooyoung, et al.
Veröffentlicht: (2022)
A Rigorous Evaluation of LLM Data Generation Strategies for Low-Resource Languages
von: Anikina, Tatiana, et al.
Veröffentlicht: (2025)
von: Anikina, Tatiana, et al.
Veröffentlicht: (2025)
Multipath parsing in the brain
von: Franzluebbers, Berta, et al.
Veröffentlicht: (2024)
von: Franzluebbers, Berta, et al.
Veröffentlicht: (2024)
GlotScript: A Resource and Tool for Low Resource Writing System Identification
von: Kargaran, Amir Hossein, et al.
Veröffentlicht: (2023)
von: Kargaran, Amir Hossein, et al.
Veröffentlicht: (2023)
STAR: Boosting Low-Resource Information Extraction by Structure-to-Text Data Generation with Large Language Models
von: Ma, Mingyu Derek, et al.
Veröffentlicht: (2023)
von: Ma, Mingyu Derek, et al.
Veröffentlicht: (2023)
Unlocking LLMs: Addressing Scarce Data and Bias Challenges in Mental Health
von: Kumar, Vivek, et al.
Veröffentlicht: (2024)
von: Kumar, Vivek, et al.
Veröffentlicht: (2024)
Massively Multilingual Text Translation For Low-Resource Languages
von: Zhou, Zhong
Veröffentlicht: (2024)
von: Zhou, Zhong
Veröffentlicht: (2024)
Modeling Low-Resource Health Coaching Dialogues via Neuro-Symbolic Goal Summarization and Text-Units-Text Generation
von: Zhou, Yue, et al.
Veröffentlicht: (2024)
von: Zhou, Yue, et al.
Veröffentlicht: (2024)
Automated Clinical Report Generation for Remote Cognitive Remediation: Comparing Knowledge-Engineered Templates and LLMs in Low-Resource Settings
von: Zhou, Yongxin, et al.
Veröffentlicht: (2026)
von: Zhou, Yongxin, et al.
Veröffentlicht: (2026)
Paired by the Teacher: Turning Unpaired Data into High-Fidelity Pairs for Low-Resource Text Generation
von: Lu, Yen-Ju, et al.
Veröffentlicht: (2025)
von: Lu, Yen-Ju, et al.
Veröffentlicht: (2025)
$\mathcal{V}isi\mathcal{P}runer$: Decoding Discontinuous Cross-Modal Dynamics for Efficient Multimodal LLMs
von: Fan, Yingqi, et al.
Veröffentlicht: (2025)
von: Fan, Yingqi, et al.
Veröffentlicht: (2025)
Scaling Low-Resource MT via Synthetic Data Generation with LLMs
von: de Gibert, Ona, et al.
Veröffentlicht: (2025)
von: de Gibert, Ona, et al.
Veröffentlicht: (2025)
Bias Attribution in Filipino Language Models: Extending a Bias Interpretability Metric for Application on Agglutinative Languages
von: Gamboa, Lance Calvin Lim, et al.
Veröffentlicht: (2025)
von: Gamboa, Lance Calvin Lim, et al.
Veröffentlicht: (2025)
Conflicts in Texts: Data, Implications and Challenges
von: Liu, Siyi, et al.
Veröffentlicht: (2025)
von: Liu, Siyi, et al.
Veröffentlicht: (2025)
GlotLID: Language Identification for Low-Resource Languages
von: Kargaran, Amir Hossein, et al.
Veröffentlicht: (2023)
von: Kargaran, Amir Hossein, et al.
Veröffentlicht: (2023)
CoDa: Constrained Generation based Data Augmentation for Low-Resource NLP
von: Evuru, Chandra Kiran Reddy, et al.
Veröffentlicht: (2024)
von: Evuru, Chandra Kiran Reddy, et al.
Veröffentlicht: (2024)
Generative-Adversarial Networks for Low-Resource Language Data Augmentation in Machine Translation
von: Zeng, Linda
Veröffentlicht: (2024)
von: Zeng, Linda
Veröffentlicht: (2024)
A fully automated and scalable Parallel Data Augmentation for Low Resource Languages using Image and Text Analytics
von: Sharma, Prawaal, et al.
Veröffentlicht: (2025)
von: Sharma, Prawaal, et al.
Veröffentlicht: (2025)
Rethinking MUSHRA: Addressing Modern Challenges in Text-to-Speech Evaluation
von: Varadhan, Praveen Srinivasa, et al.
Veröffentlicht: (2024)
von: Varadhan, Praveen Srinivasa, et al.
Veröffentlicht: (2024)
Algorithm for Semantic Network Generation from Texts of Low Resource Languages Such as Kiswahili
von: Wanjawa, Barack Wamkaya, et al.
Veröffentlicht: (2025)
von: Wanjawa, Barack Wamkaya, et al.
Veröffentlicht: (2025)
A Unified Data Augmentation Framework for Low-Resource Multi-Domain Dialogue Generation
von: Liu, Yongkang, et al.
Veröffentlicht: (2024)
von: Liu, Yongkang, et al.
Veröffentlicht: (2024)
LLM-Based Data Generation and Clinical Skills Evaluation for Low-Resource French OSCEs
von: Huang, Tian, et al.
Veröffentlicht: (2026)
von: Huang, Tian, et al.
Veröffentlicht: (2026)
TopXGen: Topic-Diverse Parallel Data Generation for Low-Resource Machine Translation
von: Zebaze, Armel, et al.
Veröffentlicht: (2025)
von: Zebaze, Armel, et al.
Veröffentlicht: (2025)
Transformer-Driven Triple Fusion Framework for Enhanced Multimodal Author Intent Classification in Low-Resource Bangla
von: Islam, Ariful, et al.
Veröffentlicht: (2025)
von: Islam, Ariful, et al.
Veröffentlicht: (2025)
Better as Generators Than Classifiers: Leveraging LLMs and Synthetic Data for Low-Resource Multilingual Classification
von: Pecher, Branislav, et al.
Veröffentlicht: (2026)
von: Pecher, Branislav, et al.
Veröffentlicht: (2026)
Efficient Biomedical Entity Linking: Clinical Text Standardization with Low-Resource Techniques
von: Achara, Akshit, et al.
Veröffentlicht: (2024)
von: Achara, Akshit, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
BabyLMs for isiXhosa: Data-Efficient Language Modelling in a Low-Resource Context
von: Matzopoulos, Alexis, et al.
Veröffentlicht: (2025) -
Automatically assessing oral narratives of Afrikaans and isiXhosa children
von: Louw, Retief, et al.
Veröffentlicht: (2025) -
Text Detoxification in isiXhosa and Yorùbá: A Cross-Lingual Machine Learning Approach for Low-Resource African Languages
von: Agbeyangi, Abayomi O.
Veröffentlicht: (2026) -
Feature-based analysis of oral narratives from Afrikaans and isiXhosa children
von: Sharratt, Emma, et al.
Veröffentlicht: (2025) -
Speech Recognition for Automatically Assessing Afrikaans and isiXhosa Preschool Oral Narratives
von: Jacobs, Christiaan, et al.
Veröffentlicht: (2025)