MALT: Mechanistic Ablation of Lossy Translation in LLMs for a Low-Resource Language: Urdu
Fuente:
arXiv
Guardado en:
| Autor principal: | Bajwa, Taaha Saleem |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
UrduFactCheck: An Agentic Fact-Checking Framework for Urdu with Evidence Boosting and Benchmarking
por: Ahmad, Sarfraz, et al.
Publicado: (2025)
por: Ahmad, Sarfraz, et al.
Publicado: (2025)
UrduBench: An Urdu Reasoning Benchmark using Contextually Ensembled Translations with Human-in-the-Loop
por: Shafique, Muhammad Ali, et al.
Publicado: (2026)
por: Shafique, Muhammad Ali, et al.
Publicado: (2026)
Reflective Translation: Improving Low-Resource Machine Translation via Structured Self-Reflection
por: Cheng, Nicholas
Publicado: (2026)
por: Cheng, Nicholas
Publicado: (2026)
UrduLLaMA 1.0: Dataset Curation, Preprocessing, and Evaluation in Low-Resource Settings
por: Fiaz, Layba, et al.
Publicado: (2025)
por: Fiaz, Layba, et al.
Publicado: (2025)
PersistBench: When Should Long-Term Memories Be Forgotten by LLMs?
por: Pulipaka, Sidharth, et al.
Publicado: (2026)
por: Pulipaka, Sidharth, et al.
Publicado: (2026)
Benchmarking the Performance of Pre-trained LLMs across Urdu NLP Tasks
por: Tahir, Munief Hassan, et al.
Publicado: (2024)
por: Tahir, Munief Hassan, et al.
Publicado: (2024)
Leveraging Large Language Models for Accurate Sign Language Translation in Low-Resource Scenarios
por: Bulla, Luana, et al.
Publicado: (2025)
por: Bulla, Luana, et al.
Publicado: (2025)
Machine Translation Hallucination Detection for Low and High Resource Languages using Large Language Models
por: Benkirane, Kenza, et al.
Publicado: (2024)
por: Benkirane, Kenza, et al.
Publicado: (2024)
Algorithm for Semantic Network Generation from Texts of Low Resource Languages Such as Kiswahili
por: Wanjawa, Barack Wamkaya, et al.
Publicado: (2025)
por: Wanjawa, Barack Wamkaya, et al.
Publicado: (2025)
A Multi-Task Benchmark for Abusive Language Detection in Low-Resource Settings
por: Gaim, Fitsum, et al.
Publicado: (2025)
por: Gaim, Fitsum, et al.
Publicado: (2025)
Enabling Low-Resource Language Retrieval: Establishing Baselines for Urdu MS MARCO
por: Butt, Umer, et al.
Publicado: (2024)
por: Butt, Umer, et al.
Publicado: (2024)
Dialect Matters: Cross-Lingual ASR Transfer for Low-Resource Indic Language Varieties
por: Dhasmana, Akriti, et al.
Publicado: (2026)
por: Dhasmana, Akriti, et al.
Publicado: (2026)
Multi-Method Validation of Large Language Model Medical Translation Across High- and Low-Resource Languages
por: Anyaegbuna, Chukwuebuka, et al.
Publicado: (2026)
por: Anyaegbuna, Chukwuebuka, et al.
Publicado: (2026)
Hard to Be Heard: Phoneme-Level ASR Analysis of Phonologically Complex, Low-Resource Endangered Languages
por: Akavarapu, V. S. D. S. Mahesh, et al.
Publicado: (2026)
por: Akavarapu, V. S. D. S. Mahesh, et al.
Publicado: (2026)
SITA: Learning Speaker-Invariant and Tone-Aware Speech Representations for Low-Resource Tonal Languages
por: Xu, Tianyi, et al.
Publicado: (2026)
por: Xu, Tianyi, et al.
Publicado: (2026)
Yes-MT's Submission to the Low-Resource Indic Language Translation Shared Task in WMT 2024
por: Bhaskar, Yash, et al.
Publicado: (2025)
por: Bhaskar, Yash, et al.
Publicado: (2025)
Beyond Many-Shot Translation: Scaling In-Context Demonstrations For Low-Resource Machine Translation
por: Salim, Luis Frentzen, et al.
Publicado: (2026)
por: Salim, Luis Frentzen, et al.
Publicado: (2026)
Pivot Language for Low-Resource Machine Translation
por: Talwar, Abhimanyu, et al.
Publicado: (2025)
por: Talwar, Abhimanyu, et al.
Publicado: (2025)
Heidelberg-Boston @ SIGTYP 2024 Shared Task: Enhancing Low-Resource Language Analysis With Character-Aware Hierarchical Transformers
por: Riemenschneider, Frederick, et al.
Publicado: (2024)
por: Riemenschneider, Frederick, et al.
Publicado: (2024)
Omnilingual MT: Machine Translation for 1,600 Languages
por: Omnilingual MT Team, et al.
Publicado: (2026)
por: Omnilingual MT Team, et al.
Publicado: (2026)
RomanLens: The Role Of Latent Romanization In Multilinguality In LLMs
por: Saji, Alan, et al.
Publicado: (2025)
por: Saji, Alan, et al.
Publicado: (2025)
Large Language Models for Persian $ \leftrightarrow $ English Idiom Translation
por: Rezaeimanesh, Sara, et al.
Publicado: (2024)
por: Rezaeimanesh, Sara, et al.
Publicado: (2024)
KinyaColBERT: A Lexically Grounded Retrieval Model for Low-Resource Retrieval-Augmented Generation
por: Nzeyimana, Antoine, et al.
Publicado: (2025)
por: Nzeyimana, Antoine, et al.
Publicado: (2025)
BOUQuET: dataset, Benchmark and Open initiative for Universal Quality Evaluation in Translation
por: The Omnilingual MT Team, et al.
Publicado: (2025)
por: The Omnilingual MT Team, et al.
Publicado: (2025)
Multilingual and Multimodal LLMs in the Wild: Building for Low-Resource Languages
por: Alam, Firoj, et al.
Publicado: (2026)
por: Alam, Firoj, et al.
Publicado: (2026)
Low-Resource Court Judgment Summarization for Common Law Systems
por: Liu, Shuaiqi, et al.
Publicado: (2024)
por: Liu, Shuaiqi, et al.
Publicado: (2024)
CRISP: Persistent Concept Unlearning via Sparse Autoencoders
por: Ashuach, Tomer, et al.
Publicado: (2025)
por: Ashuach, Tomer, et al.
Publicado: (2025)
Mitigating Translationese in Low-resource Languages: The Storyboard Approach
por: Kuwanto, Garry, et al.
Publicado: (2024)
por: Kuwanto, Garry, et al.
Publicado: (2024)
MedFuzz: Exploring the Robustness of Large Language Models in Medical Question Answering
por: Ness, Robert Osazuwa, et al.
Publicado: (2024)
por: Ness, Robert Osazuwa, et al.
Publicado: (2024)
A Case Study of Cross-Lingual Zero-Shot Generalization for Classical Languages in LLMs
por: Akavarapu, V. S. D. S. Mahesh, et al.
Publicado: (2025)
por: Akavarapu, V. S. D. S. Mahesh, et al.
Publicado: (2025)
Learning Translations via Matrix Completion
por: Wijaya, Derry, et al.
Publicado: (2024)
por: Wijaya, Derry, et al.
Publicado: (2024)
LoRS: Efficient Low-Rank Adaptation for Sparse Large Language Model
por: Hu, Yuxuan, et al.
Publicado: (2025)
por: Hu, Yuxuan, et al.
Publicado: (2025)
Causally Grounded Mechanistic Interpretability for LLMs with Faithful Natural-Language Explanations
por: Mahale, Ajay Pravin
Publicado: (2026)
por: Mahale, Ajay Pravin
Publicado: (2026)
Test-Time Scaling of Reasoning Models for Machine Translation
por: Li, Zihao, et al.
Publicado: (2025)
por: Li, Zihao, et al.
Publicado: (2025)
Whether, Not Which: Mechanistic Interpretability Reveals Dissociable Affect Reception and Emotion Categorization in LLMs
por: Keeman, Michael
Publicado: (2026)
por: Keeman, Michael
Publicado: (2026)
Do Large Language Models Speak All Languages Equally? A Comparative Study in Low-Resource Settings
por: Hasan, Md. Arid, et al.
Publicado: (2024)
por: Hasan, Md. Arid, et al.
Publicado: (2024)
Improving Retrieval-Augmented Neural Machine Translation with Monolingual Data
por: Bouthors, Maxime, et al.
Publicado: (2025)
por: Bouthors, Maxime, et al.
Publicado: (2025)
COMET-poly: Machine Translation Metric Grounded in Other Candidates
por: Züfle, Maike, et al.
Publicado: (2025)
por: Züfle, Maike, et al.
Publicado: (2025)
Multilingual LLMs Inherently Reward In-Language Time-Sensitive Semantic Alignment for Low-Resource Languages
por: Bajpai, Ashutosh, et al.
Publicado: (2024)
por: Bajpai, Ashutosh, et al.
Publicado: (2024)
Towards Fundamental Language Models: Does Linguistic Competence Scale with Model Size?
por: Collado-Montañez, Jaime, et al.
Publicado: (2025)
por: Collado-Montañez, Jaime, et al.
Publicado: (2025)
Ejemplares similares
-
UrduFactCheck: An Agentic Fact-Checking Framework for Urdu with Evidence Boosting and Benchmarking
por: Ahmad, Sarfraz, et al.
Publicado: (2025) -
UrduBench: An Urdu Reasoning Benchmark using Contextually Ensembled Translations with Human-in-the-Loop
por: Shafique, Muhammad Ali, et al.
Publicado: (2026) -
Reflective Translation: Improving Low-Resource Machine Translation via Structured Self-Reflection
por: Cheng, Nicholas
Publicado: (2026) -
UrduLLaMA 1.0: Dataset Curation, Preprocessing, and Evaluation in Low-Resource Settings
por: Fiaz, Layba, et al.
Publicado: (2025) -
PersistBench: When Should Long-Term Memories Be Forgotten by LLMs?
por: Pulipaka, Sidharth, et al.
Publicado: (2026)