SumTablets: A Transliteration Dataset of Sumerian Tablets
Fuente:
arXiv
Salvato in:
| Autori principali: | Simmons, Cole, Martinez, Richard Diehl, Jurafsky, Dan |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2026
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Transcribe, Translate, or Transliterate: An Investigation of Intermediate Representations in Spoken Language Models
di: Ògúnrèmí, Tolúlopé, et al.
Pubblicazione: (2025)
di: Ògúnrèmí, Tolúlopé, et al.
Pubblicazione: (2025)
Data Checklist: On Unit-Testing Datasets with Usable Information
di: Zhang, Heidi C., et al.
Pubblicazione: (2024)
di: Zhang, Heidi C., et al.
Pubblicazione: (2024)
A layer-wise analysis of Mandarin and English suprasegmentals in SSL speech models
di: de la Fuente, Antón, et al.
Pubblicazione: (2024)
di: de la Fuente, Antón, et al.
Pubblicazione: (2024)
BanTH: A Multi-label Hate Speech Detection Dataset for Transliterated Bangla
di: Haider, Fabiha, et al.
Pubblicazione: (2024)
di: Haider, Fabiha, et al.
Pubblicazione: (2024)
Learning Concepts, Not Tokens: Self-Supervised Semantic Alignment for Language Models
di: Zhang, Christine, et al.
Pubblicazione: (2026)
di: Zhang, Christine, et al.
Pubblicazione: (2026)
Do "New Snow Tablets" Contain Snow? Large Language Models Over-Rely on Names to Identify Ingredients of Chinese Drugs
di: Li, Sifan, et al.
Pubblicazione: (2025)
di: Li, Sifan, et al.
Pubblicazione: (2025)
HumT DumT: Measuring and controlling human-like language in LLMs
di: Cheng, Myra, et al.
Pubblicazione: (2025)
di: Cheng, Myra, et al.
Pubblicazione: (2025)
How Transliterations Improve Crosslingual Alignment
di: Liu, Yihong, et al.
Pubblicazione: (2024)
di: Liu, Yihong, et al.
Pubblicazione: (2024)
Othering and low status framing of immigrant cuisines in US restaurant reviews and large language models
di: Luo, Yiwei, et al.
Pubblicazione: (2023)
di: Luo, Yiwei, et al.
Pubblicazione: (2023)
Accommodation and Epistemic Vigilance: A Pragmatic Account of Why LLMs Fail to Challenge Harmful Beliefs
di: Cheng, Myra, et al.
Pubblicazione: (2026)
di: Cheng, Myra, et al.
Pubblicazione: (2026)
Humans overrely on overconfident language models, across languages
di: Rathi, Neil, et al.
Pubblicazione: (2025)
di: Rathi, Neil, et al.
Pubblicazione: (2025)
Connecting the Persian-speaking World through Transliteration
di: Merchant, Rayyan, et al.
Pubblicazione: (2025)
di: Merchant, Rayyan, et al.
Pubblicazione: (2025)
False Friends Are Not Foes: Investigating Vocabulary Overlap in Multilingual Language Models
di: Kallini, Julie, et al.
Pubblicazione: (2025)
di: Kallini, Julie, et al.
Pubblicazione: (2025)
AnthroScore: A Computational Linguistic Measure of Anthropomorphism
di: Cheng, Myra, et al.
Pubblicazione: (2024)
di: Cheng, Myra, et al.
Pubblicazione: (2024)
CES 2011: Tablet Crazy
di: Rapp, David
Pubblicazione: (2011)
di: Rapp, David
Pubblicazione: (2011)
Jailbreaking LLMs with Arabic Transliteration and Arabizi
di: Ghanim, Mansour Al, et al.
Pubblicazione: (2024)
di: Ghanim, Mansour Al, et al.
Pubblicazione: (2024)
ParsTranslit: Truly Versatile Tajik-Farsi Transliteration
di: Merchant, Rayyan, et al.
Pubblicazione: (2025)
di: Merchant, Rayyan, et al.
Pubblicazione: (2025)
Swa Bhasha: Message-Based Singlish to Sinhala Transliteration
di: Athukorala, Maneesha U., et al.
Pubblicazione: (2024)
di: Athukorala, Maneesha U., et al.
Pubblicazione: (2024)
Language Detection for Transliterated Content
di: S, Selva Kumar, et al.
Pubblicazione: (2024)
di: S, Selva Kumar, et al.
Pubblicazione: (2024)
CausalGym: Benchmarking causal interpretability methods on linguistic tasks
di: Arora, Aryaman, et al.
Pubblicazione: (2024)
di: Arora, Aryaman, et al.
Pubblicazione: (2024)
Scripts Through Time: A Survey of the Evolving Role of Transliteration in NLP
di: Jayakumar, Thanmay, et al.
Pubblicazione: (2026)
di: Jayakumar, Thanmay, et al.
Pubblicazione: (2026)
A Tale of Two Scripts: Transliteration and Post-Correction for Judeo-Arabic
di: Gonzalez, Juan Moreno, et al.
Pubblicazione: (2025)
di: Gonzalez, Juan Moreno, et al.
Pubblicazione: (2025)
A Benchmark for Learning to Translate a New Language from One Grammar Book
di: Tanzer, Garrett, et al.
Pubblicazione: (2023)
di: Tanzer, Garrett, et al.
Pubblicazione: (2023)
Beyond Tokens: Concept-Level Training Objectives for LLMs
di: Iyer, Laya, et al.
Pubblicazione: (2026)
di: Iyer, Laya, et al.
Pubblicazione: (2026)
The Roots of Performance Disparity in Multilingual Language Models: Intrinsic Modeling Difficulty or Design Choices?
di: Shani, Chen, et al.
Pubblicazione: (2026)
di: Shani, Chen, et al.
Pubblicazione: (2026)
AyutthayaAlpha: A Thai-Latin Script Transliteration Transformer
di: Lauc, Davor, et al.
Pubblicazione: (2024)
di: Lauc, Davor, et al.
Pubblicazione: (2024)
Happiness is Sharing a Vocabulary: A Study of Transliteration Methods
di: Jung, Haeji, et al.
Pubblicazione: (2025)
di: Jung, Haeji, et al.
Pubblicazione: (2025)
Romanized to Native Malayalam Script Transliteration Using an Encoder-Decoder Framework
di: Baiju, Bajiyo, et al.
Pubblicazione: (2024)
di: Baiju, Bajiyo, et al.
Pubblicazione: (2024)
Tending Towards Stability: Convergence Challenges in Small Language Models
di: Martinez, Richard Diehl, et al.
Pubblicazione: (2024)
di: Martinez, Richard Diehl, et al.
Pubblicazione: (2024)
Dialect prejudice predicts AI decisions about people's character, employability, and criminality
di: Hofmann, Valentin, et al.
Pubblicazione: (2024)
di: Hofmann, Valentin, et al.
Pubblicazione: (2024)
Beyond Specialization: Benchmarking LLMs for Transliteration of Indian Languages
di: Azam, Gulfarogh, et al.
Pubblicazione: (2025)
di: Azam, Gulfarogh, et al.
Pubblicazione: (2025)
Linear Script Representations in Speech Foundation Models Enable Zero-Shot Transliteration
di: Shim, Ryan Soh-Eun, et al.
Pubblicazione: (2026)
di: Shim, Ryan Soh-Eun, et al.
Pubblicazione: (2026)
Data Augmentation for Maltese NLP using Transliterated and Machine Translated Arabic Data
di: Micallef, Kurt, et al.
Pubblicazione: (2025)
di: Micallef, Kurt, et al.
Pubblicazione: (2025)
Rethinking Word Similarity: Semantic Similarity through Classification Confusion
di: Zhou, Kaitlyn, et al.
Pubblicazione: (2025)
di: Zhou, Kaitlyn, et al.
Pubblicazione: (2025)
Generation Space Size: Understanding and Calibrating Open-Endedness of LLM Generations
di: Yu, Sunny, et al.
Pubblicazione: (2025)
di: Yu, Sunny, et al.
Pubblicazione: (2025)
Dynamic Cheatsheet: Test-Time Learning with Adaptive Memory
di: Suzgun, Mirac, et al.
Pubblicazione: (2025)
di: Suzgun, Mirac, et al.
Pubblicazione: (2025)
ChakmaNMT: Machine Translation for a Low-Resource and Endangered Language via Transliteration
di: Chakma, Aunabil, et al.
Pubblicazione: (2024)
di: Chakma, Aunabil, et al.
Pubblicazione: (2024)
Swa-bhasha Resource Hub: Romanized Sinhala to Sinhala Transliteration Systems and Data Resources
di: Sumanathilaka, Deshan, et al.
Pubblicazione: (2025)
di: Sumanathilaka, Deshan, et al.
Pubblicazione: (2025)
Fractured Tablets
di: Balberg, Mira
Pubblicazione: (2023)
di: Balberg, Mira
Pubblicazione: (2023)
NADIR: Differential Attention Flow for Non-Autoregressive Transliteration in Indic Languages
di: Tomar, Lakshya, et al.
Pubblicazione: (2026)
di: Tomar, Lakshya, et al.
Pubblicazione: (2026)
Documenti analoghi
-
Transcribe, Translate, or Transliterate: An Investigation of Intermediate Representations in Spoken Language Models
di: Ògúnrèmí, Tolúlopé, et al.
Pubblicazione: (2025) -
Data Checklist: On Unit-Testing Datasets with Usable Information
di: Zhang, Heidi C., et al.
Pubblicazione: (2024) -
A layer-wise analysis of Mandarin and English suprasegmentals in SSL speech models
di: de la Fuente, Antón, et al.
Pubblicazione: (2024) -
BanTH: A Multi-label Hate Speech Detection Dataset for Transliterated Bangla
di: Haider, Fabiha, et al.
Pubblicazione: (2024) -
Learning Concepts, Not Tokens: Self-Supervised Semantic Alignment for Language Models
di: Zhang, Christine, et al.
Pubblicazione: (2026)