Enregistré dans:
| Auteurs principaux: | Simmons, Cole, Martinez, Richard Diehl, Jurafsky, Dan |
|---|---|
| Format: | Preprint |
| Publié: |
2026
|
| Sujets: | |
| Accès en ligne: | https://arxiv.org/abs/2602.22200 |
| Tags: |
Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
|
Documents similaires
Transcribe, Translate, or Transliterate: An Investigation of Intermediate Representations in Spoken Language Models
par: Ògúnrèmí, Tolúlopé, et autres
Publié: (2025)
par: Ògúnrèmí, Tolúlopé, et autres
Publié: (2025)
Data Checklist: On Unit-Testing Datasets with Usable Information
par: Zhang, Heidi C., et autres
Publié: (2024)
par: Zhang, Heidi C., et autres
Publié: (2024)
A layer-wise analysis of Mandarin and English suprasegmentals in SSL speech models
par: de la Fuente, Antón, et autres
Publié: (2024)
par: de la Fuente, Antón, et autres
Publié: (2024)
Learning Concepts, Not Tokens: Self-Supervised Semantic Alignment for Language Models
par: Zhang, Christine, et autres
Publié: (2026)
par: Zhang, Christine, et autres
Publié: (2026)
HumT DumT: Measuring and controlling human-like language in LLMs
par: Cheng, Myra, et autres
Publié: (2025)
par: Cheng, Myra, et autres
Publié: (2025)
BanTH: A Multi-label Hate Speech Detection Dataset for Transliterated Bangla
par: Haider, Fabiha, et autres
Publié: (2024)
par: Haider, Fabiha, et autres
Publié: (2024)
Othering and low status framing of immigrant cuisines in US restaurant reviews and large language models
par: Luo, Yiwei, et autres
Publié: (2023)
par: Luo, Yiwei, et autres
Publié: (2023)
Accommodation and Epistemic Vigilance: A Pragmatic Account of Why LLMs Fail to Challenge Harmful Beliefs
par: Cheng, Myra, et autres
Publié: (2026)
par: Cheng, Myra, et autres
Publié: (2026)
Humans overrely on overconfident language models, across languages
par: Rathi, Neil, et autres
Publié: (2025)
par: Rathi, Neil, et autres
Publié: (2025)
Do "New Snow Tablets" Contain Snow? Large Language Models Over-Rely on Names to Identify Ingredients of Chinese Drugs
par: Li, Sifan, et autres
Publié: (2025)
par: Li, Sifan, et autres
Publié: (2025)
AnthroScore: A Computational Linguistic Measure of Anthropomorphism
par: Cheng, Myra, et autres
Publié: (2024)
par: Cheng, Myra, et autres
Publié: (2024)
How Transliterations Improve Crosslingual Alignment
par: Liu, Yihong, et autres
Publié: (2024)
par: Liu, Yihong, et autres
Publié: (2024)
False Friends Are Not Foes: Investigating Vocabulary Overlap in Multilingual Language Models
par: Kallini, Julie, et autres
Publié: (2025)
par: Kallini, Julie, et autres
Publié: (2025)
CausalGym: Benchmarking causal interpretability methods on linguistic tasks
par: Arora, Aryaman, et autres
Publié: (2024)
par: Arora, Aryaman, et autres
Publié: (2024)
CES 2011: Tablet Crazy
par: Rapp, David
Publié: (2011)
par: Rapp, David
Publié: (2011)
Connecting the Persian-speaking World through Transliteration
par: Merchant, Rayyan, et autres
Publié: (2025)
par: Merchant, Rayyan, et autres
Publié: (2025)
Jailbreaking LLMs with Arabic Transliteration and Arabizi
par: Ghanim, Mansour Al, et autres
Publié: (2024)
par: Ghanim, Mansour Al, et autres
Publié: (2024)
Language Detection for Transliterated Content
par: S, Selva Kumar, et autres
Publié: (2024)
par: S, Selva Kumar, et autres
Publié: (2024)
ParsTranslit: Truly Versatile Tajik-Farsi Transliteration
par: Merchant, Rayyan, et autres
Publié: (2025)
par: Merchant, Rayyan, et autres
Publié: (2025)
Swa Bhasha: Message-Based Singlish to Sinhala Transliteration
par: Athukorala, Maneesha U., et autres
Publié: (2024)
par: Athukorala, Maneesha U., et autres
Publié: (2024)
A Benchmark for Learning to Translate a New Language from One Grammar Book
par: Tanzer, Garrett, et autres
Publié: (2023)
par: Tanzer, Garrett, et autres
Publié: (2023)
Dialect prejudice predicts AI decisions about people's character, employability, and criminality
par: Hofmann, Valentin, et autres
Publié: (2024)
par: Hofmann, Valentin, et autres
Publié: (2024)
Beyond Tokens: Concept-Level Training Objectives for LLMs
par: Iyer, Laya, et autres
Publié: (2026)
par: Iyer, Laya, et autres
Publié: (2026)
The Roots of Performance Disparity in Multilingual Language Models: Intrinsic Modeling Difficulty or Design Choices?
par: Shani, Chen, et autres
Publié: (2026)
par: Shani, Chen, et autres
Publié: (2026)
Scripts Through Time: A Survey of the Evolving Role of Transliteration in NLP
par: Jayakumar, Thanmay, et autres
Publié: (2026)
par: Jayakumar, Thanmay, et autres
Publié: (2026)
A Tale of Two Scripts: Transliteration and Post-Correction for Judeo-Arabic
par: Gonzalez, Juan Moreno, et autres
Publié: (2025)
par: Gonzalez, Juan Moreno, et autres
Publié: (2025)
Tending Towards Stability: Convergence Challenges in Small Language Models
par: Martinez, Richard Diehl, et autres
Publié: (2024)
par: Martinez, Richard Diehl, et autres
Publié: (2024)
AyutthayaAlpha: A Thai-Latin Script Transliteration Transformer
par: Lauc, Davor, et autres
Publié: (2024)
par: Lauc, Davor, et autres
Publié: (2024)
Happiness is Sharing a Vocabulary: A Study of Transliteration Methods
par: Jung, Haeji, et autres
Publié: (2025)
par: Jung, Haeji, et autres
Publié: (2025)
Rethinking Word Similarity: Semantic Similarity through Classification Confusion
par: Zhou, Kaitlyn, et autres
Publié: (2025)
par: Zhou, Kaitlyn, et autres
Publié: (2025)
Beyond Specialization: Benchmarking LLMs for Transliteration of Indian Languages
par: Azam, Gulfarogh, et autres
Publié: (2025)
par: Azam, Gulfarogh, et autres
Publié: (2025)
Generation Space Size: Understanding and Calibrating Open-Endedness of LLM Generations
par: Yu, Sunny, et autres
Publié: (2025)
par: Yu, Sunny, et autres
Publié: (2025)
Dynamic Cheatsheet: Test-Time Learning with Adaptive Memory
par: Suzgun, Mirac, et autres
Publié: (2025)
par: Suzgun, Mirac, et autres
Publié: (2025)
Bayesian scaling laws for in-context learning
par: Arora, Aryaman, et autres
Publié: (2024)
par: Arora, Aryaman, et autres
Publié: (2024)
Romanized to Native Malayalam Script Transliteration Using an Encoder-Decoder Framework
par: Baiju, Bajiyo, et autres
Publié: (2024)
par: Baiju, Bajiyo, et autres
Publié: (2024)
Learning the meanings of function words from grounded language using a visual question answering model
par: Portelance, Eva, et autres
Publié: (2023)
par: Portelance, Eva, et autres
Publié: (2023)
Fractured Tablets
par: Balberg, Mira
Publié: (2023)
par: Balberg, Mira
Publié: (2023)
NLP Systems That Can't Tell Use from Mention Censor Counterspeech, but Teaching the Distinction Helps
par: Gligoric, Kristina, et autres
Publié: (2024)
par: Gligoric, Kristina, et autres
Publié: (2024)
Grounding Gaps in Language Model Generations
par: Shaikh, Omar, et autres
Publié: (2023)
par: Shaikh, Omar, et autres
Publié: (2023)
Linear Script Representations in Speech Foundation Models Enable Zero-Shot Transliteration
par: Shim, Ryan Soh-Eun, et autres
Publié: (2026)
par: Shim, Ryan Soh-Eun, et autres
Publié: (2026)
Documents similaires
-
Transcribe, Translate, or Transliterate: An Investigation of Intermediate Representations in Spoken Language Models
par: Ògúnrèmí, Tolúlopé, et autres
Publié: (2025) -
Data Checklist: On Unit-Testing Datasets with Usable Information
par: Zhang, Heidi C., et autres
Publié: (2024) -
A layer-wise analysis of Mandarin and English suprasegmentals in SSL speech models
par: de la Fuente, Antón, et autres
Publié: (2024) -
Learning Concepts, Not Tokens: Self-Supervised Semantic Alignment for Language Models
par: Zhang, Christine, et autres
Publié: (2026) -
HumT DumT: Measuring and controlling human-like language in LLMs
par: Cheng, Myra, et autres
Publié: (2025)