Training Models on Dialects of Translationese Shows How Lexical Diversity and Source-Target Syntactic Similarity Shape Learning
Fuente:
arXiv
Gespeichert in:
| 1. Verfasser: | Kunz, Jenny |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
A Dataset for Probing Translationese Preferences in English-to-Swedish Translation
von: Kunz, Jenny, et al.
Veröffentlicht: (2026)
von: Kunz, Jenny, et al.
Veröffentlicht: (2026)
Lost in Literalism: How Supervised Training Shapes Translationese in LLMs
von: Li, Yafu, et al.
Veröffentlicht: (2025)
von: Li, Yafu, et al.
Veröffentlicht: (2025)
Translationese-index: Using Likelihood Ratios for Graded and Generalizable Measurement of Translationese
von: Liu, Yikang, et al.
Veröffentlicht: (2025)
von: Liu, Yikang, et al.
Veröffentlicht: (2025)
Train More Parameters But Mind Their Placement: Insights into Language Adaptation with PEFT
von: Kunz, Jenny
Veröffentlicht: (2024)
von: Kunz, Jenny
Veröffentlicht: (2024)
Pretraining Language Models Using Translationese
von: Doshi, Meet, et al.
Veröffentlicht: (2024)
von: Doshi, Meet, et al.
Veröffentlicht: (2024)
Incorporating Lexical and Syntactic Knowledge for Unsupervised Cross-Lingual Transfer
von: Zheng, Jianyu, et al.
Veröffentlicht: (2024)
von: Zheng, Jianyu, et al.
Veröffentlicht: (2024)
Extracting Lexical Features from Dialects via Interpretable Dialect Classifiers
von: Xie, Roy, et al.
Veröffentlicht: (2024)
von: Xie, Roy, et al.
Veröffentlicht: (2024)
A Study on How Attention Scores in the BERT Model are Aware of Lexical Categories in Syntactic and Semantic Tasks on the GLUE Benchmark
von: Jang, Dongjun, et al.
Veröffentlicht: (2024)
von: Jang, Dongjun, et al.
Veröffentlicht: (2024)
How to Tune a Multilingual Encoder Model for Germanic Languages: A Study of PEFT, Full Fine-Tuning, and Language Adapters
von: Oji, Romina, et al.
Veröffentlicht: (2025)
von: Oji, Romina, et al.
Veröffentlicht: (2025)
Same Meaning, Different Scores: Lexical and Syntactic Sensitivity in LLM Evaluation
von: Kostić, Bogdan, et al.
Veröffentlicht: (2026)
von: Kostić, Bogdan, et al.
Veröffentlicht: (2026)
Crowdsourcing Lexical Diversity
von: Khalilia, Hadi, et al.
Veröffentlicht: (2024)
von: Khalilia, Hadi, et al.
Veröffentlicht: (2024)
Preferences for Idiomatic Language are Acquired Slowly -- and Forgotten Quickly: A Case Study on Swedish
von: Kunz, Jenny
Veröffentlicht: (2026)
von: Kunz, Jenny
Veröffentlicht: (2026)
A Diagnostic Benchmark for Sweden-Related Factual Knowledge
von: Kunz, Jenny
Veröffentlicht: (2025)
von: Kunz, Jenny
Veröffentlicht: (2025)
Measuring Spurious Correlation in Classification: 'Clever Hans' in Translationese
von: Borah, Angana, et al.
Veröffentlicht: (2023)
von: Borah, Angana, et al.
Veröffentlicht: (2023)
Language Model Re-rankers are Fooled by Lexical Similarities
von: Hagström, Lovisa, et al.
Veröffentlicht: (2025)
von: Hagström, Lovisa, et al.
Veröffentlicht: (2025)
A Hypothesis-Driven Framework for the Analysis of Self-Rationalising Models
von: Braun, Marc, et al.
Veröffentlicht: (2024)
von: Braun, Marc, et al.
Veröffentlicht: (2024)
Targeted Syntactic Evaluation of Language Models on Georgian Case Alignment
von: Gallagher, Daniel, et al.
Veröffentlicht: (2026)
von: Gallagher, Daniel, et al.
Veröffentlicht: (2026)
ParaFusion: A Large-Scale LLM-Driven English Paraphrase Dataset Infused with High-Quality Lexical and Syntactic Diversity
von: Jayawardena, Lasal, et al.
Veröffentlicht: (2024)
von: Jayawardena, Lasal, et al.
Veröffentlicht: (2024)
Decoding Machine Translationese in English-Chinese News: LLMs vs. NMTs
von: Kong, Delu, et al.
Veröffentlicht: (2025)
von: Kong, Delu, et al.
Veröffentlicht: (2025)
The Comparison of Translationese in Machine Translation and Human Transation in terms of Translation Relations
von: Zhou, Fan
Veröffentlicht: (2024)
von: Zhou, Fan
Veröffentlicht: (2024)
Lost in Translationese? Reducing Translation Effect Using Abstract Meaning Representation
von: Wein, Shira, et al.
Veröffentlicht: (2023)
von: Wein, Shira, et al.
Veröffentlicht: (2023)
Fusing Semantic, Lexical, and Domain Perspectives for Recipe Similarity Estimation
von: Kjorvezir, Denica, et al.
Veröffentlicht: (2026)
von: Kjorvezir, Denica, et al.
Veröffentlicht: (2026)
Mitigating Translationese in Low-resource Languages: The Storyboard Approach
von: Kuwanto, Garry, et al.
Veröffentlicht: (2024)
von: Kuwanto, Garry, et al.
Veröffentlicht: (2024)
The Impact of Language Adapters in Cross-Lingual Transfer for NLU
von: Kunz, Jenny, et al.
Veröffentlicht: (2024)
von: Kunz, Jenny, et al.
Veröffentlicht: (2024)
How Lexical is Bilingual Lexicon Induction?
von: Kohli, Harsh, et al.
Veröffentlicht: (2024)
von: Kohli, Harsh, et al.
Veröffentlicht: (2024)
Vector Retrieval with Similarity and Diversity: How Hard Is It?
von: Gao, Hang, et al.
Veröffentlicht: (2024)
von: Gao, Hang, et al.
Veröffentlicht: (2024)
Code-Mixed Probes Show How Pre-Trained Models Generalise On Code-Switched Text
von: De Leon, Frances A. Laureano, et al.
Veröffentlicht: (2024)
von: De Leon, Frances A. Laureano, et al.
Veröffentlicht: (2024)
The Structural Sources of Verb Meaning Revisited: Large Language Models Display Syntactic Bootstrapping
von: Zhu, Xiaomeng, et al.
Veröffentlicht: (2025)
von: Zhu, Xiaomeng, et al.
Veröffentlicht: (2025)
Mitigating Frequency Bias and Anisotropy in Language Model Pre-Training with Syntactic Smoothing
von: Martinez, Richard Diehl, et al.
Veröffentlicht: (2024)
von: Martinez, Richard Diehl, et al.
Veröffentlicht: (2024)
Towards Tailored Recovery of Lexical Diversity in Literary Machine Translation
von: Ploeger, Esther, et al.
Veröffentlicht: (2024)
von: Ploeger, Esther, et al.
Veröffentlicht: (2024)
Probing Large Language Models for Scalar Adjective Lexical Semantics and Scalar Diversity Pragmatics
von: Lin, Fangru, et al.
Veröffentlicht: (2024)
von: Lin, Fangru, et al.
Veröffentlicht: (2024)
Modelling Child Learning and Parsing of Long-range Syntactic Dependencies
von: Mahon, Louis, et al.
Veröffentlicht: (2025)
von: Mahon, Louis, et al.
Veröffentlicht: (2025)
Transfer Learning for an Endangered Slavic Variety: Dependency Parsing in Pomak Across Contact-Shaped Dialects
von: Karakaş, Sercan
Veröffentlicht: (2026)
von: Karakaş, Sercan
Veröffentlicht: (2026)
DialUp! Modeling the Language Continuum by Adapting Models to Dialects and Dialects to Models
von: Bafna, Niyati, et al.
Veröffentlicht: (2025)
von: Bafna, Niyati, et al.
Veröffentlicht: (2025)
Mitigating Translationese Bias in Multilingual LLM-as-a-Judge via Disentangled Information Bottleneck
von: Zhang, Hongbin, et al.
Veröffentlicht: (2026)
von: Zhang, Hongbin, et al.
Veröffentlicht: (2026)
How Trustworthy are Open-Source LLMs? An Assessment under Malicious Demonstrations Shows their Vulnerabilities
von: Mo, Lingbo, et al.
Veröffentlicht: (2023)
von: Mo, Lingbo, et al.
Veröffentlicht: (2023)
Learning the Wrong Lessons: Syntactic-Domain Spurious Correlations in Language Models
von: Shaib, Chantal, et al.
Veröffentlicht: (2025)
von: Shaib, Chantal, et al.
Veröffentlicht: (2025)
How Stylistic Similarity Shapes Preferences in Dialogue Dataset with User and Third Party Evaluations
von: Numaya, Ikumi, et al.
Veröffentlicht: (2025)
von: Numaya, Ikumi, et al.
Veröffentlicht: (2025)
ArabicDialectHub: A Cross-Dialectal Arabic Learning Resource and Platform
von: Lahlou, Salem
Veröffentlicht: (2026)
von: Lahlou, Salem
Veröffentlicht: (2026)
Mechanisms vs. Outcomes: Probing for Syntax Fails to Explain Performance on Targeted Syntactic Evaluations
von: Agarwal, Ananth, et al.
Veröffentlicht: (2025)
von: Agarwal, Ananth, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
A Dataset for Probing Translationese Preferences in English-to-Swedish Translation
von: Kunz, Jenny, et al.
Veröffentlicht: (2026) -
Lost in Literalism: How Supervised Training Shapes Translationese in LLMs
von: Li, Yafu, et al.
Veröffentlicht: (2025) -
Translationese-index: Using Likelihood Ratios for Graded and Generalizable Measurement of Translationese
von: Liu, Yikang, et al.
Veröffentlicht: (2025) -
Train More Parameters But Mind Their Placement: Insights into Language Adaptation with PEFT
von: Kunz, Jenny
Veröffentlicht: (2024) -
Pretraining Language Models Using Translationese
von: Doshi, Meet, et al.
Veröffentlicht: (2024)