Sigmoid Head for Quality Estimation under Language Ambiguity
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Dinh, Tu Anh, Niehues, Jan |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Are Generative Models Underconfident? Better Quality Estimation with Boosted Model Probability
von: Dinh, Tu Anh, et al.
Veröffentlicht: (2025)
von: Dinh, Tu Anh, et al.
Veröffentlicht: (2025)
Quality Estimation with $k$-nearest Neighbors and Automatic Evaluation for Model-specific Quality Estimation
von: Dinh, Tu Anh, et al.
Veröffentlicht: (2024)
von: Dinh, Tu Anh, et al.
Veröffentlicht: (2024)
Knockout LLM Assessment: Using Large Language Models for Evaluations through Iterative Pairwise Comparisons
von: Sandan, Isik Baran, et al.
Veröffentlicht: (2025)
von: Sandan, Isik Baran, et al.
Veröffentlicht: (2025)
COMET-poly: Machine Translation Metric Grounded in Other Candidates
von: Züfle, Maike, et al.
Veröffentlicht: (2025)
von: Züfle, Maike, et al.
Veröffentlicht: (2025)
SciEx: Benchmarking Large Language Models on Scientific Exams with Human Expert Grading and Automatic Grading
von: Dinh, Tu Anh, et al.
Veröffentlicht: (2024)
von: Dinh, Tu Anh, et al.
Veröffentlicht: (2024)
Tokenization and Morphology in Multilingual Language Models: A Comparative Analysis of mT5 and ByT5
von: Dang, Thao Anh, et al.
Veröffentlicht: (2024)
von: Dang, Thao Anh, et al.
Veröffentlicht: (2024)
SECURA: Sigmoid-Enhanced CUR Decomposition with Uninterrupted Retention and Low-Rank Adaptation in Large Language Models
von: Zhang, Yuxuan
Veröffentlicht: (2025)
von: Zhang, Yuxuan
Veröffentlicht: (2025)
CRISP: Persistent Concept Unlearning via Sparse Autoencoders
von: Ashuach, Tomer, et al.
Veröffentlicht: (2025)
von: Ashuach, Tomer, et al.
Veröffentlicht: (2025)
Is Less More? Quality, Quantity and Context in Idiom Processing with Natural Language Models
von: Knietaite, Agne, et al.
Veröffentlicht: (2024)
von: Knietaite, Agne, et al.
Veröffentlicht: (2024)
Towards Human Understanding of Paraphrase Types in Large Language Models
von: Meier, Dominik, et al.
Veröffentlicht: (2024)
von: Meier, Dominik, et al.
Veröffentlicht: (2024)
Ambiguity in LLMs is a concept missing problem
von: Hu, Zhibo, et al.
Veröffentlicht: (2025)
von: Hu, Zhibo, et al.
Veröffentlicht: (2025)
Is Textual Similarity Invariant under Machine Translation? Evidence Based on the Political Manifesto Corpus
von: Boratyn, Daria, et al.
Veröffentlicht: (2026)
von: Boratyn, Daria, et al.
Veröffentlicht: (2026)
German Text Simplification: Finetuning Large Language Models with Semi-Synthetic Data
von: Klöser, Lars, et al.
Veröffentlicht: (2024)
von: Klöser, Lars, et al.
Veröffentlicht: (2024)
Towards Fundamental Language Models: Does Linguistic Competence Scale with Model Size?
von: Collado-Montañez, Jaime, et al.
Veröffentlicht: (2025)
von: Collado-Montañez, Jaime, et al.
Veröffentlicht: (2025)
BOUQuET: dataset, Benchmark and Open initiative for Universal Quality Evaluation in Translation
von: The Omnilingual MT Team, et al.
Veröffentlicht: (2025)
von: The Omnilingual MT Team, et al.
Veröffentlicht: (2025)
Cross-lingual Human-Preference Alignment for Neural Machine Translation with Direct Quality Optimization
von: Uhlig, Kaden, et al.
Veröffentlicht: (2024)
von: Uhlig, Kaden, et al.
Veröffentlicht: (2024)
Encoder-Decoder Framework for Interactive Free Verses with Generation with Controllable High-Quality Rhyming
von: Pasini, Tommaso, et al.
Veröffentlicht: (2024)
von: Pasini, Tommaso, et al.
Veröffentlicht: (2024)
RomanLens: The Role Of Latent Romanization In Multilinguality In LLMs
von: Saji, Alan, et al.
Veröffentlicht: (2025)
von: Saji, Alan, et al.
Veröffentlicht: (2025)
Text-Based Approaches to Item Difficulty Modeling in Large-Scale Assessments: A Systematic Review
von: Peters, Sydney, et al.
Veröffentlicht: (2025)
von: Peters, Sydney, et al.
Veröffentlicht: (2025)
Whose Facts Win? LLM Source Preferences under Knowledge Conflicts
von: Schuster, Jakob, et al.
Veröffentlicht: (2026)
von: Schuster, Jakob, et al.
Veröffentlicht: (2026)
Head-Specific Intervention Can Induce Misaligned AI Coordination in Large Language Models
von: Darm, Paul, et al.
Veröffentlicht: (2025)
von: Darm, Paul, et al.
Veröffentlicht: (2025)
Budget-Xfer: Budget-Constrained Source Language Selection for Cross-Lingual Transfer to African Languages
von: Idris, Tewodros Kederalah, et al.
Veröffentlicht: (2026)
von: Idris, Tewodros Kederalah, et al.
Veröffentlicht: (2026)
Calibrated Confidence Estimation for Tabular Question Answering
von: Voss, Lukas
Veröffentlicht: (2026)
von: Voss, Lukas
Veröffentlicht: (2026)
Large Language Models for Biomedical Article Classification
von: Proboszcz, Jakub, et al.
Veröffentlicht: (2026)
von: Proboszcz, Jakub, et al.
Veröffentlicht: (2026)
Precise Length Control in Large Language Models
von: Butcher, Bradley, et al.
Veröffentlicht: (2024)
von: Butcher, Bradley, et al.
Veröffentlicht: (2024)
A Review of the Marathi Natural Language Processing
von: Dani, Asang, et al.
Veröffentlicht: (2024)
von: Dani, Asang, et al.
Veröffentlicht: (2024)
What Drives Performance in Multilingual Language Models?
von: Nezhad, Sina Bagheri, et al.
Veröffentlicht: (2024)
von: Nezhad, Sina Bagheri, et al.
Veröffentlicht: (2024)
Omnilingual MT: Machine Translation for 1,600 Languages
von: Omnilingual MT Team, et al.
Veröffentlicht: (2026)
von: Omnilingual MT Team, et al.
Veröffentlicht: (2026)
Strategy Adaptation in Large Language Model Werewolf Agents
von: Nakamori, Fuya, et al.
Veröffentlicht: (2025)
von: Nakamori, Fuya, et al.
Veröffentlicht: (2025)
Mitigating Translationese in Low-resource Languages: The Storyboard Approach
von: Kuwanto, Garry, et al.
Veröffentlicht: (2024)
von: Kuwanto, Garry, et al.
Veröffentlicht: (2024)
Contextualising Levels of Language Resourcedness that affect NLP tasks
von: Keet, C. Maria, et al.
Veröffentlicht: (2023)
von: Keet, C. Maria, et al.
Veröffentlicht: (2023)
PL-Guard: Benchmarking Language Model Safety for Polish
von: Krasnodębska, Aleksandra, et al.
Veröffentlicht: (2025)
von: Krasnodębska, Aleksandra, et al.
Veröffentlicht: (2025)
Socially Responsible Data for Large Multilingual Language Models
von: Smart, Andrew, et al.
Veröffentlicht: (2024)
von: Smart, Andrew, et al.
Veröffentlicht: (2024)
SeLeRoSa: Sentence-Level Romanian Satire Detection Dataset
von: Smădu, Răzvan-Alexandru, et al.
Veröffentlicht: (2025)
von: Smădu, Răzvan-Alexandru, et al.
Veröffentlicht: (2025)
RUQuant: Towards Refining Uniform Quantization for Large Language Models
von: Liu, Han, et al.
Veröffentlicht: (2026)
von: Liu, Han, et al.
Veröffentlicht: (2026)
Large Language Models for Persian $ \leftrightarrow $ English Idiom Translation
von: Rezaeimanesh, Sara, et al.
Veröffentlicht: (2024)
von: Rezaeimanesh, Sara, et al.
Veröffentlicht: (2024)
Qomhra: A Bilingual Irish and English Large Language Model
von: McInerney, Joseph, et al.
Veröffentlicht: (2025)
von: McInerney, Joseph, et al.
Veröffentlicht: (2025)
Dialect Normalization using Large Language Models and Morphological Rules
von: Dimakis, Antonios, et al.
Veröffentlicht: (2025)
von: Dimakis, Antonios, et al.
Veröffentlicht: (2025)
Morphological Analysis for the Maltese Language: The Challenges of a Hybrid System
von: Borg, Claudia, et al.
Veröffentlicht: (2017)
von: Borg, Claudia, et al.
Veröffentlicht: (2017)
Synthetic Voice Data for Automatic Speech Recognition in African Languages
von: DeRenzi, Brian, et al.
Veröffentlicht: (2025)
von: DeRenzi, Brian, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Are Generative Models Underconfident? Better Quality Estimation with Boosted Model Probability
von: Dinh, Tu Anh, et al.
Veröffentlicht: (2025) -
Quality Estimation with $k$-nearest Neighbors and Automatic Evaluation for Model-specific Quality Estimation
von: Dinh, Tu Anh, et al.
Veröffentlicht: (2024) -
Knockout LLM Assessment: Using Large Language Models for Evaluations through Iterative Pairwise Comparisons
von: Sandan, Isik Baran, et al.
Veröffentlicht: (2025) -
COMET-poly: Machine Translation Metric Grounded in Other Candidates
von: Züfle, Maike, et al.
Veröffentlicht: (2025) -
SciEx: Benchmarking Large Language Models on Scientific Exams with Human Expert Grading and Automatic Grading
von: Dinh, Tu Anh, et al.
Veröffentlicht: (2024)