Towards Reliable Machine Translation: Scaling LLMs for Critical Error Detection and Safety
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Chopra, Muskaan, Sparrenberg, Lorenz, Sifa, Rafet |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
SynCED-EnDe 2025: A Synthetic and Curated English - German Dataset for Critical Error Detection in Machine Translation
von: Chopra, Muskaan, et al.
Veröffentlicht: (2025)
von: Chopra, Muskaan, et al.
Veröffentlicht: (2025)
How Small Can You Go? Compact Language Models for On-Device Critical Error Detection in Machine Translation
von: Chopra, Muskaan, et al.
Veröffentlicht: (2025)
von: Chopra, Muskaan, et al.
Veröffentlicht: (2025)
Knowing When Not to Predict: Self Supervised Learning and Abstention for Safer DR Screening
von: Chopra, Muskaan, et al.
Veröffentlicht: (2026)
von: Chopra, Muskaan, et al.
Veröffentlicht: (2026)
From Retinal Pixels to Patients: Evolution of Deep Learning Research in Diabetic Retinopathy Screening
von: Chopra, Muskaan, et al.
Veröffentlicht: (2025)
von: Chopra, Muskaan, et al.
Veröffentlicht: (2025)
A Survey on Current Trends and Recent Advances in Text Anonymization
von: Deußer, Tobias, et al.
Veröffentlicht: (2025)
von: Deußer, Tobias, et al.
Veröffentlicht: (2025)
Generalizing Abstention for Noise-Robust Learning in Medical Image Segmentation
von: Moustafa, Wesam, et al.
Veröffentlicht: (2026)
von: Moustafa, Wesam, et al.
Veröffentlicht: (2026)
Towards Unified Multimodal Financial Forecasting: Integrating Sentiment Embeddings and Market Indicators via Cross-Modal Attention
von: Khanna, Sarthak, et al.
Veröffentlicht: (2025)
von: Khanna, Sarthak, et al.
Veröffentlicht: (2025)
History Rhymes: Macro-Contextual Retrieval for Robust Financial Forecasting
von: Khanna, Sarthak, et al.
Veröffentlicht: (2025)
von: Khanna, Sarthak, et al.
Veröffentlicht: (2025)
[Vision Paper] PRObot: Enhancing Patient-Reported Outcome Measures for Diabetic Retinopathy using Chatbots and Generative AI
von: Pielka, Maren, et al.
Veröffentlicht: (2024)
von: Pielka, Maren, et al.
Veröffentlicht: (2024)
Pointer-Guided Pre-Training: Infusing Large Language Models with Paragraph-Level Contextual Awareness
von: Hillebrand, Lars, et al.
Veröffentlicht: (2024)
von: Hillebrand, Lars, et al.
Veröffentlicht: (2024)
Interpretable Topic Extraction and Word Embedding Learning using row-stochastic DEDICOM
von: Hillebrand, Lars, et al.
Veröffentlicht: (2025)
von: Hillebrand, Lars, et al.
Veröffentlicht: (2025)
Multi-Modal Vision vs. Text-Based Parsing: Benchmarking LLM Strategies for Invoice Processing
von: Berghaus, David, et al.
Veröffentlicht: (2025)
von: Berghaus, David, et al.
Veröffentlicht: (2025)
Can LLMs Detect Intrinsic Hallucinations in Paraphrasing and Machine Translation?
von: Gogoulou, Evangelia, et al.
Veröffentlicht: (2025)
von: Gogoulou, Evangelia, et al.
Veröffentlicht: (2025)
Is continuous CoT better suited for multi-lingual reasoning?
von: Bashir, Ali Hamza, et al.
Veröffentlicht: (2026)
von: Bashir, Ali Hamza, et al.
Veröffentlicht: (2026)
Should We be Pedantic About Reasoning Errors in Machine Translation?
von: Bao, Calvin, et al.
Veröffentlicht: (2026)
von: Bao, Calvin, et al.
Veröffentlicht: (2026)
Advancing Risk and Quality Assurance: A RAG Chatbot for Improved Regulatory Compliance
von: Hillebrand, Lars, et al.
Veröffentlicht: (2025)
von: Hillebrand, Lars, et al.
Veröffentlicht: (2025)
Improving Non-autoregressive Machine Translation with Error Exposure and Consistency Regularization
von: Chen, Xinran, et al.
Veröffentlicht: (2024)
von: Chen, Xinran, et al.
Veröffentlicht: (2024)
Towards Reliable Evaluation of Behavior Steering Interventions in LLMs
von: Pres, Itamar, et al.
Veröffentlicht: (2024)
von: Pres, Itamar, et al.
Veröffentlicht: (2024)
Reasoning LLMs in the Medical Domain: A Literature Survey
von: Berger, Armin, et al.
Veröffentlicht: (2025)
von: Berger, Armin, et al.
Veröffentlicht: (2025)
Guiding Large Language Models to Post-Edit Machine Translation with Error Annotations
von: Ki, Dayeon, et al.
Veröffentlicht: (2024)
von: Ki, Dayeon, et al.
Veröffentlicht: (2024)
Domain-Adaptation through Synthetic Data: Fine-Tuning Large Language Models for German Law
von: Bashir, Ali Hamza, et al.
Veröffentlicht: (2026)
von: Bashir, Ali Hamza, et al.
Veröffentlicht: (2026)
Towards Automated Regulatory Compliance Verification in Financial Auditing with Large Language Models
von: Berger, Armin, et al.
Veröffentlicht: (2025)
von: Berger, Armin, et al.
Veröffentlicht: (2025)
CycleDistill: Bootstrapping Machine Translation using LLMs with Cyclical Distillation
von: Halder, Deepon, et al.
Veröffentlicht: (2025)
von: Halder, Deepon, et al.
Veröffentlicht: (2025)
Towards Understanding and Improving Knowledge Distillation for Neural Machine Translation
von: Zhang, Songming, et al.
Veröffentlicht: (2023)
von: Zhang, Songming, et al.
Veröffentlicht: (2023)
Minimum Bayes Risk Decoding for Error Span Detection in Reference-Free Automatic Machine Translation Evaluation
von: Lyu, Boxuan, et al.
Veröffentlicht: (2025)
von: Lyu, Boxuan, et al.
Veröffentlicht: (2025)
Double-Calibration: Towards Reliable LLMs via Calibrating Knowledge and Reasoning Confidence
von: Lu, Yuyin, et al.
Veröffentlicht: (2026)
von: Lu, Yuyin, et al.
Veröffentlicht: (2026)
Scaling Laws of Decoder-Only Models on the Multilingual Machine Translation Task
von: Caillaut, Gaëtan, et al.
Veröffentlicht: (2024)
von: Caillaut, Gaëtan, et al.
Veröffentlicht: (2024)
Scaling, Simplification, and Adaptation: Lessons from Pretraining on Machine-Translated Text
von: Velasco, Dan John, et al.
Veröffentlicht: (2025)
von: Velasco, Dan John, et al.
Veröffentlicht: (2025)
LiveCLKTBench: Towards Reliable Evaluation of Cross-Lingual Knowledge Transfer in Multilingual LLMs
von: Guo, Pei-Fu, et al.
Veröffentlicht: (2025)
von: Guo, Pei-Fu, et al.
Veröffentlicht: (2025)
Can LLMs Detect Their Confabulations? Estimating Reliability in Uncertainty-Aware Language Models
von: Zhou, Tianyi, et al.
Veröffentlicht: (2025)
von: Zhou, Tianyi, et al.
Veröffentlicht: (2025)
ICR Probe: Tracking Hidden State Dynamics for Reliable Hallucination Detection in LLMs
von: Zhang, Zhenliang, et al.
Veröffentlicht: (2025)
von: Zhang, Zhenliang, et al.
Veröffentlicht: (2025)
Tuning LLMs with Contrastive Alignment Instructions for Machine Translation in Unseen, Low-resource Languages
von: Mao, Zhuoyuan, et al.
Veröffentlicht: (2024)
von: Mao, Zhuoyuan, et al.
Veröffentlicht: (2024)
The Homogenization Problem in LLMs: Towards Meaningful Diversity in AI Safety
von: Rios-Sialer, Ian
Veröffentlicht: (2026)
von: Rios-Sialer, Ian
Veröffentlicht: (2026)
GeoResponder: Towards Building Geospatial LLMs for Time-Critical Disaster Response
von: Zguir, Ahmed El Fekih, et al.
Veröffentlicht: (2025)
von: Zguir, Ahmed El Fekih, et al.
Veröffentlicht: (2025)
How Good Are LLMs for Literary Translation, Really? Literary Translation Evaluation with Humans and LLMs
von: Zhang, Ran, et al.
Veröffentlicht: (2024)
von: Zhang, Ran, et al.
Veröffentlicht: (2024)
Towards Privacy-Preserving Machine Translation at the Inference Stage: A New Task and Benchmark
von: Shao, Wei, et al.
Veröffentlicht: (2026)
von: Shao, Wei, et al.
Veröffentlicht: (2026)
PKU-SafeRLHF: Towards Multi-Level Safety Alignment for LLMs with Human Preference
von: Ji, Jiaming, et al.
Veröffentlicht: (2024)
von: Ji, Jiaming, et al.
Veröffentlicht: (2024)
Languages Still Left Behind: Toward a Better Multilingual Machine Translation Benchmark
von: Taguchi, Chihiro, et al.
Veröffentlicht: (2025)
von: Taguchi, Chihiro, et al.
Veröffentlicht: (2025)
This Is Your Doge, If It Please You: Exploring Deception and Robustness in Mixture of LLMs
von: Wolf, Lorenz, et al.
Veröffentlicht: (2025)
von: Wolf, Lorenz, et al.
Veröffentlicht: (2025)
Can LLMs Evaluate What They Cannot Annotate? Revisiting LLM Reliability in Hate Speech Detection
von: Piot, Paloma, et al.
Veröffentlicht: (2025)
von: Piot, Paloma, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
SynCED-EnDe 2025: A Synthetic and Curated English - German Dataset for Critical Error Detection in Machine Translation
von: Chopra, Muskaan, et al.
Veröffentlicht: (2025) -
How Small Can You Go? Compact Language Models for On-Device Critical Error Detection in Machine Translation
von: Chopra, Muskaan, et al.
Veröffentlicht: (2025) -
Knowing When Not to Predict: Self Supervised Learning and Abstention for Safer DR Screening
von: Chopra, Muskaan, et al.
Veröffentlicht: (2026) -
From Retinal Pixels to Patients: Evolution of Deep Learning Research in Diabetic Retinopathy Screening
von: Chopra, Muskaan, et al.
Veröffentlicht: (2025) -
A Survey on Current Trends and Recent Advances in Text Anonymization
von: Deußer, Tobias, et al.
Veröffentlicht: (2025)