Understanding and Mitigating the Uncertainty in Zero-Shot Translation
Fuente:
arXiv
Salvato in:
| Autori principali: | Wang, Wenxuan, Jiao, Wenxiang, Wang, Shuo, Tu, Zhaopeng, Lyu, Michael R. |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2022
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
On the Shortcut Learning in Multilingual Neural Machine Translation
di: Wang, Wenxuan, et al.
Pubblicazione: (2024)
di: Wang, Wenxuan, et al.
Pubblicazione: (2024)
Not All Countries Celebrate Thanksgiving: On the Cultural Dominance in Large Language Models
di: Wang, Wenxuan, et al.
Pubblicazione: (2023)
di: Wang, Wenxuan, et al.
Pubblicazione: (2023)
All Languages Matter: On the Multilingual Safety of Large Language Models
di: Wang, Wenxuan, et al.
Pubblicazione: (2023)
di: Wang, Wenxuan, et al.
Pubblicazione: (2023)
Improving Machine Translation with Human Feedback: An Exploration of Quality Estimation as a Reward Model
di: He, Zhiwei, et al.
Pubblicazione: (2024)
di: He, Zhiwei, et al.
Pubblicazione: (2024)
MT-R1-Zero: Advancing LLM-based Machine Translation via R1-Zero-like Reinforcement Learning
di: Feng, Zhaopeng, et al.
Pubblicazione: (2025)
di: Feng, Zhaopeng, et al.
Pubblicazione: (2025)
MT$^{3}$: Scaling MLLM-based Text Image Machine Translation via Multi-Task Reinforcement Learning
di: Feng, Zhaopeng, et al.
Pubblicazione: (2025)
di: Feng, Zhaopeng, et al.
Pubblicazione: (2025)
Language Representation Favored Zero-Shot Cross-Domain Cognitive Diagnosis
di: Liu, Shuo, et al.
Pubblicazione: (2025)
di: Liu, Shuo, et al.
Pubblicazione: (2025)
Ladder: A Model-Agnostic Framework Boosting LLM-based Machine Translation to the Next Level
di: Feng, Zhaopeng, et al.
Pubblicazione: (2024)
di: Feng, Zhaopeng, et al.
Pubblicazione: (2024)
How Far Are We on the Decision-Making of LLMs? Evaluating LLMs' Gaming Ability in Multi-Agent Environments
di: Huang, Jen-tse, et al.
Pubblicazione: (2024)
di: Huang, Jen-tse, et al.
Pubblicazione: (2024)
Refuse Whenever You Feel Unsafe: Improving Safety in LLMs via Decoupled Refusal Training
di: Yuan, Youliang, et al.
Pubblicazione: (2024)
di: Yuan, Youliang, et al.
Pubblicazione: (2024)
Agent Instructs Large Language Models to be General Zero-Shot Reasoners
di: Crispino, Nicholas, et al.
Pubblicazione: (2023)
di: Crispino, Nicholas, et al.
Pubblicazione: (2023)
SSR-Zero: Simple Self-Rewarding Reinforcement Learning for Machine Translation
di: Yang, Wenjie, et al.
Pubblicazione: (2025)
di: Yang, Wenjie, et al.
Pubblicazione: (2025)
Iterative Data Smoothing: Mitigating Reward Overfitting and Overoptimization in RLHF
di: Zhu, Banghua, et al.
Pubblicazione: (2024)
di: Zhu, Banghua, et al.
Pubblicazione: (2024)
TourPlanner: A Competitive Consensus Framework with Constraint-Gated Reinforcement Learning for Travel Planning
di: Wang, Yinuo, et al.
Pubblicazione: (2026)
di: Wang, Yinuo, et al.
Pubblicazione: (2026)
Emergent Communication Pretraining for Few-Shot Machine Translation
di: Li, Yaoyiran, et al.
Pubblicazione: (2020)
di: Li, Yaoyiran, et al.
Pubblicazione: (2020)
Understanding and Mitigating Bias Inheritance in LLM-based Data Augmentation on Downstream Tasks
di: Li, Miaomiao, et al.
Pubblicazione: (2025)
di: Li, Miaomiao, et al.
Pubblicazione: (2025)
LLMs are not Zero-Shot Reasoners for Biomedical Information Extraction
di: Nagar, Aishik, et al.
Pubblicazione: (2024)
di: Nagar, Aishik, et al.
Pubblicazione: (2024)
Understanding and Mitigating Spurious Signal Amplification in Test-Time Reinforcement Learning for Math Reasoning
di: Yu, Yongcan, et al.
Pubblicazione: (2026)
di: Yu, Yongcan, et al.
Pubblicazione: (2026)
GLiREL -- Generalist Model for Zero-Shot Relation Extraction
di: Boylan, Jack, et al.
Pubblicazione: (2025)
di: Boylan, Jack, et al.
Pubblicazione: (2025)
Leveraging Zero-Shot Prompting for Efficient Language Model Distillation
di: Vöge, Lukas, et al.
Pubblicazione: (2024)
di: Vöge, Lukas, et al.
Pubblicazione: (2024)
SPC: Evolving Self-Play Critic via Adversarial Games for LLM Reasoning
di: Chen, Jiaqi, et al.
Pubblicazione: (2025)
di: Chen, Jiaqi, et al.
Pubblicazione: (2025)
Identifying the Achilles' Heel: An Iterative Method for Dynamically Uncovering Factual Errors in Large Language Models
di: Wang, Wenxuan, et al.
Pubblicazione: (2024)
di: Wang, Wenxuan, et al.
Pubblicazione: (2024)
Understanding and Mitigating Tokenization Bias in Language Models
di: Phan, Buu, et al.
Pubblicazione: (2024)
di: Phan, Buu, et al.
Pubblicazione: (2024)
Understanding and Mitigating Dataset Corruption in LLM Steering
di: Anderson, Cullen, et al.
Pubblicazione: (2026)
di: Anderson, Cullen, et al.
Pubblicazione: (2026)
ZeroUnlearn: Few-Shot Knowledge Unlearning in Large Language Models
di: Lin, Yujie, et al.
Pubblicazione: (2026)
di: Lin, Yujie, et al.
Pubblicazione: (2026)
Spotting LLMs With Binoculars: Zero-Shot Detection of Machine-Generated Text
di: Hans, Abhimanyu, et al.
Pubblicazione: (2024)
di: Hans, Abhimanyu, et al.
Pubblicazione: (2024)
LoopTool: Closing the Data-Training Loop for Robust LLM Tool Calls
di: Zhang, Kangning, et al.
Pubblicazione: (2025)
di: Zhang, Kangning, et al.
Pubblicazione: (2025)
Prompting Large Language Models for Zero-Shot Clinical Prediction with Structured Longitudinal Electronic Health Record Data
di: Zhu, Yinghao, et al.
Pubblicazione: (2024)
di: Zhu, Yinghao, et al.
Pubblicazione: (2024)
Visualizing Uncertainty in Translation Tasks: An Evaluation of LLM Performance and Confidence Metrics
di: Park, Jin Hyun, et al.
Pubblicazione: (2025)
di: Park, Jin Hyun, et al.
Pubblicazione: (2025)
Critical Tokens Matter: Token-Level Contrastive Estimation Enhances LLM's Reasoning Capability
di: Lin, Zicheng, et al.
Pubblicazione: (2024)
di: Lin, Zicheng, et al.
Pubblicazione: (2024)
An Analysis and Mitigation of the Reversal Curse
di: Lv, Ang, et al.
Pubblicazione: (2023)
di: Lv, Ang, et al.
Pubblicazione: (2023)
Mitigating Hallucinated Translations in Large Language Models with Hallucination-focused Preference Optimization
di: Tang, Zilu, et al.
Pubblicazione: (2025)
di: Tang, Zilu, et al.
Pubblicazione: (2025)
Evidence-Focused Fact Summarization for Knowledge-Augmented Zero-Shot Question Answering
di: Ko, Sungho, et al.
Pubblicazione: (2024)
di: Ko, Sungho, et al.
Pubblicazione: (2024)
Layer Swapping for Zero-Shot Cross-Lingual Transfer in Large Language Models
di: Bandarkar, Lucas, et al.
Pubblicazione: (2024)
di: Bandarkar, Lucas, et al.
Pubblicazione: (2024)
Learning to Reason from Feedback at Test-Time
di: Li, Yanyang, et al.
Pubblicazione: (2025)
di: Li, Yanyang, et al.
Pubblicazione: (2025)
Uncertainty-Aware Fusion: An Ensemble Framework for Mitigating Hallucinations in Large Language Models
di: Dey, Prasenjit, et al.
Pubblicazione: (2025)
di: Dey, Prasenjit, et al.
Pubblicazione: (2025)
The Right Time Matters: Data Arrangement Affects Zero-Shot Generalization in Instruction Tuning
di: He, Bingxiang, et al.
Pubblicazione: (2024)
di: He, Bingxiang, et al.
Pubblicazione: (2024)
CCRS: A Zero-Shot LLM-as-a-Judge Framework for Comprehensive RAG Evaluation
di: Muhamed, Aashiq
Pubblicazione: (2025)
di: Muhamed, Aashiq
Pubblicazione: (2025)
CURE: Controlled Unlearning for Robust Embeddings -- Mitigating Conceptual Shortcuts in Pre-Trained Language Models
di: Kocak, Aysenur, et al.
Pubblicazione: (2025)
di: Kocak, Aysenur, et al.
Pubblicazione: (2025)
Quantifying and Mitigating Self-Preference Bias of LLM Judges
di: Yang, Jinming, et al.
Pubblicazione: (2026)
di: Yang, Jinming, et al.
Pubblicazione: (2026)
Documenti analoghi
-
On the Shortcut Learning in Multilingual Neural Machine Translation
di: Wang, Wenxuan, et al.
Pubblicazione: (2024) -
Not All Countries Celebrate Thanksgiving: On the Cultural Dominance in Large Language Models
di: Wang, Wenxuan, et al.
Pubblicazione: (2023) -
All Languages Matter: On the Multilingual Safety of Large Language Models
di: Wang, Wenxuan, et al.
Pubblicazione: (2023) -
Improving Machine Translation with Human Feedback: An Exploration of Quality Estimation as a Reward Model
di: He, Zhiwei, et al.
Pubblicazione: (2024) -
MT-R1-Zero: Advancing LLM-based Machine Translation via R1-Zero-like Reinforcement Learning
di: Feng, Zhaopeng, et al.
Pubblicazione: (2025)