Contrastive Preference Optimization: Pushing the Boundaries of LLM Performance in Machine Translation
Fuente:
arXiv
Salvato in:
| Autori principali: | Xu, Haoran, Sharaf, Amr, Chen, Yunmo, Tan, Weiting, Shen, Lingfeng, Van Durme, Benjamin, Murray, Kenton, Kim, Young Jin |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
A Paradigm Shift in Machine Translation: Boosting Translation Performance of Large Language Models
di: Xu, Haoran, et al.
Pubblicazione: (2023)
di: Xu, Haoran, et al.
Pubblicazione: (2023)
Streaming Sequence Transduction through Dynamic Compression
di: Tan, Weiting, et al.
Pubblicazione: (2024)
di: Tan, Weiting, et al.
Pubblicazione: (2024)
Upsample or Upweight? Balanced Training on Heavily Imbalanced Datasets
di: Li, Tianjian, et al.
Pubblicazione: (2024)
di: Li, Tianjian, et al.
Pubblicazione: (2024)
The Language Barrier: Dissecting Safety Challenges of LLMs in Multilingual Contexts
di: Shen, Lingfeng, et al.
Pubblicazione: (2024)
di: Shen, Lingfeng, et al.
Pubblicazione: (2024)
DiffNorm: Self-Supervised Normalization for Non-autoregressive Speech-to-speech Translation
di: Tan, Weiting, et al.
Pubblicazione: (2024)
di: Tan, Weiting, et al.
Pubblicazione: (2024)
Unified Multimodal Uncertain Inference
di: Zhang, Dengjia, et al.
Pubblicazione: (2026)
di: Zhang, Dengjia, et al.
Pubblicazione: (2026)
MultiMUC: Multilingual Template Filling on MUC-4
di: Gantt, William, et al.
Pubblicazione: (2024)
di: Gantt, William, et al.
Pubblicazione: (2024)
Learning to Retrieve Iteratively for In-Context Learning
di: Chen, Yunmo, et al.
Pubblicazione: (2024)
di: Chen, Yunmo, et al.
Pubblicazione: (2024)
Exploring Representational Disparities Between Multilingual and Bilingual Translation Models
di: Verma, Neha, et al.
Pubblicazione: (2023)
di: Verma, Neha, et al.
Pubblicazione: (2023)
X-ALMA: Plug & Play Modules and Adaptive Rejection for Quality Translation at Scale
di: Xu, Haoran, et al.
Pubblicazione: (2024)
di: Xu, Haoran, et al.
Pubblicazione: (2024)
HLTCOE Evaluation Team at TREC 2025: VQA Track
di: Zhang, Dengjia, et al.
Pubblicazione: (2025)
di: Zhang, Dengjia, et al.
Pubblicazione: (2025)
RE-AdaptIR: Improving Information Retrieval through Reverse Engineered Adaptation
di: Fleshman, William, et al.
Pubblicazione: (2024)
di: Fleshman, William, et al.
Pubblicazione: (2024)
Bonsai: Interpretable Tree-Adaptive Grounded Reasoning
di: Sanders, Kate, et al.
Pubblicazione: (2025)
di: Sanders, Kate, et al.
Pubblicazione: (2025)
SEQR: Secure and Efficient QR-based LoRA Routing
di: Fleshman, William, et al.
Pubblicazione: (2025)
di: Fleshman, William, et al.
Pubblicazione: (2025)
LoRA-Augmented Generation (LAG) for Knowledge-Intensive Language Tasks
di: Fleshman, William, et al.
Pubblicazione: (2025)
di: Fleshman, William, et al.
Pubblicazione: (2025)
RE-Adapt: Reverse Engineered Adaptation of Large Language Models
di: Fleshman, William, et al.
Pubblicazione: (2024)
di: Fleshman, William, et al.
Pubblicazione: (2024)
Compactor: Calibrated Query-Agnostic KV Cache Compression with Approximate Leverage Scores
di: Chari, Vivek, et al.
Pubblicazione: (2025)
di: Chari, Vivek, et al.
Pubblicazione: (2025)
A Survey of Video Datasets for Grounded Event Understanding
di: Sanders, Kate, et al.
Pubblicazione: (2024)
di: Sanders, Kate, et al.
Pubblicazione: (2024)
SpectR: Dynamically Composing LM Experts with Spectral Routing
di: Fleshman, William, et al.
Pubblicazione: (2025)
di: Fleshman, William, et al.
Pubblicazione: (2025)
Compressed Chain of Thought: Efficient Reasoning Through Dense Representations
di: Cheng, Jeffrey, et al.
Pubblicazione: (2024)
di: Cheng, Jeffrey, et al.
Pubblicazione: (2024)
DOTResize: Reducing LLM Width via Discrete Optimal Transport-based Neuron Merging
di: Verma, Neha, et al.
Pubblicazione: (2025)
di: Verma, Neha, et al.
Pubblicazione: (2025)
The Alignment Waltz: Jointly Training Agents to Collaborate for Safety
di: Zhang, Jingyu, et al.
Pubblicazione: (2025)
di: Zhang, Jingyu, et al.
Pubblicazione: (2025)
It Takes Two: On the Seamlessness between Reward and Policy Model in RLHF
di: Lu, Taiming, et al.
Pubblicazione: (2024)
di: Lu, Taiming, et al.
Pubblicazione: (2024)
CodeRosetta: Pushing the Boundaries of Unsupervised Code Translation for Parallel Programming
di: TehraniJamsaz, Ali, et al.
Pubblicazione: (2024)
di: TehraniJamsaz, Ali, et al.
Pubblicazione: (2024)
LLMs Provide Unstable Answers to Legal Questions
di: Blair-Stanek, Andrew, et al.
Pubblicazione: (2025)
di: Blair-Stanek, Andrew, et al.
Pubblicazione: (2025)
Pushing the Boundary: Specialising Deep Configuration Performance Learning
di: Gong, Jingzhi
Pubblicazione: (2024)
di: Gong, Jingzhi
Pubblicazione: (2024)
Process Supervision of Confidence Margin for Calibrated LLM Reasoning
di: Wang, Liaoyaqi, et al.
Pubblicazione: (2026)
di: Wang, Liaoyaqi, et al.
Pubblicazione: (2026)
Certified Mitigation of Worst-Case LLM Copyright Infringement
di: Zhang, Jingyu, et al.
Pubblicazione: (2025)
di: Zhang, Jingyu, et al.
Pubblicazione: (2025)
Pushing the Boundaries of Interpretability: Incremental Enhancements to the Explainable Boosting Machine
di: Liyanage, Isara, et al.
Pubblicazione: (2025)
di: Liyanage, Isara, et al.
Pubblicazione: (2025)
Error Norm Truncation: Robust Training in the Presence of Data Noise for Text Generation Models
di: Li, Tianjian, et al.
Pubblicazione: (2023)
di: Li, Tianjian, et al.
Pubblicazione: (2023)
CRPO: Confidence-Reward Driven Preference Optimization for Machine Translation
di: Cui, Guofeng, et al.
Pubblicazione: (2025)
di: Cui, Guofeng, et al.
Pubblicazione: (2025)
Backtranslation Augmented Direct Preference Optimization for Neural Machine Translation
di: Ghassabi, Mehrdad, et al.
Pubblicazione: (2026)
di: Ghassabi, Mehrdad, et al.
Pubblicazione: (2026)
Remedy: Learning Machine Translation Evaluation from Human Preferences with Reward Modeling
di: Tan, Shaomu, et al.
Pubblicazione: (2025)
di: Tan, Shaomu, et al.
Pubblicazione: (2025)
Personalized LLM Decoding via Contrasting Personal Preference
di: Bu, Hyungjune, et al.
Pubblicazione: (2025)
di: Bu, Hyungjune, et al.
Pubblicazione: (2025)
PMMT: Preference Alignment in Multilingual Machine Translation via LLM Distillation
di: Sun, Shuqiao, et al.
Pubblicazione: (2024)
di: Sun, Shuqiao, et al.
Pubblicazione: (2024)
SocialNLI: A Dialogue-Centric Social Inference Dataset
di: Deo, Akhil, et al.
Pubblicazione: (2025)
di: Deo, Akhil, et al.
Pubblicazione: (2025)
LM Agents for Coordinating Multi-User Information Gathering
di: Jhamtani, Harsh, et al.
Pubblicazione: (2025)
di: Jhamtani, Harsh, et al.
Pubblicazione: (2025)
A Replicability Study of XTR
di: Jha, Rohan, et al.
Pubblicazione: (2026)
di: Jha, Rohan, et al.
Pubblicazione: (2026)
NevIR: Negation in Neural Information Retrieval
di: Weller, Orion, et al.
Pubblicazione: (2023)
di: Weller, Orion, et al.
Pubblicazione: (2023)
Is That Your Final Answer? Test-Time Scaling Improves Selective Question Answering
di: Jurayj, William, et al.
Pubblicazione: (2025)
di: Jurayj, William, et al.
Pubblicazione: (2025)
Documenti analoghi
-
A Paradigm Shift in Machine Translation: Boosting Translation Performance of Large Language Models
di: Xu, Haoran, et al.
Pubblicazione: (2023) -
Streaming Sequence Transduction through Dynamic Compression
di: Tan, Weiting, et al.
Pubblicazione: (2024) -
Upsample or Upweight? Balanced Training on Heavily Imbalanced Datasets
di: Li, Tianjian, et al.
Pubblicazione: (2024) -
The Language Barrier: Dissecting Safety Challenges of LLMs in Multilingual Contexts
di: Shen, Lingfeng, et al.
Pubblicazione: (2024) -
DiffNorm: Self-Supervised Normalization for Non-autoregressive Speech-to-speech Translation
di: Tan, Weiting, et al.
Pubblicazione: (2024)