Salvato in:
| Autori principali: | Wang, Weixuan, Haddow, Barry, Wu, Minghao, Peng, Wei, Birch, Alexandra |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | https://arxiv.org/abs/2406.09265 |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
ExpertSteer: Intervening in LLMs through Expert Knowledge
di: Wang, Weixuan, et al.
Pubblicazione: (2025)
di: Wang, Weixuan, et al.
Pubblicazione: (2025)
Bridging the Language Gaps in Large Language Models with Inference-Time Cross-Lingual Intervention
di: Wang, Weixuan, et al.
Pubblicazione: (2024)
di: Wang, Weixuan, et al.
Pubblicazione: (2024)
HBO: Hierarchical Balancing Optimization for Fine-Tuning Large Language Models
di: Wang, Weixuan, et al.
Pubblicazione: (2025)
di: Wang, Weixuan, et al.
Pubblicazione: (2025)
Demystifying Multilingual Chain-of-Thought in Process Reward Modeling
di: Wang, Weixuan, et al.
Pubblicazione: (2025)
di: Wang, Weixuan, et al.
Pubblicazione: (2025)
Learning to Summarize by Learning to Quiz: Adversarial Agentic Collaboration for Long Document Summarization
di: Wang, Weixuan, et al.
Pubblicazione: (2025)
di: Wang, Weixuan, et al.
Pubblicazione: (2025)
Multilingual Retrieval-Augmented Generation for Knowledge-Intensive Task
di: Ranaldi, Leonardo, et al.
Pubblicazione: (2025)
di: Ranaldi, Leonardo, et al.
Pubblicazione: (2025)
The Ups and Downs of Large Language Model Inference with Vocabulary Trimming by Language Heuristics
di: Bogoychev, Nikolay, et al.
Pubblicazione: (2023)
di: Bogoychev, Nikolay, et al.
Pubblicazione: (2023)
When Does Monolingual Data Help Multilingual Translation: The Role of Domain and Model Scale
di: Baziotis, Christos, et al.
Pubblicazione: (2023)
di: Baziotis, Christos, et al.
Pubblicazione: (2023)
MGen: Millions of Naturally Occurring Generics in Context
di: Cilleruelo, Gustavo, et al.
Pubblicazione: (2025)
di: Cilleruelo, Gustavo, et al.
Pubblicazione: (2025)
Compact Speech Translation Models via Discrete Speech Units Pretraining
di: Lam, Tsz Kin, et al.
Pubblicazione: (2024)
di: Lam, Tsz Kin, et al.
Pubblicazione: (2024)
Prosody in Cascade and Direct Speech-to-Text Translation: a case study on Korean Wh-Phrases
di: Zhou, Giulio, et al.
Pubblicazione: (2024)
di: Zhou, Giulio, et al.
Pubblicazione: (2024)
The Prosody of Emojis
di: Zhou, Giulio, et al.
Pubblicazione: (2025)
di: Zhou, Giulio, et al.
Pubblicazione: (2025)
Improving Multilingual Retrieval-Augmented Language Models through Dialectic Reasoning Argumentations
di: Ranaldi, Leonardo, et al.
Pubblicazione: (2025)
di: Ranaldi, Leonardo, et al.
Pubblicazione: (2025)
Generics are puzzling. Can language models find the missing piece?
di: Calderón, Gustavo Cilleruelo, et al.
Pubblicazione: (2024)
di: Calderón, Gustavo Cilleruelo, et al.
Pubblicazione: (2024)
Quality or Quantity? On Data Scale and Diversity in Adapting Large Language Models for Low-Resource Translation
di: Iyer, Vivek, et al.
Pubblicazione: (2024)
di: Iyer, Vivek, et al.
Pubblicazione: (2024)
Liaozhai through the Looking-Glass: On Paratextual Explicitation of Culture-Bound Terms in Machine Translation
di: Shen, Sherrie, et al.
Pubblicazione: (2025)
di: Shen, Sherrie, et al.
Pubblicazione: (2025)
In-game Toxic Language Detection: Shared Task and Attention Residuals
di: Jia, Yuanzhe, et al.
Pubblicazione: (2022)
di: Jia, Yuanzhe, et al.
Pubblicazione: (2022)
Context and System Fusion in Post-ASR Emotion Recognition with Large Language Models
di: Stepachev, Pavel, et al.
Pubblicazione: (2024)
di: Stepachev, Pavel, et al.
Pubblicazione: (2024)
Semantics-Adaptive Activation Intervention for LLMs via Dynamic Steering Vectors
di: Wang, Weixuan, et al.
Pubblicazione: (2024)
di: Wang, Weixuan, et al.
Pubblicazione: (2024)
Iterative Translation Refinement with Large Language Models
di: Chen, Pinzhen, et al.
Pubblicazione: (2023)
di: Chen, Pinzhen, et al.
Pubblicazione: (2023)
Controlling What You Share: Assessing Language Model Adherence to Privacy Preferences
di: Ramírez, Guillem, et al.
Pubblicazione: (2025)
di: Ramírez, Guillem, et al.
Pubblicazione: (2025)
Is It Good Data for Multilingual Instruction Tuning or Just Bad Multilingual Evaluation for Large Language Models?
di: Chen, Pinzhen, et al.
Pubblicazione: (2024)
di: Chen, Pinzhen, et al.
Pubblicazione: (2024)
Fine-Tuning Large Language Models to Translate: Will a Touch of Noisy Data in Misaligned Languages Suffice?
di: Zhu, Dawei, et al.
Pubblicazione: (2024)
di: Zhu, Dawei, et al.
Pubblicazione: (2024)
Kakugo: Distillation of Low-Resource Languages into Small Language Models
di: Devine, Peter, et al.
Pubblicazione: (2026)
di: Devine, Peter, et al.
Pubblicazione: (2026)
No Train but Gain: Language Arithmetic for training-free Language Adapters enhancement
di: Klimaszewski, Mateusz, et al.
Pubblicazione: (2024)
di: Klimaszewski, Mateusz, et al.
Pubblicazione: (2024)
MatheMagic: Generating Dynamic Mathematics Benchmarks Robust to Memorization
di: O'Brien, Dayyán, et al.
Pubblicazione: (2025)
di: O'Brien, Dayyán, et al.
Pubblicazione: (2025)
Pitfalls and Outlooks in Using COMET
di: Zouhar, Vilém, et al.
Pubblicazione: (2024)
di: Zouhar, Vilém, et al.
Pubblicazione: (2024)
Optimising Calls to Large Language Models with Uncertainty-Based Two-Tier Selection
di: Ramírez, Guillem, et al.
Pubblicazione: (2024)
di: Ramírez, Guillem, et al.
Pubblicazione: (2024)
Catastrophic Forgetting in LLMs: A Comparative Analysis Across Language Tasks
di: Haque, Naimul
Pubblicazione: (2025)
di: Haque, Naimul
Pubblicazione: (2025)
Do LLMs and VLMs Share Neurons for Inference? Evidence and Mechanisms of Cross-Modal Transfer
di: Cui, Chenhang, et al.
Pubblicazione: (2026)
di: Cui, Chenhang, et al.
Pubblicazione: (2026)
How Programming Concepts and Neurons Are Shared in Code Language Models
di: Kargaran, Amir Hossein, et al.
Pubblicazione: (2025)
di: Kargaran, Amir Hossein, et al.
Pubblicazione: (2025)
EuroLLM: Multilingual Language Models for Europe
di: Martins, Pedro Henrique, et al.
Pubblicazione: (2024)
di: Martins, Pedro Henrique, et al.
Pubblicazione: (2024)
LF-Steering: Latent Feature Activation Steering for Enhancing Semantic Consistency in Large Language Models
di: Yang, Jingyuan, et al.
Pubblicazione: (2025)
di: Yang, Jingyuan, et al.
Pubblicazione: (2025)
Findings of the WMT 2024 Shared Task on Discourse-Level Literary Translation
di: Wang, Longyue, et al.
Pubblicazione: (2024)
di: Wang, Longyue, et al.
Pubblicazione: (2024)
Evaluating the IWSLT2023 Speech Translation Tasks: Human Annotations, Automatic Metrics, and Segmentation
di: Sperber, Matthias, et al.
Pubblicazione: (2024)
di: Sperber, Matthias, et al.
Pubblicazione: (2024)
Monolingual or Multilingual Instruction Tuning: Which Makes a Better Alpaca
di: Chen, Pinzhen, et al.
Pubblicazione: (2023)
di: Chen, Pinzhen, et al.
Pubblicazione: (2023)
Identifying Good and Bad Neurons for Task-Level Controllable LLMs
di: Li, Wenjie, et al.
Pubblicazione: (2026)
di: Li, Wenjie, et al.
Pubblicazione: (2026)
The Semantic Hub Hypothesis: Language Models Share Semantic Representations Across Languages and Modalities
di: Wu, Zhaofeng, et al.
Pubblicazione: (2024)
di: Wu, Zhaofeng, et al.
Pubblicazione: (2024)
Prepending or Cross-Attention for Speech-to-Text? An Empirical Comparison
di: Lam, Tsz Kin, et al.
Pubblicazione: (2025)
di: Lam, Tsz Kin, et al.
Pubblicazione: (2025)
DocHPLT: A Massively Multilingual Document-Level Translation Dataset
di: O'Brien, Dayyán, et al.
Pubblicazione: (2025)
di: O'Brien, Dayyán, et al.
Pubblicazione: (2025)
Documenti analoghi
-
ExpertSteer: Intervening in LLMs through Expert Knowledge
di: Wang, Weixuan, et al.
Pubblicazione: (2025) -
Bridging the Language Gaps in Large Language Models with Inference-Time Cross-Lingual Intervention
di: Wang, Weixuan, et al.
Pubblicazione: (2024) -
HBO: Hierarchical Balancing Optimization for Fine-Tuning Large Language Models
di: Wang, Weixuan, et al.
Pubblicazione: (2025) -
Demystifying Multilingual Chain-of-Thought in Process Reward Modeling
di: Wang, Weixuan, et al.
Pubblicazione: (2025) -
Learning to Summarize by Learning to Quiz: Adversarial Agentic Collaboration for Long Document Summarization
di: Wang, Weixuan, et al.
Pubblicazione: (2025)