Gespeichert in:
| Hauptverfasser: | Wang, Weixuan, Wu, Minghao, Haddow, Barry, Birch, Alexandra |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | https://arxiv.org/abs/2502.12663 |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
HBO: Hierarchical Balancing Optimization for Fine-Tuning Large Language Models
von: Wang, Weixuan, et al.
Veröffentlicht: (2025)
von: Wang, Weixuan, et al.
Veröffentlicht: (2025)
Bridging the Language Gaps in Large Language Models with Inference-Time Cross-Lingual Intervention
von: Wang, Weixuan, et al.
Veröffentlicht: (2024)
von: Wang, Weixuan, et al.
Veröffentlicht: (2024)
ExpertSteer: Intervening in LLMs through Expert Knowledge
von: Wang, Weixuan, et al.
Veröffentlicht: (2025)
von: Wang, Weixuan, et al.
Veröffentlicht: (2025)
Learning to Summarize by Learning to Quiz: Adversarial Agentic Collaboration for Long Document Summarization
von: Wang, Weixuan, et al.
Veröffentlicht: (2025)
von: Wang, Weixuan, et al.
Veröffentlicht: (2025)
Sharing Matters: Analysing Neurons Across Languages and Tasks in LLMs
von: Wang, Weixuan, et al.
Veröffentlicht: (2024)
von: Wang, Weixuan, et al.
Veröffentlicht: (2024)
Multilingual Retrieval-Augmented Generation for Knowledge-Intensive Task
von: Ranaldi, Leonardo, et al.
Veröffentlicht: (2025)
von: Ranaldi, Leonardo, et al.
Veröffentlicht: (2025)
When Does Monolingual Data Help Multilingual Translation: The Role of Domain and Model Scale
von: Baziotis, Christos, et al.
Veröffentlicht: (2023)
von: Baziotis, Christos, et al.
Veröffentlicht: (2023)
Improving Multilingual Retrieval-Augmented Language Models through Dialectic Reasoning Argumentations
von: Ranaldi, Leonardo, et al.
Veröffentlicht: (2025)
von: Ranaldi, Leonardo, et al.
Veröffentlicht: (2025)
The Ups and Downs of Large Language Model Inference with Vocabulary Trimming by Language Heuristics
von: Bogoychev, Nikolay, et al.
Veröffentlicht: (2023)
von: Bogoychev, Nikolay, et al.
Veröffentlicht: (2023)
Compact Speech Translation Models via Discrete Speech Units Pretraining
von: Lam, Tsz Kin, et al.
Veröffentlicht: (2024)
von: Lam, Tsz Kin, et al.
Veröffentlicht: (2024)
MGen: Millions of Naturally Occurring Generics in Context
von: Cilleruelo, Gustavo, et al.
Veröffentlicht: (2025)
von: Cilleruelo, Gustavo, et al.
Veröffentlicht: (2025)
The Prosody of Emojis
von: Zhou, Giulio, et al.
Veröffentlicht: (2025)
von: Zhou, Giulio, et al.
Veröffentlicht: (2025)
Prosody in Cascade and Direct Speech-to-Text Translation: a case study on Korean Wh-Phrases
von: Zhou, Giulio, et al.
Veröffentlicht: (2024)
von: Zhou, Giulio, et al.
Veröffentlicht: (2024)
Generics are puzzling. Can language models find the missing piece?
von: Calderón, Gustavo Cilleruelo, et al.
Veröffentlicht: (2024)
von: Calderón, Gustavo Cilleruelo, et al.
Veröffentlicht: (2024)
Quality or Quantity? On Data Scale and Diversity in Adapting Large Language Models for Low-Resource Translation
von: Iyer, Vivek, et al.
Veröffentlicht: (2024)
von: Iyer, Vivek, et al.
Veröffentlicht: (2024)
Is It Good Data for Multilingual Instruction Tuning or Just Bad Multilingual Evaluation for Large Language Models?
von: Chen, Pinzhen, et al.
Veröffentlicht: (2024)
von: Chen, Pinzhen, et al.
Veröffentlicht: (2024)
Liaozhai through the Looking-Glass: On Paratextual Explicitation of Culture-Bound Terms in Machine Translation
von: Shen, Sherrie, et al.
Veröffentlicht: (2025)
von: Shen, Sherrie, et al.
Veröffentlicht: (2025)
Demystifying Chains, Trees, and Graphs of Thoughts
von: Besta, Maciej, et al.
Veröffentlicht: (2024)
von: Besta, Maciej, et al.
Veröffentlicht: (2024)
Demystifying Long Chain-of-Thought Reasoning in LLMs
von: Yeo, Edward, et al.
Veröffentlicht: (2025)
von: Yeo, Edward, et al.
Veröffentlicht: (2025)
EuroLLM: Multilingual Language Models for Europe
von: Martins, Pedro Henrique, et al.
Veröffentlicht: (2024)
von: Martins, Pedro Henrique, et al.
Veröffentlicht: (2024)
Context and System Fusion in Post-ASR Emotion Recognition with Large Language Models
von: Stepachev, Pavel, et al.
Veröffentlicht: (2024)
von: Stepachev, Pavel, et al.
Veröffentlicht: (2024)
Monolingual or Multilingual Instruction Tuning: Which Makes a Better Alpaca
von: Chen, Pinzhen, et al.
Veröffentlicht: (2023)
von: Chen, Pinzhen, et al.
Veröffentlicht: (2023)
DocHPLT: A Massively Multilingual Document-Level Translation Dataset
von: O'Brien, Dayyán, et al.
Veröffentlicht: (2025)
von: O'Brien, Dayyán, et al.
Veröffentlicht: (2025)
Pushing on Multilingual Reasoning Models with Language-Mixed Chain-of-Thought
von: Son, Guijin, et al.
Veröffentlicht: (2025)
von: Son, Guijin, et al.
Veröffentlicht: (2025)
The Bitter Lesson Learned from 2,000+ Multilingual Benchmarks
von: Wu, Minghao, et al.
Veröffentlicht: (2025)
von: Wu, Minghao, et al.
Veröffentlicht: (2025)
Iterative Translation Refinement with Large Language Models
von: Chen, Pinzhen, et al.
Veröffentlicht: (2023)
von: Chen, Pinzhen, et al.
Veröffentlicht: (2023)
EMMA-500: Enhancing Massively Multilingual Adaptation of Large Language Models
von: Ji, Shaoxiong, et al.
Veröffentlicht: (2024)
von: Ji, Shaoxiong, et al.
Veröffentlicht: (2024)
Teaching Models to Verbalize Reward Hacking in Chain-of-Thought Reasoning
von: Turpin, Miles, et al.
Veröffentlicht: (2025)
von: Turpin, Miles, et al.
Veröffentlicht: (2025)
ETR: Entropy Trend Reward for Efficient Chain-of-Thought Reasoning
von: Xiong, Xuan, et al.
Veröffentlicht: (2026)
von: Xiong, Xuan, et al.
Veröffentlicht: (2026)
Demystifying Instruction Mixing for Fine-tuning Large Language Models
von: Wang, Renxi, et al.
Veröffentlicht: (2023)
von: Wang, Renxi, et al.
Veröffentlicht: (2023)
On the Representational Capacity of Neural Language Models with Chain-of-Thought Reasoning
von: Nowak, Franz, et al.
Veröffentlicht: (2024)
von: Nowak, Franz, et al.
Veröffentlicht: (2024)
Question Translation Training for Better Multilingual Reasoning
von: Zhu, Wenhao, et al.
Veröffentlicht: (2024)
von: Zhu, Wenhao, et al.
Veröffentlicht: (2024)
Fine-Tuning Large Language Models to Translate: Will a Touch of Noisy Data in Misaligned Languages Suffice?
von: Zhu, Dawei, et al.
Veröffentlicht: (2024)
von: Zhu, Dawei, et al.
Veröffentlicht: (2024)
MatheMagic: Generating Dynamic Mathematics Benchmarks Robust to Memorization
von: O'Brien, Dayyán, et al.
Veröffentlicht: (2025)
von: O'Brien, Dayyán, et al.
Veröffentlicht: (2025)
Pitfalls and Outlooks in Using COMET
von: Zouhar, Vilém, et al.
Veröffentlicht: (2024)
von: Zouhar, Vilém, et al.
Veröffentlicht: (2024)
Kakugo: Distillation of Low-Resource Languages into Small Language Models
von: Devine, Peter, et al.
Veröffentlicht: (2026)
von: Devine, Peter, et al.
Veröffentlicht: (2026)
The Molecular Structure of Thought: Mapping the Topology of Long Chain-of-Thought Reasoning
von: Chen, Qiguang, et al.
Veröffentlicht: (2026)
von: Chen, Qiguang, et al.
Veröffentlicht: (2026)
The Power of Question Translation Training in Multilingual Reasoning: Broadened Scope and Deepened Insights
von: Zhu, Wenhao, et al.
Veröffentlicht: (2024)
von: Zhu, Wenhao, et al.
Veröffentlicht: (2024)
Enhancing Chain of Thought Prompting in Large Language Models via Reasoning Patterns
von: Zhang, Yufeng, et al.
Veröffentlicht: (2024)
von: Zhang, Yufeng, et al.
Veröffentlicht: (2024)
Scaling Code-Assisted Chain-of-Thoughts and Instructions for Model Reasoning
von: Lin, Honglin, et al.
Veröffentlicht: (2025)
von: Lin, Honglin, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
HBO: Hierarchical Balancing Optimization for Fine-Tuning Large Language Models
von: Wang, Weixuan, et al.
Veröffentlicht: (2025) -
Bridging the Language Gaps in Large Language Models with Inference-Time Cross-Lingual Intervention
von: Wang, Weixuan, et al.
Veröffentlicht: (2024) -
ExpertSteer: Intervening in LLMs through Expert Knowledge
von: Wang, Weixuan, et al.
Veröffentlicht: (2025) -
Learning to Summarize by Learning to Quiz: Adversarial Agentic Collaboration for Long Document Summarization
von: Wang, Weixuan, et al.
Veröffentlicht: (2025) -
Sharing Matters: Analysing Neurons Across Languages and Tasks in LLMs
von: Wang, Weixuan, et al.
Veröffentlicht: (2024)