Simultaneous Masking, Not Prompting Optimization: A Paradigm Shift in Fine-tuning LLMs for Simultaneous Translation
Fuente:
arXiv
Guardado en:
| Autores principales: | Raffel, Matthew, Agostinelli, Victor, Chen, Lizhong |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
BeaverTalk: Oregon State University's IWSLT 2025 Simultaneous Speech Translation System
por: Raffel, Matthew, et al.
Publicado: (2025)
por: Raffel, Matthew, et al.
Publicado: (2025)
LeaPformer: Enabling Linear Transformers for Autoregressive and Simultaneous Tasks via Learned Proportions
por: Agostinelli, Victor, et al.
Publicado: (2024)
por: Agostinelli, Victor, et al.
Publicado: (2024)
Simul-LLM: A Framework for Exploring High-Quality Simultaneous Translation with Large Language Models
por: Agostinelli, Victor, et al.
Publicado: (2023)
por: Agostinelli, Victor, et al.
Publicado: (2023)
FlashKAT: Understanding and Addressing Performance Bottlenecks in the Kolmogorov-Arnold Transformer
por: Raffel, Matthew, et al.
Publicado: (2025)
por: Raffel, Matthew, et al.
Publicado: (2025)
Memento: Fine-tuning LLM Agents without Fine-tuning LLMs
por: Zhou, Huichi, et al.
Publicado: (2025)
por: Zhou, Huichi, et al.
Publicado: (2025)
MetaCluster: Enabling Deep Compression of Kolmogorov-Arnold Network
por: Raffel, Matthew, et al.
Publicado: (2025)
por: Raffel, Matthew, et al.
Publicado: (2025)
Simultaneous Computation and Memory Efficient Zeroth-Order Optimizer for Fine-Tuning Large Language Models
por: Wang, Fei, et al.
Publicado: (2024)
por: Wang, Fei, et al.
Publicado: (2024)
SeqPO-SiMT: Sequential Policy Optimization for Simultaneous Machine Translation
por: Xu, Ting, et al.
Publicado: (2025)
por: Xu, Ting, et al.
Publicado: (2025)
FineInstructions: Scaling Synthetic Instructions to Pre-Training Scale
por: Patel, Ajay, et al.
Publicado: (2026)
por: Patel, Ajay, et al.
Publicado: (2026)
Keeping LLMs Aligned After Fine-tuning: The Crucial Role of Prompt Templates
por: Lyu, Kaifeng, et al.
Publicado: (2024)
por: Lyu, Kaifeng, et al.
Publicado: (2024)
Topic Modeling with Fine-tuning LLMs and Bag of Sentences
por: Schneider, Johannes
Publicado: (2024)
por: Schneider, Johannes
Publicado: (2024)
Fine-tuning Large Language Models for Domain-specific Machine Translation
por: Zheng, Jiawei, et al.
Publicado: (2024)
por: Zheng, Jiawei, et al.
Publicado: (2024)
BitFit: Simple Parameter-efficient Fine-tuning for Transformer-based Masked Language-models
por: Ben-Zaken, Elad, et al.
Publicado: (2021)
por: Ben-Zaken, Elad, et al.
Publicado: (2021)
Random Masking Finds Winning Tickets for Parameter Efficient Fine-tuning
por: Xu, Jing, et al.
Publicado: (2024)
por: Xu, Jing, et al.
Publicado: (2024)
SimulTron: On-Device Simultaneous Speech to Speech Translation
por: Agranovich, Alex, et al.
Publicado: (2024)
por: Agranovich, Alex, et al.
Publicado: (2024)
Robust and Efficient Fine-tuning of LLMs with Bayesian Reparameterization of Low-Rank Adaptation
por: Sengupta, Ayan, et al.
Publicado: (2024)
por: Sengupta, Ayan, et al.
Publicado: (2024)
Simplify-This: A Comparative Analysis of Prompt-Based and Fine-Tuned LLMs
por: Cohen, Eilam, et al.
Publicado: (2026)
por: Cohen, Eilam, et al.
Publicado: (2026)
REINA: Regularized Entropy Information-Based Loss for Efficient Simultaneous Speech Translation
por: Hirschkind, Nameer, et al.
Publicado: (2025)
por: Hirschkind, Nameer, et al.
Publicado: (2025)
Fine-tuning and Prompt Engineering with Cognitive Knowledge Graphs for Scholarly Knowledge Organization
por: Rabby, Gollam, et al.
Publicado: (2024)
por: Rabby, Gollam, et al.
Publicado: (2024)
Merging by Matching Models in Task Parameter Subspaces
por: Tam, Derek, et al.
Publicado: (2023)
por: Tam, Derek, et al.
Publicado: (2023)
DataDreamer: A Tool for Synthetic Data Generation and Reproducible LLM Workflows
por: Patel, Ajay, et al.
Publicado: (2024)
por: Patel, Ajay, et al.
Publicado: (2024)
RLSF: Fine-tuning LLMs via Symbolic Feedback
por: Jha, Piyush, et al.
Publicado: (2024)
por: Jha, Piyush, et al.
Publicado: (2024)
Multitask Mayhem: Unveiling and Mitigating Safety Gaps in LLMs Fine-tuning
por: Jan, Essa, et al.
Publicado: (2024)
por: Jan, Essa, et al.
Publicado: (2024)
AlignAtt: Using Attention-based Audio-Translation Alignments as a Guide for Simultaneous Speech Translation
por: Papi, Sara, et al.
Publicado: (2023)
por: Papi, Sara, et al.
Publicado: (2023)
PromptIntern: Saving Inference Costs by Internalizing Recurrent Prompt during Large Language Model Fine-tuning
por: Zou, Jiaru, et al.
Publicado: (2024)
por: Zou, Jiaru, et al.
Publicado: (2024)
Prompting and Fine-tuning Large Language Models for Automated Code Review Comment Generation
por: Haider, Md. Asif, et al.
Publicado: (2024)
por: Haider, Md. Asif, et al.
Publicado: (2024)
Zero- and Few-Shot Prompting with LLMs: A Comparative Study with Fine-tuned Models for Bangla Sentiment Analysis
por: Hasan, Md. Arid, et al.
Publicado: (2023)
por: Hasan, Md. Arid, et al.
Publicado: (2023)
Obliviate: Neutralizing Task-agnostic Backdoors within the Parameter-efficient Fine-tuning Paradigm
por: Kim, Jaehan, et al.
Publicado: (2024)
por: Kim, Jaehan, et al.
Publicado: (2024)
Why LLM Safety Guardrails Collapse After Fine-tuning: A Similarity Analysis Between Alignment and Fine-tuning Datasets
por: Hsiung, Lei, et al.
Publicado: (2025)
por: Hsiung, Lei, et al.
Publicado: (2025)
Position: The Most Expensive Part of an LLM should be its Training Data
por: Kandpal, Nikhil, et al.
Publicado: (2025)
por: Kandpal, Nikhil, et al.
Publicado: (2025)
Principled Fine-tuning of LLMs from User-Edits: A Medley of Preference, Supervision, and Reward
por: Misra, Dipendra, et al.
Publicado: (2026)
por: Misra, Dipendra, et al.
Publicado: (2026)
KcMF: A Knowledge-compliant Framework for Schema and Entity Matching with Fine-tuning-free LLMs
por: Xu, Yongqin, et al.
Publicado: (2024)
por: Xu, Yongqin, et al.
Publicado: (2024)
Fine-tuning Quantized Neural Networks with Zeroth-order Optimization
por: Shang, Sifeng, et al.
Publicado: (2025)
por: Shang, Sifeng, et al.
Publicado: (2025)
Private Fine-tuning of Large Language Models with Zeroth-order Optimization
por: Tang, Xinyu, et al.
Publicado: (2024)
por: Tang, Xinyu, et al.
Publicado: (2024)
Improving Rare Word Translation With Dictionaries and Attention Masking
por: Sible, Kenneth J., et al.
Publicado: (2024)
por: Sible, Kenneth J., et al.
Publicado: (2024)
On the Loss of Context-awareness in General Instruction Fine-tuning
por: Wang, Yihan, et al.
Publicado: (2024)
por: Wang, Yihan, et al.
Publicado: (2024)
SEE: Continual Fine-tuning with Sequential Ensemble of Experts
por: Wang, Zhilin, et al.
Publicado: (2025)
por: Wang, Zhilin, et al.
Publicado: (2025)
DePT: Decomposed Prompt Tuning for Parameter-Efficient Fine-tuning
por: Shi, Zhengxiang, et al.
Publicado: (2023)
por: Shi, Zhengxiang, et al.
Publicado: (2023)
What Makes and Breaks Safety Fine-tuning? A Mechanistic Study
por: Jain, Samyak, et al.
Publicado: (2024)
por: Jain, Samyak, et al.
Publicado: (2024)
A General Framework to Enhance Fine-tuning-based LLM Unlearning
por: Ren, Jie, et al.
Publicado: (2025)
por: Ren, Jie, et al.
Publicado: (2025)
Ejemplares similares
-
BeaverTalk: Oregon State University's IWSLT 2025 Simultaneous Speech Translation System
por: Raffel, Matthew, et al.
Publicado: (2025) -
LeaPformer: Enabling Linear Transformers for Autoregressive and Simultaneous Tasks via Learned Proportions
por: Agostinelli, Victor, et al.
Publicado: (2024) -
Simul-LLM: A Framework for Exploring High-Quality Simultaneous Translation with Large Language Models
por: Agostinelli, Victor, et al.
Publicado: (2023) -
FlashKAT: Understanding and Addressing Performance Bottlenecks in the Kolmogorov-Arnold Transformer
por: Raffel, Matthew, et al.
Publicado: (2025) -
Memento: Fine-tuning LLM Agents without Fine-tuning LLMs
por: Zhou, Huichi, et al.
Publicado: (2025)