Steering Large Language Models for Machine Translation Personalization
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Scalena, Daniel, Sarti, Gabriele, Bisazza, Arianna, Fersini, Elisabetta, Nissim, Malvina |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Multi-property Steering of Large Language Models with Dynamic Activation Composition
von: Scalena, Daniel, et al.
Veröffentlicht: (2024)
von: Scalena, Daniel, et al.
Veröffentlicht: (2024)
Quantifying the Plausibility of Context Reliance in Neural Machine Translation
von: Sarti, Gabriele, et al.
Veröffentlicht: (2023)
von: Sarti, Gabriele, et al.
Veröffentlicht: (2023)
A gentle push funziona benissimo: making instructed models in Italian via contrastive activation steering
von: Scalena, Daniel, et al.
Veröffentlicht: (2024)
von: Scalena, Daniel, et al.
Veröffentlicht: (2024)
Non Verbis, Sed Rebus: Large Language Models are Weak Solvers of Italian Rebuses
von: Sarti, Gabriele, et al.
Veröffentlicht: (2024)
von: Sarti, Gabriele, et al.
Veröffentlicht: (2024)
Unsupervised Word-level Quality Estimation for Machine Translation Through the Lens of Annotators (Dis)agreement
von: Sarti, Gabriele, et al.
Veröffentlicht: (2025)
von: Sarti, Gabriele, et al.
Veröffentlicht: (2025)
Model Internals-based Answer Attribution for Trustworthy Retrieval-Augmented Generation
von: Qi, Jirui, et al.
Veröffentlicht: (2024)
von: Qi, Jirui, et al.
Veröffentlicht: (2024)
Cross-Lingual Consistency of Factual Knowledge in Multilingual Language Models
von: Qi, Jirui, et al.
Veröffentlicht: (2023)
von: Qi, Jirui, et al.
Veröffentlicht: (2023)
IT5: Text-to-text Pretraining for Italian Language Understanding and Generation
von: Sarti, Gabriele, et al.
Veröffentlicht: (2022)
von: Sarti, Gabriele, et al.
Veröffentlicht: (2022)
Compositional Steering of Large Language Models with Steering Tokens
von: Radevski, Gorjan, et al.
Veröffentlicht: (2026)
von: Radevski, Gorjan, et al.
Veröffentlicht: (2026)
QE4PE: Word-level Quality Estimation for Human Post-Editing
von: Sarti, Gabriele, et al.
Veröffentlicht: (2025)
von: Sarti, Gabriele, et al.
Veröffentlicht: (2025)
Pointwise Mutual Information as a Performance Gauge for Retrieval-Augmented Generation
von: Liu, Tianyu, et al.
Veröffentlicht: (2024)
von: Liu, Tianyu, et al.
Veröffentlicht: (2024)
TACLer: Tailored Curriculum Reinforcement Learning for Efficient Reasoning
von: Lai, Huiyuan, et al.
Veröffentlicht: (2026)
von: Lai, Huiyuan, et al.
Veröffentlicht: (2026)
COLD-Steer: Steering Large Language Models via In-Context One-step Learning Dynamics
von: Sharma, Kartik, et al.
Veröffentlicht: (2026)
von: Sharma, Kartik, et al.
Veröffentlicht: (2026)
FineSteer: A Unified Framework for Fine-Grained Inference-Time Steering in Large Language Models
von: Weng, Zixuan, et al.
Veröffentlicht: (2026)
von: Weng, Zixuan, et al.
Veröffentlicht: (2026)
Can General-Purpose Large Language Models Generalize to English-Thai Machine Translation ?
von: Chiaranaipanich, Jirat, et al.
Veröffentlicht: (2024)
von: Chiaranaipanich, Jirat, et al.
Veröffentlicht: (2024)
Interpretable Steering of Large Language Models with Feature Guided Activation Additions
von: Soo, Samuel, et al.
Veröffentlicht: (2025)
von: Soo, Samuel, et al.
Veröffentlicht: (2025)
ContextFocus: Activation Steering for Contextual Faithfulness in Large Language Models
von: Anand, Nikhil, et al.
Veröffentlicht: (2026)
von: Anand, Nikhil, et al.
Veröffentlicht: (2026)
Do Personality Traits Interfere? Geometric Limitations of Steering in Large Language Models
von: Bhandari, Pranav, et al.
Veröffentlicht: (2026)
von: Bhandari, Pranav, et al.
Veröffentlicht: (2026)
Integrating Pre-trained Language Model into Neural Machine Translation
von: Hwang, Soon-Jae, et al.
Veröffentlicht: (2023)
von: Hwang, Soon-Jae, et al.
Veröffentlicht: (2023)
Word Embeddings Are Steers for Language Models
von: Han, Chi, et al.
Veröffentlicht: (2023)
von: Han, Chi, et al.
Veröffentlicht: (2023)
Endogenous Resistance to Activation Steering in Language Models
von: McKenzie, Alex, et al.
Veröffentlicht: (2026)
von: McKenzie, Alex, et al.
Veröffentlicht: (2026)
Personalized Wireless Federated Learning for Large Language Models
von: Jiang, Feibo, et al.
Veröffentlicht: (2024)
von: Jiang, Feibo, et al.
Veröffentlicht: (2024)
Practising responsibility: Ethics in NLP as a hands-on course
von: Nissim, Malvina, et al.
Veröffentlicht: (2025)
von: Nissim, Malvina, et al.
Veröffentlicht: (2025)
SAEMark: Steering Personalized Multilingual LLM Watermarks with Sparse Autoencoders
von: Yu, Zhuohao, et al.
Veröffentlicht: (2025)
von: Yu, Zhuohao, et al.
Veröffentlicht: (2025)
Mitigating Overthinking in Large Reasoning Models via Manifold Steering
von: Huang, Yao, et al.
Veröffentlicht: (2025)
von: Huang, Yao, et al.
Veröffentlicht: (2025)
Large Language Models Streamline Automated Machine Learning for Clinical Studies
von: Arasteh, Soroosh Tayebi, et al.
Veröffentlicht: (2023)
von: Arasteh, Soroosh Tayebi, et al.
Veröffentlicht: (2023)
Using Large Language Models to Automate and Expedite Reinforcement Learning with Reward Machine
von: Alsadat, Shayan Meshkat, et al.
Veröffentlicht: (2024)
von: Alsadat, Shayan Meshkat, et al.
Veröffentlicht: (2024)
Are Character-level Translations Worth the Wait? Comparing ByT5 and mT5 for Machine Translation
von: Edman, Lukas, et al.
Veröffentlicht: (2023)
von: Edman, Lukas, et al.
Veröffentlicht: (2023)
Multi-Attribute Steering of Language Models via Targeted Intervention
von: Nguyen, Duy, et al.
Veröffentlicht: (2025)
von: Nguyen, Duy, et al.
Veröffentlicht: (2025)
Improving Instruction-Following in Language Models through Activation Steering
von: Stolfo, Alessandro, et al.
Veröffentlicht: (2024)
von: Stolfo, Alessandro, et al.
Veröffentlicht: (2024)
Linear Representation Transferability Hypothesis: Leveraging Small Models to Steer Large Models
von: Bello, Femi, et al.
Veröffentlicht: (2025)
von: Bello, Femi, et al.
Veröffentlicht: (2025)
GenTranslate: Large Language Models are Generative Multilingual Speech and Machine Translators
von: Hu, Yuchen, et al.
Veröffentlicht: (2024)
von: Hu, Yuchen, et al.
Veröffentlicht: (2024)
Do Large Language Models Know How Much They Know?
von: Prato, Gabriele, et al.
Veröffentlicht: (2025)
von: Prato, Gabriele, et al.
Veröffentlicht: (2025)
Mitigating Hallucinated Translations in Large Language Models with Hallucination-focused Preference Optimization
von: Tang, Zilu, et al.
Veröffentlicht: (2025)
von: Tang, Zilu, et al.
Veröffentlicht: (2025)
A Behavioural and Representational Evaluation of Goal-Directedness in Language Model Agents
von: Arghal, Raghu, et al.
Veröffentlicht: (2026)
von: Arghal, Raghu, et al.
Veröffentlicht: (2026)
Automating Steering for Safe Multimodal Large Language Models
von: Wu, Lyucheng, et al.
Veröffentlicht: (2025)
von: Wu, Lyucheng, et al.
Veröffentlicht: (2025)
Effect of Document Packing on the Latent Multi-Hop Reasoning Capabilities of Large Language Models
von: Prato, Gabriele, et al.
Veröffentlicht: (2025)
von: Prato, Gabriele, et al.
Veröffentlicht: (2025)
Beyond Linear Steering: Unified Multi-Attribute Control for Language Models
von: Oozeer, Narmeen, et al.
Veröffentlicht: (2025)
von: Oozeer, Narmeen, et al.
Veröffentlicht: (2025)
ILRR: Inference-Time Steering Method for Masked Diffusion Language Models
von: Avrahami, Eden, et al.
Veröffentlicht: (2026)
von: Avrahami, Eden, et al.
Veröffentlicht: (2026)
On Translating Technical Terminology: A Translation Workflow for Machine-Translated Acronyms
von: Yue, Richard, et al.
Veröffentlicht: (2024)
von: Yue, Richard, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Multi-property Steering of Large Language Models with Dynamic Activation Composition
von: Scalena, Daniel, et al.
Veröffentlicht: (2024) -
Quantifying the Plausibility of Context Reliance in Neural Machine Translation
von: Sarti, Gabriele, et al.
Veröffentlicht: (2023) -
A gentle push funziona benissimo: making instructed models in Italian via contrastive activation steering
von: Scalena, Daniel, et al.
Veröffentlicht: (2024) -
Non Verbis, Sed Rebus: Large Language Models are Weak Solvers of Italian Rebuses
von: Sarti, Gabriele, et al.
Veröffentlicht: (2024) -
Unsupervised Word-level Quality Estimation for Machine Translation Through the Lens of Annotators (Dis)agreement
von: Sarti, Gabriele, et al.
Veröffentlicht: (2025)