DIETA: A Decoder-only transformer-based model for Italian-English machine TrAnslation
Fuente:
arXiv
Salvato in:
| Autori principali: | Kasela, Pranav, Braga, Marco, Ghiotto, Alessandro, Pilzer, Andrea, Viviani, Marco, Raganato, Alessandro |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2026
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Investigating Task Arithmetic for Zero-Shot Information Retrieval
di: Braga, Marco, et al.
Pubblicazione: (2025)
di: Braga, Marco, et al.
Pubblicazione: (2025)
Synthetic Data Generation with Large Language Models for Personalized Community Question Answering
di: Braga, Marco, et al.
Pubblicazione: (2024)
di: Braga, Marco, et al.
Pubblicazione: (2024)
When Good and Reproducible Results are a Giant with Feet of Clay: The Importance of Software Quality in NLP
di: Papi, Sara, et al.
Pubblicazione: (2023)
di: Papi, Sara, et al.
Pubblicazione: (2023)
Reasoning Capabilities and Invariability of Large Language Models
di: Raganato, Alessandro, et al.
Pubblicazione: (2025)
di: Raganato, Alessandro, et al.
Pubblicazione: (2025)
FAMA: The First Large-Scale Open-Science Speech Foundation Model for English and Italian
di: Papi, Sara, et al.
Pubblicazione: (2025)
di: Papi, Sara, et al.
Pubblicazione: (2025)
Harnessing LLMs for Educational Content-Driven Italian Crossword Generation
di: Zeinalipour, Kamyar, et al.
Pubblicazione: (2024)
di: Zeinalipour, Kamyar, et al.
Pubblicazione: (2024)
WordAlchemy: A transformer-based Reverse Dictionary
di: Madaswar, Kanhaiya, et al.
Pubblicazione: (2022)
di: Madaswar, Kanhaiya, et al.
Pubblicazione: (2022)
StableMask: Refining Causal Masking in Decoder-only Transformer
di: Yin, Qingyu, et al.
Pubblicazione: (2024)
di: Yin, Qingyu, et al.
Pubblicazione: (2024)
Uncovering the Potential Risks in Unlearning: Danger of English-only Unlearning in Multilingual LLMs
di: Hwang, Kyomin, et al.
Pubblicazione: (2025)
di: Hwang, Kyomin, et al.
Pubblicazione: (2025)
TrInk: Ink Generation with Transformer Network
di: Jin, Zezhong, et al.
Pubblicazione: (2025)
di: Jin, Zezhong, et al.
Pubblicazione: (2025)
The Disparate Impacts of Speculative Decoding
di: Sandler, Jameson, et al.
Pubblicazione: (2025)
di: Sandler, Jameson, et al.
Pubblicazione: (2025)
Investigating Large Language Models' Linguistic Abilities for Text Preprocessing
di: Braga, Marco, et al.
Pubblicazione: (2025)
di: Braga, Marco, et al.
Pubblicazione: (2025)
Investigating Mixture of Experts in Dense Retrieval
di: Sokli, Effrosyni, et al.
Pubblicazione: (2024)
di: Sokli, Effrosyni, et al.
Pubblicazione: (2024)
Faithfulness Evaluation for Decoder-only LLM Attributions with Controlled Retained Information
di: Huang, Xin, et al.
Pubblicazione: (2026)
di: Huang, Xin, et al.
Pubblicazione: (2026)
How word semantics and phonology affect handwriting of Alzheimer's patients: a machine learning based analysis
di: Cilia, Nicole Dalia, et al.
Pubblicazione: (2023)
di: Cilia, Nicole Dalia, et al.
Pubblicazione: (2023)
Causal2Vec: Improving Decoder-only LLMs as Embedding Models through a Contextual Token
di: Lin, Ailiang, et al.
Pubblicazione: (2025)
di: Lin, Ailiang, et al.
Pubblicazione: (2025)
Identifying Gender Stereotypes and Biases in Automated Translation from English to Italian using Similarity Networks
di: Mohammadi, Fatemeh, et al.
Pubblicazione: (2025)
di: Mohammadi, Fatemeh, et al.
Pubblicazione: (2025)
Hyperdimensional Probe: Decoding LLM Representations via Vector Symbolic Architectures
di: Bronzini, Marco, et al.
Pubblicazione: (2025)
di: Bronzini, Marco, et al.
Pubblicazione: (2025)
Spanish TrOCR: Leveraging Transfer Learning for Language Adaptation
di: Lauar, Filipe, et al.
Pubblicazione: (2024)
di: Lauar, Filipe, et al.
Pubblicazione: (2024)
Collaborative Storytelling and LLM: A Linguistic Analysis of Automatically-Generated Role-Playing Game Sessions
di: Maisto, Alessandro
Pubblicazione: (2025)
di: Maisto, Alessandro
Pubblicazione: (2025)
Approximately Aligned Decoding
di: Melcer, Daniel, et al.
Pubblicazione: (2024)
di: Melcer, Daniel, et al.
Pubblicazione: (2024)
Prompting Encoder Models for Zero-Shot Classification: A Cross-Domain Study in Italian
di: Auriemma, Serena, et al.
Pubblicazione: (2024)
di: Auriemma, Serena, et al.
Pubblicazione: (2024)
Handling Ontology Gaps in Semantic Parsing
di: Bacciu, Andrea, et al.
Pubblicazione: (2024)
di: Bacciu, Andrea, et al.
Pubblicazione: (2024)
Automated evaluation of LLMs for effective machine translation of Mandarin Chinese to English
di: Zhang, Yue, et al.
Pubblicazione: (2026)
di: Zhang, Yue, et al.
Pubblicazione: (2026)
Decoder-only Conformer with Modality-aware Sparse Mixtures of Experts for ASR
di: Lee, Jaeyoung, et al.
Pubblicazione: (2026)
di: Lee, Jaeyoung, et al.
Pubblicazione: (2026)
Memorization in Attention-only Transformers
di: Dana, Léo, et al.
Pubblicazione: (2024)
di: Dana, Léo, et al.
Pubblicazione: (2024)
Dynamic Depth Decoding: Faster Speculative Decoding for LLMs
di: Brown, Oscar, et al.
Pubblicazione: (2024)
di: Brown, Oscar, et al.
Pubblicazione: (2024)
Disce aut Deficere: Evaluating LLMs Proficiency on the INVALSI Italian Benchmark
di: Mercorio, Fabio, et al.
Pubblicazione: (2024)
di: Mercorio, Fabio, et al.
Pubblicazione: (2024)
How Humans and LLMs Organize Conceptual Knowledge: Exploring Subordinate Categories in Italian
di: Pedrotti, Andrea, et al.
Pubblicazione: (2025)
di: Pedrotti, Andrea, et al.
Pubblicazione: (2025)
Efficient Conformance Checking of Rich Data-Aware Declare Specifications (Extended)
di: Casas-Ramos, Jacobo, et al.
Pubblicazione: (2025)
di: Casas-Ramos, Jacobo, et al.
Pubblicazione: (2025)
A Survey on Spoken Italian Datasets and Corpora
di: Giordano, Marco, et al.
Pubblicazione: (2025)
di: Giordano, Marco, et al.
Pubblicazione: (2025)
How to Blend Concepts in Diffusion Models
di: Olearo, Lorenzo, et al.
Pubblicazione: (2024)
di: Olearo, Lorenzo, et al.
Pubblicazione: (2024)
Evil twins are not that evil: Qualitative insights into machine-generated prompts
di: Rakotonirina, Nathanaël Carraz, et al.
Pubblicazione: (2024)
di: Rakotonirina, Nathanaël Carraz, et al.
Pubblicazione: (2024)
Ukrainian-to-English folktale corpus: Parallel corpus creation and augmentation for machine translation in low-resource languages
di: Burda-Lassen, Olena
Pubblicazione: (2024)
di: Burda-Lassen, Olena
Pubblicazione: (2024)
Enhancing Debunking Effectiveness through LLM-based Personality Adaptation
di: Dell'Oglio, Pietro, et al.
Pubblicazione: (2026)
di: Dell'Oglio, Pietro, et al.
Pubblicazione: (2026)
A decoder-only foundation model for time-series forecasting
di: Das, Abhimanyu, et al.
Pubblicazione: (2023)
di: Das, Abhimanyu, et al.
Pubblicazione: (2023)
Benchmarking EngGPT2-16B-A3B against Comparable Italian and International Open-source LLMs
di: Sassella, Andrea, et al.
Pubblicazione: (2026)
di: Sassella, Andrea, et al.
Pubblicazione: (2026)
Decoder-based Sense Knowledge Distillation
di: Wang, Qitong, et al.
Pubblicazione: (2026)
di: Wang, Qitong, et al.
Pubblicazione: (2026)
Introducing TrGLUE and SentiTurca: A Comprehensive Benchmark for Turkish General Language Understanding and Sentiment Analysis
di: Altinok, Duygu
Pubblicazione: (2025)
di: Altinok, Duygu
Pubblicazione: (2025)
LIMBA: An Open-Source Framework for the Preservation and Valorization of Low-Resource Languages using Generative Models
di: Carta, Salvatore Mario, et al.
Pubblicazione: (2024)
di: Carta, Salvatore Mario, et al.
Pubblicazione: (2024)
Documenti analoghi
-
Investigating Task Arithmetic for Zero-Shot Information Retrieval
di: Braga, Marco, et al.
Pubblicazione: (2025) -
Synthetic Data Generation with Large Language Models for Personalized Community Question Answering
di: Braga, Marco, et al.
Pubblicazione: (2024) -
When Good and Reproducible Results are a Giant with Feet of Clay: The Importance of Software Quality in NLP
di: Papi, Sara, et al.
Pubblicazione: (2023) -
Reasoning Capabilities and Invariability of Large Language Models
di: Raganato, Alessandro, et al.
Pubblicazione: (2025) -
FAMA: The First Large-Scale Open-Science Speech Foundation Model for English and Italian
di: Papi, Sara, et al.
Pubblicazione: (2025)