IT5: Text-to-text Pretraining for Italian Language Understanding and Generation
Fuente:
arXiv
Salvato in:
| Autori principali: | Sarti, Gabriele, Nissim, Malvina |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2022
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Non Verbis, Sed Rebus: Large Language Models are Weak Solvers of Italian Rebuses
di: Sarti, Gabriele, et al.
Pubblicazione: (2024)
di: Sarti, Gabriele, et al.
Pubblicazione: (2024)
Multi-property Steering of Large Language Models with Dynamic Activation Composition
di: Scalena, Daniel, et al.
Pubblicazione: (2024)
di: Scalena, Daniel, et al.
Pubblicazione: (2024)
Unsupervised Word-level Quality Estimation for Machine Translation Through the Lens of Annotators (Dis)agreement
di: Sarti, Gabriele, et al.
Pubblicazione: (2025)
di: Sarti, Gabriele, et al.
Pubblicazione: (2025)
Steering Large Language Models for Machine Translation Personalization
di: Scalena, Daniel, et al.
Pubblicazione: (2025)
di: Scalena, Daniel, et al.
Pubblicazione: (2025)
Quantifying the Plausibility of Context Reliance in Neural Machine Translation
di: Sarti, Gabriele, et al.
Pubblicazione: (2023)
di: Sarti, Gabriele, et al.
Pubblicazione: (2023)
mCoT: Multilingual Instruction Tuning for Reasoning Consistency in Language Models
di: Lai, Huiyuan, et al.
Pubblicazione: (2024)
di: Lai, Huiyuan, et al.
Pubblicazione: (2024)
A gentle push funziona benissimo: making instructed models in Italian via contrastive activation steering
di: Scalena, Daniel, et al.
Pubblicazione: (2024)
di: Scalena, Daniel, et al.
Pubblicazione: (2024)
QE4PE: Word-level Quality Estimation for Human Post-Editing
di: Sarti, Gabriele, et al.
Pubblicazione: (2025)
di: Sarti, Gabriele, et al.
Pubblicazione: (2025)
Puzzled By ChatGPT? No more! A Jigsaw Puzzle to Promote AI Literacy and Awareness
di: Padovani, Francesca, et al.
Pubblicazione: (2026)
di: Padovani, Francesca, et al.
Pubblicazione: (2026)
Multidimensional Consistency Improves Reasoning in Language Models
di: Lai, Huiyuan, et al.
Pubblicazione: (2025)
di: Lai, Huiyuan, et al.
Pubblicazione: (2025)
When Harry Meets Superman: The Role of The Interlocutor in Persona-Based Dialogue Generation
di: Occhipinti, Daniela, et al.
Pubblicazione: (2025)
di: Occhipinti, Daniela, et al.
Pubblicazione: (2025)
TACLer: Tailored Curriculum Reinforcement Learning for Efficient Reasoning
di: Lai, Huiyuan, et al.
Pubblicazione: (2026)
di: Lai, Huiyuan, et al.
Pubblicazione: (2026)
Choosy Babies Need One Coach: Inducing Mode-Seeking Behavior in BabyLlama with Reverse KL Divergence
di: Shi, Shaozhen, et al.
Pubblicazione: (2024)
di: Shi, Shaozhen, et al.
Pubblicazione: (2024)
Practising responsibility: Ethics in NLP as a hands-on course
di: Nissim, Malvina, et al.
Pubblicazione: (2025)
di: Nissim, Malvina, et al.
Pubblicazione: (2025)
The Role of the Availability Heuristic in Multiple-Choice Answering Behaviour
di: Zotos, Leonidas, et al.
Pubblicazione: (2026)
di: Zotos, Leonidas, et al.
Pubblicazione: (2026)
Can Model Uncertainty Function as a Proxy for Multiple-Choice Question Item Difficulty?
di: Zotos, Leonidas, et al.
Pubblicazione: (2024)
di: Zotos, Leonidas, et al.
Pubblicazione: (2024)
Are You Doubtful? Oh, It Might Be Difficult Then! Exploring the Use of Model Uncertainty for Question Difficulty Estimation
di: Zotos, Leonidas, et al.
Pubblicazione: (2024)
di: Zotos, Leonidas, et al.
Pubblicazione: (2024)
ARGUS: Seeing the Influence of Narrative Features on Persuasion in Argumentative Texts
di: Nabhani, Sara, et al.
Pubblicazione: (2026)
di: Nabhani, Sara, et al.
Pubblicazione: (2026)
Reveal-Bangla: A Dataset for Cross-Lingual Multi-Step Reasoning Evaluation
di: Islam, Khondoker Ittehadul, et al.
Pubblicazione: (2025)
di: Islam, Khondoker Ittehadul, et al.
Pubblicazione: (2025)
Fine-tuning with HED-IT: The impact of human post-editing for dialogical language models
di: Occhipinti, Daniela, et al.
Pubblicazione: (2024)
di: Occhipinti, Daniela, et al.
Pubblicazione: (2024)
A Primer on the Inner Workings of Transformer-based Language Models
di: Ferrando, Javier, et al.
Pubblicazione: (2024)
di: Ferrando, Javier, et al.
Pubblicazione: (2024)
Are Character-level Translations Worth the Wait? Comparing ByT5 and mT5 for Machine Translation
di: Edman, Lukas, et al.
Pubblicazione: (2023)
di: Edman, Lukas, et al.
Pubblicazione: (2023)
On The Role of Pretrained Language Models in General-Purpose Text Embeddings: A Survey
di: Zhang, Meishan, et al.
Pubblicazione: (2025)
di: Zhang, Meishan, et al.
Pubblicazione: (2025)
The Heuristic Core: Understanding Subnetwork Generalization in Pretrained Language Models
di: Bhaskar, Adithya, et al.
Pubblicazione: (2024)
di: Bhaskar, Adithya, et al.
Pubblicazione: (2024)
Table-to-Text Generation with Pretrained Diffusion Models
di: Krylov, Aleksei S., et al.
Pubblicazione: (2024)
di: Krylov, Aleksei S., et al.
Pubblicazione: (2024)
E2S2: Encoding-Enhanced Sequence-to-Sequence Pretraining for Language Understanding and Generation
di: Zhong, Qihuang, et al.
Pubblicazione: (2022)
di: Zhong, Qihuang, et al.
Pubblicazione: (2022)
Model-Generated Pretraining Signals Improves Zero-Shot Generalization of Text-to-Text Transformers
di: Gong, Linyuan, et al.
Pubblicazione: (2023)
di: Gong, Linyuan, et al.
Pubblicazione: (2023)
AST-T5: Structure-Aware Pretraining for Code Generation and Understanding
di: Gong, Linyuan, et al.
Pubblicazione: (2024)
di: Gong, Linyuan, et al.
Pubblicazione: (2024)
Model Internals-based Answer Attribution for Trustworthy Retrieval-Augmented Generation
di: Qi, Jirui, et al.
Pubblicazione: (2024)
di: Qi, Jirui, et al.
Pubblicazione: (2024)
Distilling Formal Logic into Neural Spaces: A Kernel Alignment Approach for Signal Temporal Logic
di: Candussio, Sara, et al.
Pubblicazione: (2026)
di: Candussio, Sara, et al.
Pubblicazione: (2026)
Challenging the Abilities of Large Language Models in Italian: a Community Initiative
di: Nissim, Malvina, et al.
Pubblicazione: (2025)
di: Nissim, Malvina, et al.
Pubblicazione: (2025)
Rethinking the Role of Text Complexity in Language Model Pretraining
di: Velasco, Dan John, et al.
Pubblicazione: (2025)
di: Velasco, Dan John, et al.
Pubblicazione: (2025)
NLP Methods May Actually Be Better Than Professors at Estimating Question Difficulty
di: Zotos, Leonidas, et al.
Pubblicazione: (2025)
di: Zotos, Leonidas, et al.
Pubblicazione: (2025)
Evaluating, Understanding, and Improving Constrained Text Generation for Large Language Models
di: Chen, Xiang, et al.
Pubblicazione: (2023)
di: Chen, Xiang, et al.
Pubblicazione: (2023)
Improving Estonian Text Simplification through Pretrained Language Models and Custom Datasets
di: Barbu, Eduard, et al.
Pubblicazione: (2025)
di: Barbu, Eduard, et al.
Pubblicazione: (2025)
Bridging Logic and Learning: Decoding Temporal Logic Embeddings via Transformers
di: Candussio, Sara, et al.
Pubblicazione: (2025)
di: Candussio, Sara, et al.
Pubblicazione: (2025)
Domain Embeddings for Generating Complex Descriptions of Concepts in Italian Language
di: Maisto, Alessandro
Pubblicazione: (2024)
di: Maisto, Alessandro
Pubblicazione: (2024)
NoteContrast: Contrastive Language-Diagnostic Pretraining for Medical Text
di: Kailas, Prajwal, et al.
Pubblicazione: (2024)
di: Kailas, Prajwal, et al.
Pubblicazione: (2024)
DecoderLens: Layerwise Interpretation of Encoder-Decoder Transformers
di: Langedijk, Anna, et al.
Pubblicazione: (2023)
di: Langedijk, Anna, et al.
Pubblicazione: (2023)
HRM-Text: Efficient Pretraining Beyond Scaling
di: Wang, Guan, et al.
Pubblicazione: (2026)
di: Wang, Guan, et al.
Pubblicazione: (2026)
Documenti analoghi
-
Non Verbis, Sed Rebus: Large Language Models are Weak Solvers of Italian Rebuses
di: Sarti, Gabriele, et al.
Pubblicazione: (2024) -
Multi-property Steering of Large Language Models with Dynamic Activation Composition
di: Scalena, Daniel, et al.
Pubblicazione: (2024) -
Unsupervised Word-level Quality Estimation for Machine Translation Through the Lens of Annotators (Dis)agreement
di: Sarti, Gabriele, et al.
Pubblicazione: (2025) -
Steering Large Language Models for Machine Translation Personalization
di: Scalena, Daniel, et al.
Pubblicazione: (2025) -
Quantifying the Plausibility of Context Reliance in Neural Machine Translation
di: Sarti, Gabriele, et al.
Pubblicazione: (2023)