Linguistic Knowledge Can Enhance Encoder-Decoder Models (If You Let It)
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Miaschi, Alessio, Dell'Orletta, Felice, Venturi, Giulia |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Linguistic Profiling of a Neural Language Model
von: Miaschi, Alessio, et al.
Veröffentlicht: (2020)
von: Miaschi, Alessio, et al.
Veröffentlicht: (2020)
Leveraging Encoder-only Large Language Models for Mobile App Review Feature Extraction
von: Motger, Quim, et al.
Veröffentlicht: (2024)
von: Motger, Quim, et al.
Veröffentlicht: (2024)
Linguistically-driven Selection of Correct Arcs for Dependency Parsing
von: Felice Dell’Orletta
Veröffentlicht: (2013)
von: Felice Dell’Orletta
Veröffentlicht: (2013)
Stress-testing Machine Generated Text Detection: Shifting Language Models Writing Style to Fool Detectors
von: Pedrotti, Andrea, et al.
Veröffentlicht: (2025)
von: Pedrotti, Andrea, et al.
Veröffentlicht: (2025)
Outliers Dimensions that Disrupt Transformers Are Driven by Frequency
von: Puccetti, Giovanni, et al.
Veröffentlicht: (2022)
von: Puccetti, Giovanni, et al.
Veröffentlicht: (2022)
Optimizing LLMs for Italian: Reducing Token Fertility and Enhancing Efficiency Through Vocabulary Adaptation
von: Moroni, Luca, et al.
Veröffentlicht: (2025)
von: Moroni, Luca, et al.
Veröffentlicht: (2025)
Contextualized Counterspeech: Strategies for Adaptation, Personalization, and Evaluation
von: Cima, Lorenzo, et al.
Veröffentlicht: (2024)
von: Cima, Lorenzo, et al.
Veröffentlicht: (2024)
T-FREX: A Transformer-based Feature Extraction Method from Mobile App Reviews
von: Motger, Quim, et al.
Veröffentlicht: (2024)
von: Motger, Quim, et al.
Veröffentlicht: (2024)
AI "News" Content Farms Are Easy to Make and Hard to Detect: A Case Study in Italian
von: Puccetti, Giovanni, et al.
Veröffentlicht: (2024)
von: Puccetti, Giovanni, et al.
Veröffentlicht: (2024)
Fine-tuning with HED-IT: The impact of human post-editing for dialogical language models
von: Occhipinti, Daniela, et al.
Veröffentlicht: (2024)
von: Occhipinti, Daniela, et al.
Veröffentlicht: (2024)
Encoder-Decoder or Decoder-Only? Revisiting Encoder-Decoder Large Language Model
von: Zhang, Biao, et al.
Veröffentlicht: (2025)
von: Zhang, Biao, et al.
Veröffentlicht: (2025)
Charting a Decade of Computational Linguistics in Italy: The CLiC-it Corpus
von: Alzetta, Chiara, et al.
Veröffentlicht: (2025)
von: Alzetta, Chiara, et al.
Veröffentlicht: (2025)
Beyond Decoder-only: Large Language Models Can be Good Encoders for Machine Translation
von: Luo, Yingfeng, et al.
Veröffentlicht: (2025)
von: Luo, Yingfeng, et al.
Veröffentlicht: (2025)
Language on Demand, Knowledge at Core: Composing LLMs with Encoder-Decoder Translation Models for Extensible Multilinguality
von: Bu, Mengyu, et al.
Veröffentlicht: (2026)
von: Bu, Mengyu, et al.
Veröffentlicht: (2026)
Let Me Teach You: Pedagogical Foundations of Feedback for Language Models
von: Borges, Beatriz, et al.
Veröffentlicht: (2023)
von: Borges, Beatriz, et al.
Veröffentlicht: (2023)
Efficient Encoder-Decoder Transformer Decoding for Decomposable Tasks
von: Lu, Bo-Ru, et al.
Veröffentlicht: (2024)
von: Lu, Bo-Ru, et al.
Veröffentlicht: (2024)
DecoderLens: Layerwise Interpretation of Encoder-Decoder Transformers
von: Langedijk, Anna, et al.
Veröffentlicht: (2023)
von: Langedijk, Anna, et al.
Veröffentlicht: (2023)
Modeling Professionalism in Expert Questioning through Linguistic Differentiation
von: D'Agostino, Giulia, et al.
Veröffentlicht: (2025)
von: D'Agostino, Giulia, et al.
Veröffentlicht: (2025)
Small Encoders Can Rival Large Decoders in Detecting Groundedness
von: Abbes, Istabrak, et al.
Veröffentlicht: (2025)
von: Abbes, Istabrak, et al.
Veröffentlicht: (2025)
You Only Cache Once: Decoder-Decoder Architectures for Language Models
von: Sun, Yutao, et al.
Veröffentlicht: (2024)
von: Sun, Yutao, et al.
Veröffentlicht: (2024)
All-in-one: Understanding and Generation in Multimodal Reasoning with the MAIA Benchmark
von: Testa, Davide, et al.
Veröffentlicht: (2025)
von: Testa, Davide, et al.
Veröffentlicht: (2025)
Decoders Laugh as Loud as Encoders
von: Borodach, Eli, et al.
Veröffentlicht: (2025)
von: Borodach, Eli, et al.
Veröffentlicht: (2025)
TURNA: A Turkish Encoder-Decoder Language Model for Enhanced Understanding and Generation
von: Uludoğan, Gökçe, et al.
Veröffentlicht: (2024)
von: Uludoğan, Gökçe, et al.
Veröffentlicht: (2024)
Clinical Reading Comprehension with Encoder-Decoder Models Enhanced by Direct Preference Optimization
von: Nahian, Md Sultan Al, et al.
Veröffentlicht: (2024)
von: Nahian, Md Sultan Al, et al.
Veröffentlicht: (2024)
Efficient Knowledge Feeding to Language Models: A Novel Integrated Encoder-Decoder Architecture
von: Kumar, S Santosh, et al.
Veröffentlicht: (2025)
von: Kumar, S Santosh, et al.
Veröffentlicht: (2025)
Enhanced Hybrid Transducer and Attention Encoder Decoder with Text Data
von: Tang, Yun, et al.
Veröffentlicht: (2025)
von: Tang, Yun, et al.
Veröffentlicht: (2025)
Linguistic Loops and Geometric Invariants as a Way to Pre-Verbal Thought?
von: Corradetti, Daniele, et al.
Veröffentlicht: (2025)
von: Corradetti, Daniele, et al.
Veröffentlicht: (2025)
Training and Inference Efficiency of Encoder-Decoder Speech Models
von: Żelasko, Piotr, et al.
Veröffentlicht: (2025)
von: Żelasko, Piotr, et al.
Veröffentlicht: (2025)
Okay, Let's Do This! Modeling Event Coreference with Generated Rationales and Knowledge Distillation
von: Nath, Abhijnan, et al.
Veröffentlicht: (2024)
von: Nath, Abhijnan, et al.
Veröffentlicht: (2024)
Causal Reasoning Favors Encoders: On The Limits of Decoder-Only Models
von: Roy, Amartya, et al.
Veröffentlicht: (2025)
von: Roy, Amartya, et al.
Veröffentlicht: (2025)
LinguaLens: Towards Interpreting Linguistic Mechanisms of Large Language Models via Sparse Auto-Encoder
von: Jing, Yi, et al.
Veröffentlicht: (2025)
von: Jing, Yi, et al.
Veröffentlicht: (2025)
Encoder vs Decoder: Comparative Analysis of Encoder and Decoder Language Models on Multilingual NLU Tasks
von: Nielsen, Dan Saattrup, et al.
Veröffentlicht: (2024)
von: Nielsen, Dan Saattrup, et al.
Veröffentlicht: (2024)
Linguistically Informed Tokenization Improves ASR for Underresourced Languages
von: Daul, Massimo, et al.
Veröffentlicht: (2025)
von: Daul, Massimo, et al.
Veröffentlicht: (2025)
Let it Calm: Exploratory Annealed Decoding for Verifiable Reinforcement Learning
von: Yang, Chenghao, et al.
Veröffentlicht: (2025)
von: Yang, Chenghao, et al.
Veröffentlicht: (2025)
Machine Translation with Large Language Models: Decoder Only vs. Encoder-Decoder
von: M., Abhinav P., et al.
Veröffentlicht: (2024)
von: M., Abhinav P., et al.
Veröffentlicht: (2024)
Let Guidelines Guide You: A Prescriptive Guideline-Centered Data Annotation Methodology
von: Ruggeri, Federico, et al.
Veröffentlicht: (2024)
von: Ruggeri, Federico, et al.
Veröffentlicht: (2024)
Enhancing Authorship Attribution through Embedding Fusion: A Novel Approach with Masked and Encoder-Decoder Language Models
von: Kaushik, Arjun Ramesh, et al.
Veröffentlicht: (2024)
von: Kaushik, Arjun Ramesh, et al.
Veröffentlicht: (2024)
Rethinking the adaptive relationship between Encoder Layers and Decoder Layers
von: Song, Yubo
Veröffentlicht: (2024)
von: Song, Yubo
Veröffentlicht: (2024)
SpeechGLUE: How Well Can Self-Supervised Speech Models Capture Linguistic Knowledge?
von: Ashihara, Takanori, et al.
Veröffentlicht: (2023)
von: Ashihara, Takanori, et al.
Veröffentlicht: (2023)
Adapting Decoder-Based Language Models for Diverse Encoder Downstream Tasks
von: Suganthan, Paul, et al.
Veröffentlicht: (2025)
von: Suganthan, Paul, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Linguistic Profiling of a Neural Language Model
von: Miaschi, Alessio, et al.
Veröffentlicht: (2020) -
Leveraging Encoder-only Large Language Models for Mobile App Review Feature Extraction
von: Motger, Quim, et al.
Veröffentlicht: (2024) -
Linguistically-driven Selection of Correct Arcs for Dependency Parsing
von: Felice Dell’Orletta
Veröffentlicht: (2013) -
Stress-testing Machine Generated Text Detection: Shifting Language Models Writing Style to Fool Detectors
von: Pedrotti, Andrea, et al.
Veröffentlicht: (2025) -
Outliers Dimensions that Disrupt Transformers Are Driven by Frequency
von: Puccetti, Giovanni, et al.
Veröffentlicht: (2022)