Improving Data and Parameter Efficiency of Neural Language Models Using Representation Analysis
Fuente:
arXiv
Saved in:
| Main Author: | Jukić, Josip |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
From Robustness to Improved Generalization and Calibration in Pre-trained Language Models
by: Jukić, Josip, et al.
Published: (2024)
by: Jukić, Josip, et al.
Published: (2024)
Disentangling Latent Shifts of In-Context Learning with Weak Supervision
by: Jukić, Josip, et al.
Published: (2024)
by: Jukić, Josip, et al.
Published: (2024)
Context Parametrization with Compositional Adapters
by: Jukić, Josip, et al.
Published: (2025)
by: Jukić, Josip, et al.
Published: (2025)
Characterizing Linguistic Shifts in Croatian News via Diachronic Word Embeddings
by: Dukić, David, et al.
Published: (2025)
by: Dukić, David, et al.
Published: (2025)
Out-of-Distribution Detection by Leveraging Between-Layer Transformation Smoothness
by: Jelenić, Fran, et al.
Published: (2023)
by: Jelenić, Fran, et al.
Published: (2023)
Performance Trade-offs of Optimizing Small Language Models for E-Commerce
by: Licardo, Josip Tomo, et al.
Published: (2025)
by: Licardo, Josip Tomo, et al.
Published: (2025)
Improved Representation Steering for Language Models
by: Wu, Zhengxuan, et al.
Published: (2025)
by: Wu, Zhengxuan, et al.
Published: (2025)
Navigating Brain Language Representations: A Comparative Analysis of Neural Language Models and Psychologically Plausible Models
by: Zhang, Yunhao, et al.
Published: (2024)
by: Zhang, Yunhao, et al.
Published: (2024)
Improving the Efficiency of Visually Augmented Language Models
by: Ontalvilla, Paula, et al.
Published: (2024)
by: Ontalvilla, Paula, et al.
Published: (2024)
Advancing Parameter Efficiency in Fine-tuning via Representation Editing
by: Wu, Muling, et al.
Published: (2024)
by: Wu, Muling, et al.
Published: (2024)
Representational Analysis of Binding in Language Models
by: Dai, Qin, et al.
Published: (2024)
by: Dai, Qin, et al.
Published: (2024)
ReTok: Replacing Tokenizer to Enhance Representation Efficiency in Large Language Model
by: Gu, Shuhao, et al.
Published: (2024)
by: Gu, Shuhao, et al.
Published: (2024)
ResoFilter: Fine-grained Synthetic Data Filtering for Large Language Models through Data-Parameter Resonance Analysis
by: Tu, Zeao, et al.
Published: (2024)
by: Tu, Zeao, et al.
Published: (2024)
Improving Multilingual Language Models by Aligning Representations through Steering
by: Mahmoud, Omar, et al.
Published: (2025)
by: Mahmoud, Omar, et al.
Published: (2025)
ERAGent: Enhancing Retrieval-Augmented Language Models with Improved Accuracy, Efficiency, and Personalization
by: Shi, Yunxiao, et al.
Published: (2024)
by: Shi, Yunxiao, et al.
Published: (2024)
On the Representational Capacity of Neural Language Models with Chain-of-Thought Reasoning
by: Nowak, Franz, et al.
Published: (2024)
by: Nowak, Franz, et al.
Published: (2024)
On the Representational Capacity of Recurrent Neural Language Models
by: Nowak, Franz, et al.
Published: (2023)
by: Nowak, Franz, et al.
Published: (2023)
Parameter-Efficient Tuning Large Language Models for Graph Representation Learning
by: Zhu, Qi, et al.
Published: (2024)
by: Zhu, Qi, et al.
Published: (2024)
Improving the Accuracy and Efficiency of Legal Document Tagging with Large Language Models and Instruction Prompts
by: Johnson, Emily, et al.
Published: (2025)
by: Johnson, Emily, et al.
Published: (2025)
Scaling Beyond Masked Diffusion Language Models
by: Sahoo, Subham Sekhar, et al.
Published: (2026)
by: Sahoo, Subham Sekhar, et al.
Published: (2026)
Relationship Detection on Tabular Data Using Statistical Analysis and Large Language Models
by: Koletsis, Panagiotis, et al.
Published: (2025)
by: Koletsis, Panagiotis, et al.
Published: (2025)
Non-Fluent Synthetic Target-Language Data Improve Neural Machine Translation
by: Sánchez-Cartagena, Víctor M., et al.
Published: (2024)
by: Sánchez-Cartagena, Víctor M., et al.
Published: (2024)
Improving Language Models Trained on Translated Data with Continual Pre-Training and Dictionary Learning Analysis
by: Boughorbel, Sabri, et al.
Published: (2024)
by: Boughorbel, Sabri, et al.
Published: (2024)
Vision-Language Models Align with Human Neural Representations in Concept Processing
by: Bavaresco, Anna, et al.
Published: (2024)
by: Bavaresco, Anna, et al.
Published: (2024)
Scaling Parameter-Constrained Language Models with Quality Data
by: Chang, Ernie, et al.
Published: (2024)
by: Chang, Ernie, et al.
Published: (2024)
Large Language Models as Financial Data Annotators: A Study on Effectiveness and Efficiency
by: Aguda, Toyin, et al.
Published: (2024)
by: Aguda, Toyin, et al.
Published: (2024)
Linguistic Entity Masking to Improve Cross-Lingual Representation of Multilingual Language Models for Low-Resource Languages
by: Fernando, Aloka, et al.
Published: (2025)
by: Fernando, Aloka, et al.
Published: (2025)
DiLaDiff: Distilled Latent-Augmented Diffusion for Language Modeling
by: Lemercier, Jean-Marie, et al.
Published: (2026)
by: Lemercier, Jean-Marie, et al.
Published: (2026)
Investigating Low-Rank Training in Transformer Language Models: Efficiency and Scaling Analysis
by: Wei, Xiuying, et al.
Published: (2024)
by: Wei, Xiuying, et al.
Published: (2024)
How Tokenization Limits Phonological Knowledge Representation in Language Models and How to Improve Them
by: Liao, Disen, et al.
Published: (2026)
by: Liao, Disen, et al.
Published: (2026)
Analysis and Visualization of Linguistic Structures in Large Language Models: Neural Representations of Verb-Particle Constructions in BERT
by: Kissane, Hassane, et al.
Published: (2024)
by: Kissane, Hassane, et al.
Published: (2024)
Speech Representation Learning Revisited: The Necessity of Separate Learnable Parameters and Robust Data Augmentation
by: Yadav, Hemant, et al.
Published: (2024)
by: Yadav, Hemant, et al.
Published: (2024)
On Linear Representations and Pretraining Data Frequency in Language Models
by: Merullo, Jack, et al.
Published: (2025)
by: Merullo, Jack, et al.
Published: (2025)
Task-Specific Efficiency Analysis: When Small Language Models Outperform Large Language Models
by: Cao, Jinghan, et al.
Published: (2026)
by: Cao, Jinghan, et al.
Published: (2026)
Learning to Reduce: Optimal Representations of Structured Data in Prompting Large Language Models
by: Lee, Younghun, et al.
Published: (2024)
by: Lee, Younghun, et al.
Published: (2024)
Structured Code Representations Enable Data-Efficient Adaptation of Code Language Models
by: Agarwal, Mayank, et al.
Published: (2024)
by: Agarwal, Mayank, et al.
Published: (2024)
Improving Sample Efficiency of Reinforcement Learning with Background Knowledge from Large Language Models
by: Zhang, Fuxiang, et al.
Published: (2024)
by: Zhang, Fuxiang, et al.
Published: (2024)
Improving Training Efficiency and Reducing Maintenance Costs via Language Specific Model Merging
by: Dmonte, Alphaeus, et al.
Published: (2026)
by: Dmonte, Alphaeus, et al.
Published: (2026)
Improving Data and Reward Design for Scientific Reasoning in Large Language Models
by: Chen, Zijie, et al.
Published: (2026)
by: Chen, Zijie, et al.
Published: (2026)
Decoding In-Context Learning: Neuroscience-inspired Analysis of Representations in Large Language Models
by: Yousefi, Safoora, et al.
Published: (2023)
by: Yousefi, Safoora, et al.
Published: (2023)
Similar Items
-
From Robustness to Improved Generalization and Calibration in Pre-trained Language Models
by: Jukić, Josip, et al.
Published: (2024) -
Disentangling Latent Shifts of In-Context Learning with Weak Supervision
by: Jukić, Josip, et al.
Published: (2024) -
Context Parametrization with Compositional Adapters
by: Jukić, Josip, et al.
Published: (2025) -
Characterizing Linguistic Shifts in Croatian News via Diachronic Word Embeddings
by: Dukić, David, et al.
Published: (2025) -
Out-of-Distribution Detection by Leveraging Between-Layer Transformation Smoothness
by: Jelenić, Fran, et al.
Published: (2023)