Parameter-Efficient Transformer Embeddings
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Ndubuaku, Henry, Talhi, Mouad |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
DYNAMAX: Dynamic computing for Transformers and Mamba based architectures
von: Nogales, Miguel, et al.
Veröffentlicht: (2025)
von: Nogales, Miguel, et al.
Veröffentlicht: (2025)
Smoothed Embeddings for Robust Language Models
von: Hase, Ryo, et al.
Veröffentlicht: (2025)
von: Hase, Ryo, et al.
Veröffentlicht: (2025)
Empirical analysis of binding precedent efficiency in Brazilian Supreme Court via case classification
von: Tinarrage, Raphaël, et al.
Veröffentlicht: (2024)
von: Tinarrage, Raphaël, et al.
Veröffentlicht: (2024)
Atyaephyra at SemEval-2025 Task 4: Low-Rank Negative Preference Optimization
von: Bronec, Jan, et al.
Veröffentlicht: (2025)
von: Bronec, Jan, et al.
Veröffentlicht: (2025)
A Comprehensive Survey of Small Language Models in the Era of Large Language Models: Techniques, Enhancements, Applications, Collaboration with LLMs, and Trustworthiness
von: Wang, Fali, et al.
Veröffentlicht: (2024)
von: Wang, Fali, et al.
Veröffentlicht: (2024)
Evaluating Embedding Generalization: How LLMs, LoRA, and SLERP Shape Representational Geometry
von: Kabane, Siyaxolisa
Veröffentlicht: (2025)
von: Kabane, Siyaxolisa
Veröffentlicht: (2025)
Bridging the Language Gap: Enhancing Multilingual Prompt-Based Code Generation in LLMs via Zero-Shot Cross-Lingual Transfer
von: Li, Mingda, et al.
Veröffentlicht: (2024)
von: Li, Mingda, et al.
Veröffentlicht: (2024)
A Survey on Collaborating Small and Large Language Models for Performance, Cost-effectiveness, Cloud-edge Privacy, and Trustworthiness
von: Wang, Fali, et al.
Veröffentlicht: (2025)
von: Wang, Fali, et al.
Veröffentlicht: (2025)
Do Reasoning Models Enhance Embedding Models?
von: Chan, Wun Yu, et al.
Veröffentlicht: (2026)
von: Chan, Wun Yu, et al.
Veröffentlicht: (2026)
Uncovering Uncertainty in Transformer Inference
von: Brothers, Greyson, et al.
Veröffentlicht: (2024)
von: Brothers, Greyson, et al.
Veröffentlicht: (2024)
A Language Model-Driven Semi-Supervised Ensemble Framework for Illicit Market Detection Across Deep/Dark Web and Social Platforms
von: Yazdanjue, Navid, et al.
Veröffentlicht: (2025)
von: Yazdanjue, Navid, et al.
Veröffentlicht: (2025)
OptPO: Optimal Rollout Allocation for Test-time Policy Optimization
von: Wang, Youkang, et al.
Veröffentlicht: (2025)
von: Wang, Youkang, et al.
Veröffentlicht: (2025)
InhibiDistilbert: Knowledge Distillation for a ReLU and Addition-based Transformer
von: Zhang, Tony, et al.
Veröffentlicht: (2025)
von: Zhang, Tony, et al.
Veröffentlicht: (2025)
GraphSkill: Documentation-Guided Hierarchical Retrieval-Augmented Coding for Complex Graph Reasoning
von: Wang, Fali, et al.
Veröffentlicht: (2026)
von: Wang, Fali, et al.
Veröffentlicht: (2026)
Graph Neural Networks for Brain Graph Learning: A Survey
von: Luo, Xuexiong, et al.
Veröffentlicht: (2024)
von: Luo, Xuexiong, et al.
Veröffentlicht: (2024)
Research on a hybrid LSTM-CNN-Attention model for text-based web content classification
von: Kuz, Mykola, et al.
Veröffentlicht: (2025)
von: Kuz, Mykola, et al.
Veröffentlicht: (2025)
Linguistic Collapse: Neural Collapse in (Large) Language Models
von: Wu, Robert, et al.
Veröffentlicht: (2024)
von: Wu, Robert, et al.
Veröffentlicht: (2024)
AI in Investment Analysis: LLMs for Equity Stock Ratings
von: Papasotiriou, Kassiani, et al.
Veröffentlicht: (2024)
von: Papasotiriou, Kassiani, et al.
Veröffentlicht: (2024)
Model Surgery: Modulating LLM's Behavior Via Simple Parameter Editing
von: Wang, Huanqian, et al.
Veröffentlicht: (2024)
von: Wang, Huanqian, et al.
Veröffentlicht: (2024)
Quantum NLP models on Natural Language Inference
von: Sun, Ling, et al.
Veröffentlicht: (2025)
von: Sun, Ling, et al.
Veröffentlicht: (2025)
Semantic Retention and Extreme Compression in LLMs: Can We Have Both?
von: Laborde, Stanislas, et al.
Veröffentlicht: (2025)
von: Laborde, Stanislas, et al.
Veröffentlicht: (2025)
Symphonym: Universal Phonetic Embeddings for Cross-Script Name Matching
von: Gadd, Stephen
Veröffentlicht: (2026)
von: Gadd, Stephen
Veröffentlicht: (2026)
Understanding the geometry of deep learning with decision boundary volume
von: Burfitt, Matthew, et al.
Veröffentlicht: (2026)
von: Burfitt, Matthew, et al.
Veröffentlicht: (2026)
OPENXRD: A Comprehensive Benchmark Framework for LLM/MLLM XRD Question Answering
von: Vosoughi, Ali, et al.
Veröffentlicht: (2025)
von: Vosoughi, Ali, et al.
Veröffentlicht: (2025)
Pay Attention to What You Need
von: Gao, Yifei, et al.
Veröffentlicht: (2023)
von: Gao, Yifei, et al.
Veröffentlicht: (2023)
Uncertainty Estimation and Quantification for LLMs: A Simple Supervised Approach
von: Liu, Linyu, et al.
Veröffentlicht: (2024)
von: Liu, Linyu, et al.
Veröffentlicht: (2024)
FedUNet: A Lightweight Additive U-Net Module for Federated Learning with Heterogeneous Models
von: Seo, Beomseok, et al.
Veröffentlicht: (2025)
von: Seo, Beomseok, et al.
Veröffentlicht: (2025)
Logistic Regression makes small LLMs strong and explainable "tens-of-shot" classifiers
von: Buckmann, Marcus, et al.
Veröffentlicht: (2024)
von: Buckmann, Marcus, et al.
Veröffentlicht: (2024)
Generative AI Models: Opportunities and Risks for Industry and Authorities
von: Alt, Tobias, et al.
Veröffentlicht: (2024)
von: Alt, Tobias, et al.
Veröffentlicht: (2024)
Overcoming Catastrophic Forgetting in Federated Class-Incremental Learning via Federated Global Twin Generator
von: Nguyen, Thinh, et al.
Veröffentlicht: (2024)
von: Nguyen, Thinh, et al.
Veröffentlicht: (2024)
CogniLoad: A Synthetic Natural Language Reasoning Benchmark With Tunable Length, Intrinsic Difficulty, and Distractor Density
von: Kaiser, Daniel, et al.
Veröffentlicht: (2025)
von: Kaiser, Daniel, et al.
Veröffentlicht: (2025)
Investigating Neuron Ablation in Attention Heads: The Case for Peak Activation Centering
von: Pochinkov, Nicholas, et al.
Veröffentlicht: (2024)
von: Pochinkov, Nicholas, et al.
Veröffentlicht: (2024)
Attention Maps in 3D Shape Classification for Dental Stage Estimation with Class Node Graph Attention Networks
von: Buyukcakir, Barkin, et al.
Veröffentlicht: (2025)
von: Buyukcakir, Barkin, et al.
Veröffentlicht: (2025)
Forging GEMs: Advancing Greek NLP through Quality-Based Corpus Curation
von: Apostolopoulou, Alexandra, et al.
Veröffentlicht: (2025)
von: Apostolopoulou, Alexandra, et al.
Veröffentlicht: (2025)
DanceHA: A Multi-Agent Framework for Document-Level Aspect-Based Sentiment Analysis
von: Wang, Lei, et al.
Veröffentlicht: (2026)
von: Wang, Lei, et al.
Veröffentlicht: (2026)
Handling Out-of-Distribution Data: A Survey
von: Tamang, Lakpa, et al.
Veröffentlicht: (2025)
von: Tamang, Lakpa, et al.
Veröffentlicht: (2025)
On Memory: A comparison of memory mechanisms in world models
von: Laird, Eli J., et al.
Veröffentlicht: (2025)
von: Laird, Eli J., et al.
Veröffentlicht: (2025)
Symbolic Branch Networks: Tree-Inherited Neural Models for Interpretable Multiclass Classification
von: Rodríguez-Salas, Dalia
Veröffentlicht: (2025)
von: Rodríguez-Salas, Dalia
Veröffentlicht: (2025)
Temporal Attention Evolutional Graph Convolutional Network for Multivariate Time Series Forecasting
von: Zhao, Xinlong, et al.
Veröffentlicht: (2025)
von: Zhao, Xinlong, et al.
Veröffentlicht: (2025)
Clinical Data Goes MEDS? Let's OWL make sense of it
von: Marfoglia, Alberto, et al.
Veröffentlicht: (2026)
von: Marfoglia, Alberto, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
DYNAMAX: Dynamic computing for Transformers and Mamba based architectures
von: Nogales, Miguel, et al.
Veröffentlicht: (2025) -
Smoothed Embeddings for Robust Language Models
von: Hase, Ryo, et al.
Veröffentlicht: (2025) -
Empirical analysis of binding precedent efficiency in Brazilian Supreme Court via case classification
von: Tinarrage, Raphaël, et al.
Veröffentlicht: (2024) -
Atyaephyra at SemEval-2025 Task 4: Low-Rank Negative Preference Optimization
von: Bronec, Jan, et al.
Veröffentlicht: (2025) -
A Comprehensive Survey of Small Language Models in the Era of Large Language Models: Techniques, Enhancements, Applications, Collaboration with LLMs, and Trustworthiness
von: Wang, Fali, et al.
Veröffentlicht: (2024)