Extracting Sentence Embeddings from Pretrained Transformer Models
Fuente:
arXiv
Guardado en:
| Autores principales: | Stankevičius, Lukas, Lukoševičius, Mantas |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Sentiment Analysis of Lithuanian Online Reviews Using Large Language Models
por: Vileikytė, Brigita, et al.
Publicado: (2024)
por: Vileikytė, Brigita, et al.
Publicado: (2024)
Inference acceleration for large language models using "stairs" assisted greedy generation
por: Grigaliūnas, Domas, et al.
Publicado: (2024)
por: Grigaliūnas, Domas, et al.
Publicado: (2024)
Latent Object Permanence: Topological Phase Transitions, Free-Energy Principles, and Renormalization Group Flows in Deep Transformer Manifolds
por: Alpay, Faruk, et al.
Publicado: (2026)
por: Alpay, Faruk, et al.
Publicado: (2026)
Optimizing Large Language Models for OpenAPI Code Completion
por: Petryshyn, Bohdan, et al.
Publicado: (2024)
por: Petryshyn, Bohdan, et al.
Publicado: (2024)
Mechanistic Analysis of Catastrophic Forgetting in Large Language Models During Continual Fine-tuning
por: Imanov, Olaf Yunus Laitinen
Publicado: (2026)
por: Imanov, Olaf Yunus Laitinen
Publicado: (2026)
Strategic Doctrine Language Models (sdLM): A Learning-System Framework for Doctrinal Consistency and Geopolitical Forecasting
por: Imanov, Olaf Yunus Laitinen, et al.
Publicado: (2026)
por: Imanov, Olaf Yunus Laitinen, et al.
Publicado: (2026)
When Does Content-Based Routing Work? Representation Requirements for Selective Attention in Hybrid Sequence Models
por: Basu, Abhinaba
Publicado: (2026)
por: Basu, Abhinaba
Publicado: (2026)
ReFactor GNNs: Revisiting Factorisation-based Models from a Message-Passing Perspective
por: Chen, Yihong, et al.
Publicado: (2022)
por: Chen, Yihong, et al.
Publicado: (2022)
mHC-SSM: Manifold-Constrained Hyper-Connections for State Space Language Models with Stream-Specialized Adapters
por: Mutlu, Abdulvahap, et al.
Publicado: (2026)
por: Mutlu, Abdulvahap, et al.
Publicado: (2026)
Improving Large-Scale k-Nearest Neighbor Text Categorization with Label Autoencoders
por: Ribadas-Pena, Francisco J., et al.
Publicado: (2024)
por: Ribadas-Pena, Francisco J., et al.
Publicado: (2024)
Harnessing non-adversarial robustness in large language models
por: Zhou, Qinghua, et al.
Publicado: (2026)
por: Zhou, Qinghua, et al.
Publicado: (2026)
SafeAnchor: Preventing Cumulative Safety Erosion in Continual Domain Adaptation of Large Language Models
por: Guo, Dongxin, et al.
Publicado: (2026)
por: Guo, Dongxin, et al.
Publicado: (2026)
On Self-improving Token Embeddings
por: Kubek, Mario M., et al.
Publicado: (2025)
por: Kubek, Mario M., et al.
Publicado: (2025)
NOTAI.AI: Explainable Detection of Machine-Generated Text via Curvature and Feature Attribution
por: Breneur, Oleksandr Marchenko, et al.
Publicado: (2026)
por: Breneur, Oleksandr Marchenko, et al.
Publicado: (2026)
Rethinking the Multilingual Reasoning Gap with Layer Swap
por: Lasbordes, Maxence, et al.
Publicado: (2026)
por: Lasbordes, Maxence, et al.
Publicado: (2026)
MAcPNN: Mutual Assisted Learning on Data Streams with Temporal Dependence
por: Giannini, Federico, et al.
Publicado: (2026)
por: Giannini, Federico, et al.
Publicado: (2026)
The Concept Allocation Zone: Tracking How Concepts Form Across Transformer Depth
por: Henry, James
Publicado: (2026)
por: Henry, James
Publicado: (2026)
CLMN: Concept based Language Models via Neural Symbolic Reasoning
por: Yang, Yibo
Publicado: (2025)
por: Yang, Yibo
Publicado: (2025)
Think-at-Hard: Selective Latent Iterations to Improve Reasoning Language Models
por: Fu, Tianyu, et al.
Publicado: (2025)
por: Fu, Tianyu, et al.
Publicado: (2025)
Transactional Attention: Semantic Sponsorship for KV-Cache Retention
por: Basu, Abhinaba
Publicado: (2026)
por: Basu, Abhinaba
Publicado: (2026)
Mitigating Position-Shift Failures in Text-Based Modular Arithmetic via Position Curriculum and Template Diversity
por: Yudin, Nikolay
Publicado: (2026)
por: Yudin, Nikolay
Publicado: (2026)
HEFT: A Coarse-to-Fine Hierarchy for Enhancing the Efficiency and Accuracy of Language Model Reasoning
por: Hill, Brennen
Publicado: (2025)
por: Hill, Brennen
Publicado: (2025)
How Pruning Reshapes Features: Sparse Autoencoder Analysis of Weight-Pruned Language Models
por: Borobia, Hector, et al.
Publicado: (2026)
por: Borobia, Hector, et al.
Publicado: (2026)
WebMap -- Large Language Model-assisted Semantic Link Induction in the Web
por: Pokharel, Shiraj, et al.
Publicado: (2025)
por: Pokharel, Shiraj, et al.
Publicado: (2025)
Linguistic Collapse: Neural Collapse in (Large) Language Models
por: Wu, Robert, et al.
Publicado: (2024)
por: Wu, Robert, et al.
Publicado: (2024)
Combating data scarcity in recommendation services: Integrating cognitive types of VARK and neural network technologies (LLM)
por: Zmanovskii, Nikita
Publicado: (2026)
por: Zmanovskii, Nikita
Publicado: (2026)
ProbeScale: Probing Analysis to Optimize Neural Scaling Laws for Efficient Small Language Model Inference
por: Das, Sourav
Publicado: (2026)
por: Das, Sourav
Publicado: (2026)
Reducing Labeling Costs in Sentiment Analysis via Semi-Supervised Learning
por: Jafarlou, Minoo, et al.
Publicado: (2024)
por: Jafarlou, Minoo, et al.
Publicado: (2024)
How much do LLMs learn from negative examples?
por: Hamdan, Shadi, et al.
Publicado: (2025)
por: Hamdan, Shadi, et al.
Publicado: (2025)
Vis-CoT: A Human-in-the-Loop Framework for Interactive Visualization and Intervention in LLM Chain-of-Thought Reasoning
por: Pather, Kaviraj, et al.
Publicado: (2025)
por: Pather, Kaviraj, et al.
Publicado: (2025)
Unsolvability Ceiling in Multi-LLM Routing: An Empirical Study of Evaluation Artifacts
por: Garg, Saloni, et al.
Publicado: (2026)
por: Garg, Saloni, et al.
Publicado: (2026)
Research on a hybrid LSTM-CNN-Attention model for text-based web content classification
por: Kuz, Mykola, et al.
Publicado: (2025)
por: Kuz, Mykola, et al.
Publicado: (2025)
Enhancing Ultra-Low-Bit Quantization of Large Language Models Through Saliency-Aware Partial Retraining
por: Cao, Deyu, et al.
Publicado: (2025)
por: Cao, Deyu, et al.
Publicado: (2025)
Causal Dimensionality of Transformer Representations: Measurement, Scaling, and Layer Structure
por: Sarkar, Nilesh, et al.
Publicado: (2026)
por: Sarkar, Nilesh, et al.
Publicado: (2026)
AIPsy-Affect: A Keyword-Free Clinical Stimulus Battery for Mechanistic Interpretability of Emotion in Language Models
por: Keeman, Michael
Publicado: (2026)
por: Keeman, Michael
Publicado: (2026)
ProactBench: Beyond What The User Asked For
por: Harfi, Sepehr, et al.
Publicado: (2026)
por: Harfi, Sepehr, et al.
Publicado: (2026)
Multipole Semantic Attention: A Fast Approximation of Softmax Attention for Pretraining
por: Mitchell, Rupert, et al.
Publicado: (2025)
por: Mitchell, Rupert, et al.
Publicado: (2025)
CogniLoad: A Synthetic Natural Language Reasoning Benchmark With Tunable Length, Intrinsic Difficulty, and Distractor Density
por: Kaiser, Daniel, et al.
Publicado: (2025)
por: Kaiser, Daniel, et al.
Publicado: (2025)
From Noise to Diversity: Random Embedding Injection in LLM Reasoning
por: Kim, Heejun, et al.
Publicado: (2026)
por: Kim, Heejun, et al.
Publicado: (2026)
Do Reasoning Models Enhance Embedding Models?
por: Chan, Wun Yu, et al.
Publicado: (2026)
por: Chan, Wun Yu, et al.
Publicado: (2026)
Ejemplares similares
-
Sentiment Analysis of Lithuanian Online Reviews Using Large Language Models
por: Vileikytė, Brigita, et al.
Publicado: (2024) -
Inference acceleration for large language models using "stairs" assisted greedy generation
por: Grigaliūnas, Domas, et al.
Publicado: (2024) -
Latent Object Permanence: Topological Phase Transitions, Free-Energy Principles, and Renormalization Group Flows in Deep Transformer Manifolds
por: Alpay, Faruk, et al.
Publicado: (2026) -
Optimizing Large Language Models for OpenAPI Code Completion
por: Petryshyn, Bohdan, et al.
Publicado: (2024) -
Mechanistic Analysis of Catastrophic Forgetting in Large Language Models During Continual Fine-tuning
por: Imanov, Olaf Yunus Laitinen
Publicado: (2026)