Sequence Repetition Enhances Token Embeddings and Improves Sequence Labeling with Decoder-only Language Models
Fuente:
arXiv
Saved in:
| Main Authors: | Kukić, Matija Luka, Čuljak, Marko, Dukić, David, Tutek, Martin, Šnajder, Jan |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Looking Right is Sometimes Right: Investigating the Capabilities of Decoder-only LLMs for Sequence Labeling
by: Dukić, David, et al.
Published: (2024)
by: Dukić, David, et al.
Published: (2024)
Supervised In-Context Fine-Tuning for Generative Sequence Labeling
by: Dukić, David, et al.
Published: (2025)
by: Dukić, David, et al.
Published: (2025)
Characterizing Linguistic Shifts in Croatian News via Diachronic Word Embeddings
by: Dukić, David, et al.
Published: (2025)
by: Dukić, David, et al.
Published: (2025)
Improving Transfer Learning for Sequence Labeling Tasks by Adapting Pre-trained Neural Language Models
by: Dukić, David
Published: (2025)
by: Dukić, David
Published: (2025)
Are ELECTRA's Sentence Embeddings Beyond Repair? The Case of Semantic Textual Similarity
by: Rep, Ivan, et al.
Published: (2024)
by: Rep, Ivan, et al.
Published: (2024)
Context Parametrization with Compositional Adapters
by: Jukić, Josip, et al.
Published: (2025)
by: Jukić, Josip, et al.
Published: (2025)
Leveraging Open Information Extraction for More Robust Domain Transfer of Event Trigger Detection
by: Dukić, David, et al.
Published: (2023)
by: Dukić, David, et al.
Published: (2023)
TakeLab Retriever: AI-Driven Search Engine for Articles from Croatian News Outlets
by: Dukić, David, et al.
Published: (2024)
by: Dukić, David, et al.
Published: (2024)
From Robustness to Improved Generalization and Calibration in Pre-trained Language Models
by: Jukić, Josip, et al.
Published: (2024)
by: Jukić, Josip, et al.
Published: (2024)
Out-of-Distribution Detection by Leveraging Between-Layer Transformation Smoothness
by: Jelenić, Fran, et al.
Published: (2023)
by: Jelenić, Fran, et al.
Published: (2023)
Repetition Improves Language Model Embeddings
by: Springer, Jacob Mitchell, et al.
Published: (2024)
by: Springer, Jacob Mitchell, et al.
Published: (2024)
Causal2Vec: Improving Decoder-only LLMs as Embedding Models through a Contextual Token
by: Lin, Ailiang, et al.
Published: (2025)
by: Lin, Ailiang, et al.
Published: (2025)
Compressing Sequences in the Latent Embedding Space: $K$-Token Merging for Large Language Models
by: Xu, Zihao, et al.
Published: (2026)
by: Xu, Zihao, et al.
Published: (2026)
Claim Check-Worthiness Detection: How Well do LLMs Grasp Annotation Guidelines?
by: Majer, Laura, et al.
Published: (2024)
by: Majer, Laura, et al.
Published: (2024)
CATfOOD: Counterfactual Augmented Training for Improving Out-of-Domain Performance and Calibration
by: Sachdeva, Rachneet, et al.
Published: (2023)
by: Sachdeva, Rachneet, et al.
Published: (2023)
Sequence to Sequence Reward Modeling: Improving RLHF by Language Feedback
by: Zhou, Jiayi, et al.
Published: (2024)
by: Zhou, Jiayi, et al.
Published: (2024)
Disentangling Latent Shifts of In-Context Learning with Weak Supervision
by: Jukić, Josip, et al.
Published: (2024)
by: Jukić, Josip, et al.
Published: (2024)
Dependency Graph Parsing as Sequence Labeling
by: Ezquerro, Ana, et al.
Published: (2024)
by: Ezquerro, Ana, et al.
Published: (2024)
Improving Low-Resource Sequence Labeling with Knowledge Fusion and Contextual Label Explanations
by: Lai, Peichao, et al.
Published: (2025)
by: Lai, Peichao, et al.
Published: (2025)
REVS: Unlearning Sensitive Information in Language Models via Rank Editing in the Vocabulary Space
by: Ashuach, Tomer, et al.
Published: (2024)
by: Ashuach, Tomer, et al.
Published: (2024)
ToPro: Token-Level Prompt Decomposition for Cross-Lingual Sequence Labeling Tasks
by: Ma, Bolei, et al.
Published: (2024)
by: Ma, Bolei, et al.
Published: (2024)
Bringing Emerging Architectures to Sequence Labeling in NLP
by: Ezquerro, Ana, et al.
Published: (2025)
by: Ezquerro, Ana, et al.
Published: (2025)
Temporal Tokenization Strategies for Event Sequence Modeling with Large Language Models
by: Liu, Zefang, et al.
Published: (2025)
by: Liu, Zefang, et al.
Published: (2025)
LLMs for Targeted Sentiment in News Headlines: Exploring the Descriptive-Prescriptive Dilemma
by: Juroš, Jana, et al.
Published: (2024)
by: Juroš, Jana, et al.
Published: (2024)
Advancing Cross-lingual Aspect-Based Sentiment Analysis with LLMs and Constrained Decoding for Sequence-to-Sequence Models
by: Šmíd, Jakub, et al.
Published: (2025)
by: Šmíd, Jakub, et al.
Published: (2025)
Chinese Sequence Labeling with Semi-Supervised Boundary-Aware Language Model Pre-training
by: Zhang, Longhui, et al.
Published: (2024)
by: Zhang, Longhui, et al.
Published: (2024)
E2S2: Encoding-Enhanced Sequence-to-Sequence Pretraining for Language Understanding and Generation
by: Zhong, Qihuang, et al.
Published: (2022)
by: Zhong, Qihuang, et al.
Published: (2022)
Improving Language Transfer Capability of Decoder-only Architecture in Multilingual Neural Machine Translation
by: Qu, Zhi, et al.
Published: (2024)
by: Qu, Zhi, et al.
Published: (2024)
A Global Context Mechanism for Sequence Labeling
by: Xu, Conglei, et al.
Published: (2023)
by: Xu, Conglei, et al.
Published: (2023)
Textless Dependency Parsing by Labeled Sequence Prediction
by: Kando, Shunsuke, et al.
Published: (2024)
by: Kando, Shunsuke, et al.
Published: (2024)
Semiparametric Token-Sequence Co-Supervision
by: Lee, Hyunji, et al.
Published: (2024)
by: Lee, Hyunji, et al.
Published: (2024)
Multi-word Tokenization for Sequence Compression
by: Gee, Leonidas, et al.
Published: (2024)
by: Gee, Leonidas, et al.
Published: (2024)
Repetition Neurons: How Do Language Models Produce Repetitions?
by: Hiraoka, Tatsuya, et al.
Published: (2024)
by: Hiraoka, Tatsuya, et al.
Published: (2024)
Retrieval Backward Attention without Additional Training: Enhance Embeddings of Large Language Models via Repetition
by: Duan, Yifei, et al.
Published: (2025)
by: Duan, Yifei, et al.
Published: (2025)
The Repetition Threshold for Rote Sequences
by: Ollinger, Nicolas, et al.
Published: (2024)
by: Ollinger, Nicolas, et al.
Published: (2024)
What Makes You CLIC: Detection of Croatian Clickbait Headlines
by: Anđelić, Marija, et al.
Published: (2025)
by: Anđelić, Marija, et al.
Published: (2025)
Training Language Models on Synthetic Edit Sequences Improves Code Synthesis
by: Piterbarg, Ulyana, et al.
Published: (2024)
by: Piterbarg, Ulyana, et al.
Published: (2024)
Sequence-to-Sequence Spanish Pre-trained Language Models
by: Araujo, Vladimir, et al.
Published: (2023)
by: Araujo, Vladimir, et al.
Published: (2023)
SciAnnotate: A Tool for Integrating Weak Labeling Sources for Sequence Labeling
by: Liu, Mengyang, et al.
Published: (2022)
by: Liu, Mengyang, et al.
Published: (2022)
KV-Embedding: Training-free Text Embedding via Internal KV Re-routing in Decoder-only LLMs
by: Tang, Yixuan, et al.
Published: (2026)
by: Tang, Yixuan, et al.
Published: (2026)
Similar Items
-
Looking Right is Sometimes Right: Investigating the Capabilities of Decoder-only LLMs for Sequence Labeling
by: Dukić, David, et al.
Published: (2024) -
Supervised In-Context Fine-Tuning for Generative Sequence Labeling
by: Dukić, David, et al.
Published: (2025) -
Characterizing Linguistic Shifts in Croatian News via Diachronic Word Embeddings
by: Dukić, David, et al.
Published: (2025) -
Improving Transfer Learning for Sequence Labeling Tasks by Adapting Pre-trained Neural Language Models
by: Dukić, David
Published: (2025) -
Are ELECTRA's Sentence Embeddings Beyond Repair? The Case of Semantic Textual Similarity
by: Rep, Ivan, et al.
Published: (2024)