Guardado en:
| Autores principales: | Veisi, Ali, Amirzadeh, Hamidreza, Mansourian, Amir |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | https://arxiv.org/abs/2503.08067 |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Context-aware Rotary Position Embedding
por: Veisi, Ali, et al.
Publicado: (2025)
por: Veisi, Ali, et al.
Publicado: (2025)
How Language Models Prioritize Contextual Grammatical Cues?
por: Amirzadeh, Hamidreza, et al.
Publicado: (2024)
por: Amirzadeh, Hamidreza, et al.
Publicado: (2024)
data2lang2vec: Data Driven Typological Features Completion
por: Amirzadeh, Hamidreza, et al.
Publicado: (2024)
por: Amirzadeh, Hamidreza, et al.
Publicado: (2024)
In-Context Learning (and Unlearning) of Length Biases
por: Schoch, Stephanie, et al.
Publicado: (2025)
por: Schoch, Stephanie, et al.
Publicado: (2025)
ParallelComp: Parallel Long-Context Compressor for Length Extrapolation
por: Xiong, Jing, et al.
Publicado: (2025)
por: Xiong, Jing, et al.
Publicado: (2025)
DAPE: Data-Adaptive Positional Encoding for Length Extrapolation
por: Zheng, Chuanyang, et al.
Publicado: (2024)
por: Zheng, Chuanyang, et al.
Publicado: (2024)
CLEX: Continuous Length Extrapolation for Large Language Models
por: Chen, Guanzheng, et al.
Publicado: (2023)
por: Chen, Guanzheng, et al.
Publicado: (2023)
Extrapolation by Association: Length Generalization Transfer in Transformers
por: Cai, Ziyang, et al.
Publicado: (2025)
por: Cai, Ziyang, et al.
Publicado: (2025)
Gammatonegram Representation for End-to-End Dysarthric Speech Processing Tasks: Speech Recognition, Speaker Identification, and Intelligibility Assessment
por: Farhadipour, Aref, et al.
Publicado: (2023)
por: Farhadipour, Aref, et al.
Publicado: (2023)
Bayesian Network Fusion of Large Language Models for Sentiment Analysis
por: Amirzadeh, Rasoul, et al.
Publicado: (2025)
por: Amirzadeh, Rasoul, et al.
Publicado: (2025)
Information Entropy Invariance: Enhancing Length Extrapolation in Attention Mechanisms
por: Li, Kewei, et al.
Publicado: (2025)
por: Li, Kewei, et al.
Publicado: (2025)
Length Extrapolation of Transformers: A Survey from the Perspective of Positional Encoding
por: Zhao, Liang, et al.
Publicado: (2023)
por: Zhao, Liang, et al.
Publicado: (2023)
Bayesian Attention Mechanism: A Probabilistic Framework for Positional Encoding and Context Length Extrapolation
por: Bianchessi, Arthur S., et al.
Publicado: (2025)
por: Bianchessi, Arthur S., et al.
Publicado: (2025)
From Interpolation to Extrapolation: Complete Length Generalization for Arithmetic Transformers
por: Duan, Shaoxiong, et al.
Publicado: (2023)
por: Duan, Shaoxiong, et al.
Publicado: (2023)
DAPE V2: Process Attention Score as Feature Map for Length Extrapolation
por: Zheng, Chuanyang, et al.
Publicado: (2024)
por: Zheng, Chuanyang, et al.
Publicado: (2024)
Effective Length Extrapolation via Dimension-Wise Positional Embeddings Manipulation
por: Lu, Yi, et al.
Publicado: (2025)
por: Lu, Yi, et al.
Publicado: (2025)
Enhancing Length Extrapolation in Sequential Models with Pointer-Augmented Neural Memory
por: Le, Hung, et al.
Publicado: (2024)
por: Le, Hung, et al.
Publicado: (2024)
Squeezed Attention: Accelerating Long Context Length LLM Inference
por: Hooper, Coleman, et al.
Publicado: (2024)
por: Hooper, Coleman, et al.
Publicado: (2024)
DCIS: Efficient Length Extrapolation of LLMs via Divide-and-Conquer Scaling Factor Search
por: Yang, Lei, et al.
Publicado: (2024)
por: Yang, Lei, et al.
Publicado: (2024)
KurdSTS: The Kurdish Semantic Textual Similarity
por: Abdullah, Abdulhady Abas, et al.
Publicado: (2025)
por: Abdullah, Abdulhady Abas, et al.
Publicado: (2025)
KuBERT: Central Kurdish BERT Model and Its Application for Sentiment Analysis
por: Awlla, Kozhin muhealddin, et al.
Publicado: (2025)
por: Awlla, Kozhin muhealddin, et al.
Publicado: (2025)
TokenSelect: Efficient Long-Context Inference and Length Extrapolation for LLMs via Dynamic Token-Level KV Cache Selection
por: Wu, Wei, et al.
Publicado: (2024)
por: Wu, Wei, et al.
Publicado: (2024)
The Role of Orthographic Consistency in Multilingual Embedding Models for Text Classification in Arabic-Script Languages
por: Abdullah, Abdulhady Abas, et al.
Publicado: (2025)
por: Abdullah, Abdulhady Abas, et al.
Publicado: (2025)
Position as Probability: Self-Supervised Transformers that Think Past Their Training for Length Extrapolation
por: Lee, Philip Heejun
Publicado: (2025)
por: Lee, Philip Heejun
Publicado: (2025)
Evaluating Biases in Context-Dependent Health Questions
por: Levy, Sharon, et al.
Publicado: (2024)
por: Levy, Sharon, et al.
Publicado: (2024)
Softplus Attention with Re-weighting Boosts Length Extrapolation in Large Language Models
por: Gao, Bo, et al.
Publicado: (2025)
por: Gao, Bo, et al.
Publicado: (2025)
Enhancing Kurdish Text-to-Speech with Native Corpus Training: A High-Quality WaveGlow Vocoder Approach
por: Abdullah, Abdulhady Abas, et al.
Publicado: (2024)
por: Abdullah, Abdulhady Abas, et al.
Publicado: (2024)
A Training-Free Length Extrapolation Approach for LLMs: Greedy Attention Logit Interpolation (GALI)
por: Li, Yan, et al.
Publicado: (2025)
por: Li, Yan, et al.
Publicado: (2025)
Recall with Reasoning: Chain-of-Thought Distillation for Mamba's Long-Context Memory and Extrapolation
por: Ma, Junyu, et al.
Publicado: (2025)
por: Ma, Junyu, et al.
Publicado: (2025)
Two Stones Hit One Bird: Bilevel Positional Encoding for Better Length Extrapolation
por: He, Zhenyu, et al.
Publicado: (2024)
por: He, Zhenyu, et al.
Publicado: (2024)
Systematic Biases in LLM Simulations of Debates
por: Taubenfeld, Amir, et al.
Publicado: (2024)
por: Taubenfeld, Amir, et al.
Publicado: (2024)
Extrapolation Merging: Keep Improving With Extrapolation and Merging
por: Lin, Yiguan, et al.
Publicado: (2025)
por: Lin, Yiguan, et al.
Publicado: (2025)
Base of RoPE Bounds Context Length
por: Men, Xin, et al.
Publicado: (2024)
por: Men, Xin, et al.
Publicado: (2024)
Eliminating Biased Length Reliance of Direct Preference Optimization via Down-Sampled KL Divergence
por: Lu, Junru, et al.
Publicado: (2024)
por: Lu, Junru, et al.
Publicado: (2024)
InfLLM: Training-Free Long-Context Extrapolation for LLMs with an Efficient Context Memory
por: Xiao, Chaojun, et al.
Publicado: (2024)
por: Xiao, Chaojun, et al.
Publicado: (2024)
Revisiting Context Choices for Context-aware Machine Translation
por: Rikters, Matīss, et al.
Publicado: (2021)
por: Rikters, Matīss, et al.
Publicado: (2021)
Positional Biases Shift as Inputs Approach Context Window Limits
por: Veseli, Blerta, et al.
Publicado: (2025)
por: Veseli, Blerta, et al.
Publicado: (2025)
The Impact of Role Design in In-Context Learning for Large Language Models
por: Rouzegar, Hamidreza, et al.
Publicado: (2025)
por: Rouzegar, Hamidreza, et al.
Publicado: (2025)
Bootstrap Your Own Context Length
por: Wang, Liang, et al.
Publicado: (2024)
por: Wang, Liang, et al.
Publicado: (2024)
CoBia: Constructed Conversations Can Trigger Otherwise Concealed Societal Biases in LLMs
por: Nikeghbal, Nafiseh, et al.
Publicado: (2025)
por: Nikeghbal, Nafiseh, et al.
Publicado: (2025)
Ejemplares similares
-
Context-aware Rotary Position Embedding
por: Veisi, Ali, et al.
Publicado: (2025) -
How Language Models Prioritize Contextual Grammatical Cues?
por: Amirzadeh, Hamidreza, et al.
Publicado: (2024) -
data2lang2vec: Data Driven Typological Features Completion
por: Amirzadeh, Hamidreza, et al.
Publicado: (2024) -
In-Context Learning (and Unlearning) of Length Biases
por: Schoch, Stephanie, et al.
Publicado: (2025) -
ParallelComp: Parallel Long-Context Compressor for Length Extrapolation
por: Xiong, Jing, et al.
Publicado: (2025)