Position Engineering: Boosting Large Language Models through Positional Information Manipulation
Fuente:
arXiv
Salvato in:
| Autori principali: | He, Zhiyuan, Jiang, Huiqiang, Wang, Zilong, Yang, Yuqing, Qiu, Luna, Qiu, Lili |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
LeanK: Learnable K Cache Channel Pruning for Efficient Decoding
di: Zhang, Yike, et al.
Pubblicazione: (2025)
di: Zhang, Yike, et al.
Pubblicazione: (2025)
Retrieval Augmented Generation (RAG) and Beyond: A Comprehensive Survey on How to Make your LLMs use External Data More Wisely
di: Zhao, Siyun, et al.
Pubblicazione: (2024)
di: Zhao, Siyun, et al.
Pubblicazione: (2024)
Mitigate Position Bias in Large Language Models via Scaling a Single Dimension
di: Yu, Yijiong, et al.
Pubblicazione: (2024)
di: Yu, Yijiong, et al.
Pubblicazione: (2024)
VL Norm: Rethink Loss Aggregation in RLVR
di: He, Zhiyuan, et al.
Pubblicazione: (2025)
di: He, Zhiyuan, et al.
Pubblicazione: (2025)
SortedRL: Accelerating RL Training for LLMs through Online Length-Aware Scheduling
di: Zhang, Yiqi, et al.
Pubblicazione: (2026)
di: Zhang, Yiqi, et al.
Pubblicazione: (2026)
Agent Lightning: Train ANY AI Agents with Reinforcement Learning
di: Luo, Xufang, et al.
Pubblicazione: (2025)
di: Luo, Xufang, et al.
Pubblicazione: (2025)
Position-Aware Parameter Efficient Fine-Tuning Approach for Reducing Positional Bias in LLMs
di: Zhang, Zheng, et al.
Pubblicazione: (2024)
di: Zhang, Zheng, et al.
Pubblicazione: (2024)
PoPE: Legendre Orthogonal Polynomials Based Position Encoding for Large Language Models
di: Aggarwal, Arpit
Pubblicazione: (2024)
di: Aggarwal, Arpit
Pubblicazione: (2024)
EnJa: Ensemble Jailbreak on Large Language Models
di: Zhang, Jiahao, et al.
Pubblicazione: (2024)
di: Zhang, Jiahao, et al.
Pubblicazione: (2024)
RePo: Language Models with Context Re-Positioning
di: Li, Huayang, et al.
Pubblicazione: (2025)
di: Li, Huayang, et al.
Pubblicazione: (2025)
Boosting Large Language Models with Mask Fine-Tuning
di: Zhang, Mingyuan, et al.
Pubblicazione: (2025)
di: Zhang, Mingyuan, et al.
Pubblicazione: (2025)
Self-Supervised Position Debiasing for Large Language Models
di: Liu, Zhongkun, et al.
Pubblicazione: (2024)
di: Liu, Zhongkun, et al.
Pubblicazione: (2024)
UltraFeedback: Boosting Language Models with Scaled AI Feedback
di: Cui, Ganqu, et al.
Pubblicazione: (2023)
di: Cui, Ganqu, et al.
Pubblicazione: (2023)
Position: Uncertainty Quantification Needs Reassessment for Large-language Model Agents
di: Kirchhof, Michael, et al.
Pubblicazione: (2025)
di: Kirchhof, Michael, et al.
Pubblicazione: (2025)
Aligning Large Language Models via Fine-grained Supervision
di: Xu, Dehong, et al.
Pubblicazione: (2024)
di: Xu, Dehong, et al.
Pubblicazione: (2024)
Enhancing Quantitative Reasoning Skills of Large Language Models through Dimension Perception
di: Huang, Yuncheng, et al.
Pubblicazione: (2023)
di: Huang, Yuncheng, et al.
Pubblicazione: (2023)
Two Stones Hit One Bird: Bilevel Positional Encoding for Better Length Extrapolation
di: He, Zhenyu, et al.
Pubblicazione: (2024)
di: He, Zhenyu, et al.
Pubblicazione: (2024)
E-Sparse: Boosting the Large Language Model Inference through Entropy-based N:M Sparsity
di: Li, Yun, et al.
Pubblicazione: (2023)
di: Li, Yun, et al.
Pubblicazione: (2023)
Uncertainty of Thoughts: Uncertainty-Aware Planning Enhances Information Seeking in Large Language Models
di: Hu, Zhiyuan, et al.
Pubblicazione: (2024)
di: Hu, Zhiyuan, et al.
Pubblicazione: (2024)
BoostStep: Boosting mathematical capability of Large Language Models via improved single-step reasoning
di: Zhang, Beichen, et al.
Pubblicazione: (2025)
di: Zhang, Beichen, et al.
Pubblicazione: (2025)
Position: Scaling LLM Agents Requires Asymptotic Analysis with LLM Primitives
di: Meyerson, Elliot, et al.
Pubblicazione: (2025)
di: Meyerson, Elliot, et al.
Pubblicazione: (2025)
Athena: Efficient Block-Wise Post-Training Quantization for Large Language Models Using Second-Order Matrix Derivative Information
di: Wang, Yanshu, et al.
Pubblicazione: (2024)
di: Wang, Yanshu, et al.
Pubblicazione: (2024)
Do Not Let Low-Probability Tokens Over-Dominate in RL for LLMs
di: Yang, Zhihe, et al.
Pubblicazione: (2025)
di: Yang, Zhihe, et al.
Pubblicazione: (2025)
Stepwise Self-Consistent Mathematical Reasoning with Large Language Models
di: Zhao, Zilong, et al.
Pubblicazione: (2024)
di: Zhao, Zilong, et al.
Pubblicazione: (2024)
Explicit Multi-head Attention for Inter-head Interaction in Large Language Models
di: Peng, Runyu, et al.
Pubblicazione: (2026)
di: Peng, Runyu, et al.
Pubblicazione: (2026)
Spectral Editing of Activations for Large Language Model Alignment
di: Qiu, Yifu, et al.
Pubblicazione: (2024)
di: Qiu, Yifu, et al.
Pubblicazione: (2024)
ROPO: Robust Preference Optimization for Large Language Models
di: Liang, Xize, et al.
Pubblicazione: (2024)
di: Liang, Xize, et al.
Pubblicazione: (2024)
Instruction Following by Principled Boosting Attention of Large Language Models
di: Guardieiro, Vitoria, et al.
Pubblicazione: (2025)
di: Guardieiro, Vitoria, et al.
Pubblicazione: (2025)
Segment Policy Optimization: Effective Segment-Level Credit Assignment in RL for Large Language Models
di: Guo, Yiran, et al.
Pubblicazione: (2025)
di: Guo, Yiran, et al.
Pubblicazione: (2025)
Group Representational Position Encoding
di: Zhang, Yifan, et al.
Pubblicazione: (2025)
di: Zhang, Yifan, et al.
Pubblicazione: (2025)
Revisiting the Test-Time Scaling of o1-like Models: Do they Truly Possess Test-Time Scaling Capabilities?
di: Zeng, Zhiyuan, et al.
Pubblicazione: (2025)
di: Zeng, Zhiyuan, et al.
Pubblicazione: (2025)
SConU: Selective Conformal Uncertainty in Large Language Models
di: Wang, Zhiyuan, et al.
Pubblicazione: (2025)
di: Wang, Zhiyuan, et al.
Pubblicazione: (2025)
Reinforcement Learning for Reasoning in Large Language Models with One Training Example
di: Wang, Yiping, et al.
Pubblicazione: (2025)
di: Wang, Yiping, et al.
Pubblicazione: (2025)
More Thinking, More Bias: Length-Driven Position Bias in Reasoning Models
di: Wang, Xiao
Pubblicazione: (2026)
di: Wang, Xiao
Pubblicazione: (2026)
Boosting of Thoughts: Trial-and-Error Problem Solving with Large Language Models
di: Chen, Sijia, et al.
Pubblicazione: (2024)
di: Chen, Sijia, et al.
Pubblicazione: (2024)
dLLM-Cache: Accelerating Diffusion Large Language Models with Adaptive Caching
di: Liu, Zhiyuan, et al.
Pubblicazione: (2025)
di: Liu, Zhiyuan, et al.
Pubblicazione: (2025)
Position: The Turing-Completeness of Autoregressive Transformers Relies Heavily on Context Management
di: Cui, Guanyu, et al.
Pubblicazione: (2026)
di: Cui, Guanyu, et al.
Pubblicazione: (2026)
PassNet: Scaling Large Language Models for Graph Compiler Pass Generation
di: Liu, Yiqun, et al.
Pubblicazione: (2026)
di: Liu, Yiqun, et al.
Pubblicazione: (2026)
Rethinking On-Policy Distillation of Large Language Models: Phenomenology, Mechanism, and Recipe
di: Li, Yaxuan, et al.
Pubblicazione: (2026)
di: Li, Yaxuan, et al.
Pubblicazione: (2026)
Evaluating the Formal Reasoning Capabilities of Large Language Models through Chomsky Hierarchy
di: Dong, Yihong, et al.
Pubblicazione: (2026)
di: Dong, Yihong, et al.
Pubblicazione: (2026)
Documenti analoghi
-
LeanK: Learnable K Cache Channel Pruning for Efficient Decoding
di: Zhang, Yike, et al.
Pubblicazione: (2025) -
Retrieval Augmented Generation (RAG) and Beyond: A Comprehensive Survey on How to Make your LLMs use External Data More Wisely
di: Zhao, Siyun, et al.
Pubblicazione: (2024) -
Mitigate Position Bias in Large Language Models via Scaling a Single Dimension
di: Yu, Yijiong, et al.
Pubblicazione: (2024) -
VL Norm: Rethink Loss Aggregation in RLVR
di: He, Zhiyuan, et al.
Pubblicazione: (2025) -
SortedRL: Accelerating RL Training for LLMs through Online Length-Aware Scheduling
di: Zhang, Yiqi, et al.
Pubblicazione: (2026)