Mitigate Position Bias in Large Language Models via Scaling a Single Dimension
Fuente:
arXiv
Salvato in:
| Autori principali: | Yu, Yijiong, Jiang, Huiqiang, Luo, Xufang, Wu, Qianhui, Lin, Chin-Yew, Li, Dongsheng, Yang, Yuqing, Huang, Yongfeng, Qiu, Lili |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
LongLLMLingua: Accelerating and Enhancing LLMs in Long Context Scenarios via Prompt Compression
di: Jiang, Huiqiang, et al.
Pubblicazione: (2023)
di: Jiang, Huiqiang, et al.
Pubblicazione: (2023)
On Memory Construction and Retrieval for Personalized Conversational Agents
di: Pan, Zhuoshi, et al.
Pubblicazione: (2025)
di: Pan, Zhuoshi, et al.
Pubblicazione: (2025)
MInference 1.0: Accelerating Pre-filling for Long-Context LLMs via Dynamic Sparse Attention
di: Jiang, Huiqiang, et al.
Pubblicazione: (2024)
di: Jiang, Huiqiang, et al.
Pubblicazione: (2024)
Position Engineering: Boosting Large Language Models through Positional Information Manipulation
di: He, Zhiyuan, et al.
Pubblicazione: (2024)
di: He, Zhiyuan, et al.
Pubblicazione: (2024)
MMInference: Accelerating Pre-filling for Long-Context VLMs via Modality-Aware Permutation Sparse Attention
di: Li, Yucheng, et al.
Pubblicazione: (2025)
di: Li, Yucheng, et al.
Pubblicazione: (2025)
LLMLingua-2: Data Distillation for Efficient and Faithful Task-Agnostic Prompt Compression
di: Pan, Zhuoshi, et al.
Pubblicazione: (2024)
di: Pan, Zhuoshi, et al.
Pubblicazione: (2024)
SCBench: A KV Cache-Centric Analysis of Long-Context Methods
di: Li, Yucheng, et al.
Pubblicazione: (2024)
di: Li, Yucheng, et al.
Pubblicazione: (2024)
Mitigating Hallucinations in Large Vision-Language Models via DPO: On-Policy Data Hold the Key
di: Yang, Zhihe, et al.
Pubblicazione: (2025)
di: Yang, Zhihe, et al.
Pubblicazione: (2025)
SortedRL: Accelerating RL Training for LLMs through Online Length-Aware Scheduling
di: Zhang, Yiqi, et al.
Pubblicazione: (2026)
di: Zhang, Yiqi, et al.
Pubblicazione: (2026)
Designing Network Algorithms via Large Language Models
di: He, Zhiyuan, et al.
Pubblicazione: (2024)
di: He, Zhiyuan, et al.
Pubblicazione: (2024)
LLM-RadJudge: Achieving Radiologist-Level Evaluation for X-Ray Report Generation
di: Wang, Zilong, et al.
Pubblicazione: (2024)
di: Wang, Zilong, et al.
Pubblicazione: (2024)
Unified Medical Image Pre-training in Language-Guided Common Semantic Space
di: He, Xiaoxuan, et al.
Pubblicazione: (2023)
di: He, Xiaoxuan, et al.
Pubblicazione: (2023)
Multi-Persona Thinking for Bias Mitigation in Large Language Models
di: Chen, Yuxing, et al.
Pubblicazione: (2026)
di: Chen, Yuxing, et al.
Pubblicazione: (2026)
VL Norm: Rethink Loss Aggregation in RLVR
di: He, Zhiyuan, et al.
Pubblicazione: (2025)
di: He, Zhiyuan, et al.
Pubblicazione: (2025)
Exploratory Memory-Augmented LLM Agent via Hybrid On- and Off-Policy Optimization
di: Liu, Zeyuan, et al.
Pubblicazione: (2026)
di: Liu, Zeyuan, et al.
Pubblicazione: (2026)
Accelerating Prefilling via Decoding-time Contribution Sparsity
di: He, Zhiyuan, et al.
Pubblicazione: (2025)
di: He, Zhiyuan, et al.
Pubblicazione: (2025)
Patience Is The Key to Large Language Model Reasoning
di: Yu, Yijiong
Pubblicazione: (2024)
di: Yu, Yijiong
Pubblicazione: (2024)
VisRL: Intention-Driven Visual Perception via Reinforced Reasoning
di: Chen, Zhangquan, et al.
Pubblicazione: (2025)
di: Chen, Zhangquan, et al.
Pubblicazione: (2025)
SecurityLingua: Efficient Defense of LLM Jailbreak Attacks via Security-Aware Prompt Compression
di: Li, Yucheng, et al.
Pubblicazione: (2025)
di: Li, Yucheng, et al.
Pubblicazione: (2025)
An Effective Framework to Help Large Language Models Handle Numeric-involved Long-context Tasks
di: Yu, Yijiong
Pubblicazione: (2024)
di: Yu, Yijiong
Pubblicazione: (2024)
A Large-scale Medical Visual Task Adaptation Benchmark
di: Mo, Shentong, et al.
Pubblicazione: (2024)
di: Mo, Shentong, et al.
Pubblicazione: (2024)
MTraining: Distributed Dynamic Sparse Attention for Efficient Ultra-Long Context Training
di: Li, Wenxuan, et al.
Pubblicazione: (2025)
di: Li, Wenxuan, et al.
Pubblicazione: (2025)
pMoE: Prompting Diverse Experts Together Wins More in Visual Adaptation
di: Mo, Shentong, et al.
Pubblicazione: (2026)
di: Mo, Shentong, et al.
Pubblicazione: (2026)
Training With "Paraphrasing the Original Text" Teaches LLM to Better Retrieve in Long-context Tasks
di: Yu, Yijiong, et al.
Pubblicazione: (2023)
di: Yu, Yijiong, et al.
Pubblicazione: (2023)
Understanding Reasoning in LLMs through Strategic Information Allocation under Uncertainty
di: Kim, Jeonghye, et al.
Pubblicazione: (2026)
di: Kim, Jeonghye, et al.
Pubblicazione: (2026)
Accelerate Parallelizable Reasoning via Parallel Decoding within One Sequence
di: Yu, Yijiong
Pubblicazione: (2025)
di: Yu, Yijiong
Pubblicazione: (2025)
DesignProbe: A Graphic Design Benchmark for Multimodal Large Language Models
di: Lin, Jieru, et al.
Pubblicazione: (2024)
di: Lin, Jieru, et al.
Pubblicazione: (2024)
LeanK: Learnable K Cache Channel Pruning for Efficient Decoding
di: Zhang, Yike, et al.
Pubblicazione: (2025)
di: Zhang, Yike, et al.
Pubblicazione: (2025)
Investigating Thinking Behaviours of Reasoning-Based Language Models for Social Bias Mitigation
di: Luo, Guoqing, et al.
Pubblicazione: (2025)
di: Luo, Guoqing, et al.
Pubblicazione: (2025)
Position Bias Mitigates Position Bias:Mitigate Position Bias Through Inter-Position Knowledge Distillation
di: Wang, Yifei, et al.
Pubblicazione: (2025)
di: Wang, Yifei, et al.
Pubblicazione: (2025)
LSPT: Long-term Spatial Prompt Tuning for Visual Representation Learning
di: Mo, Shentong, et al.
Pubblicazione: (2024)
di: Mo, Shentong, et al.
Pubblicazione: (2024)
A Comprehensive Information-Decomposition Analysis of Large Vision-Language Models
di: Xiu, Lixin, et al.
Pubblicazione: (2026)
di: Xiu, Lixin, et al.
Pubblicazione: (2026)
Agent Lightning: Train ANY AI Agents with Reinforcement Learning
di: Luo, Xufang, et al.
Pubblicazione: (2025)
di: Luo, Xufang, et al.
Pubblicazione: (2025)
CogBias: Measuring and Mitigating Cognitive Bias in Large Language Models
di: Huang, Fan, et al.
Pubblicazione: (2026)
di: Huang, Fan, et al.
Pubblicazione: (2026)
Unveiling and Mitigating Bias in Mental Health Analysis with Large Language Models
di: Wang, Yuqing, et al.
Pubblicazione: (2024)
di: Wang, Yuqing, et al.
Pubblicazione: (2024)
Long-context Language Models Fail in Basic Retrieval Tasks Without Sufficient Reasoning Steps
di: Yu, Yijiong, et al.
Pubblicazione: (2024)
di: Yu, Yijiong, et al.
Pubblicazione: (2024)
Bias in Large Language Models: Origin, Evaluation, and Mitigation
di: Guo, Yufei, et al.
Pubblicazione: (2024)
di: Guo, Yufei, et al.
Pubblicazione: (2024)
Do LLMs Really Think Step-by-step In Implicit Reasoning?
di: Yu, Yijiong
Pubblicazione: (2024)
di: Yu, Yijiong
Pubblicazione: (2024)
Challenges of COVID‐19 Vaccination Effectiveness in Individuals with Alzheimer's Disease and Related Dementias
di: Yijiong Yang
Pubblicazione: (2025)
di: Yijiong Yang
Pubblicazione: (2025)
Comparing overall medical costs of operative versus nonoperative treatment for femoral neck fractures among Alzheimer’s disease patients: a retrospective cohort study
di: Yijiong Yang
Pubblicazione: (2024)
di: Yijiong Yang
Pubblicazione: (2024)
Documenti analoghi
-
LongLLMLingua: Accelerating and Enhancing LLMs in Long Context Scenarios via Prompt Compression
di: Jiang, Huiqiang, et al.
Pubblicazione: (2023) -
On Memory Construction and Retrieval for Personalized Conversational Agents
di: Pan, Zhuoshi, et al.
Pubblicazione: (2025) -
MInference 1.0: Accelerating Pre-filling for Long-Context LLMs via Dynamic Sparse Attention
di: Jiang, Huiqiang, et al.
Pubblicazione: (2024) -
Position Engineering: Boosting Large Language Models through Positional Information Manipulation
di: He, Zhiyuan, et al.
Pubblicazione: (2024) -
MMInference: Accelerating Pre-filling for Long-Context VLMs via Modality-Aware Permutation Sparse Attention
di: Li, Yucheng, et al.
Pubblicazione: (2025)