The Order Effect: Investigating Prompt Sensitivity to Input Order in LLMs
Fuente:
arXiv
Salvato in:
| Autori principali: | Guan, Bryan, Roosta, Tanya, Passban, Peyman, Rezagholizadeh, Mehdi |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
How Well Do LLMs Represent Values Across Cultures? Empirical Analysis of LLM Responses Based on Hofstede Cultural Dimensions
di: Kharchenko, Julia, et al.
Pubblicazione: (2024)
di: Kharchenko, Julia, et al.
Pubblicazione: (2024)
How Order-Sensitive Are LLMs? OrderProbe for Deterministic Structural Reconstruction
di: He, Yingjie, et al.
Pubblicazione: (2026)
di: He, Yingjie, et al.
Pubblicazione: (2026)
Towards Practical Tool Usage for Continually Learning LLMs
di: Huang, Jerry, et al.
Pubblicazione: (2024)
di: Huang, Jerry, et al.
Pubblicazione: (2024)
Do Robot Snakes Dream like Electric Sheep? Investigating the Effects of Architectural Inductive Biases on Hallucination
di: Huang, Jerry, et al.
Pubblicazione: (2024)
di: Huang, Jerry, et al.
Pubblicazione: (2024)
I Think, Therefore I Am Under-Qualified? A Benchmark for Evaluating Linguistic Shibboleth Detection in LLM Hiring Evaluations
di: Kharchenko, Julia, et al.
Pubblicazione: (2025)
di: Kharchenko, Julia, et al.
Pubblicazione: (2025)
Batch-Max: Higher LLM Throughput using Larger Batch Sizes and KV Cache Compression
di: Metel, Michael R., et al.
Pubblicazione: (2024)
di: Metel, Michael R., et al.
Pubblicazione: (2024)
Synthetic Users, Real Differences: an Evaluation Framework for User Simulation in Multi-Turn Conversations
di: Liu, Yu Lu, et al.
Pubblicazione: (2026)
di: Liu, Yu Lu, et al.
Pubblicazione: (2026)
On the importance of Data Scale in Pretraining Arabic Language Models
di: Ghaddar, Abbas, et al.
Pubblicazione: (2024)
di: Ghaddar, Abbas, et al.
Pubblicazione: (2024)
ReGLA: Refining Gated Linear Attention
di: Lu, Peng, et al.
Pubblicazione: (2025)
di: Lu, Peng, et al.
Pubblicazione: (2025)
X-EcoMLA: Upcycling Pre-Trained Attention into MLA for Efficient and Extreme KV Compression
di: Li, Guihong, et al.
Pubblicazione: (2025)
di: Li, Guihong, et al.
Pubblicazione: (2025)
On the Sensitivity of Instruction-tuned LLMs to Harmful Sentences in Long Inputs
di: Ghorbanpour, Faeze, et al.
Pubblicazione: (2025)
di: Ghorbanpour, Faeze, et al.
Pubblicazione: (2025)
Input Order Shapes LLM Semantic Alignment in Multi-Document Summarization
di: Ma, Jing
Pubblicazione: (2025)
di: Ma, Jing
Pubblicazione: (2025)
Leveraging Self-Attention for Input-Dependent Soft Prompting in LLMs
di: Muppidi, Ananth, et al.
Pubblicazione: (2025)
di: Muppidi, Ananth, et al.
Pubblicazione: (2025)
Order Matters: Rethinking Prompt Construction in In-Context Learning
di: Li, Warren, et al.
Pubblicazione: (2025)
di: Li, Warren, et al.
Pubblicazione: (2025)
Draft on the Fly: Adaptive Self-Speculative Decoding using Cosine Similarity
di: Metel, Michael R., et al.
Pubblicazione: (2024)
di: Metel, Michael R., et al.
Pubblicazione: (2024)
Order Matters in Hallucination: Reasoning Order as Benchmark and Reflexive Prompting for Large-Language-Models
di: Xie, Zikai
Pubblicazione: (2024)
di: Xie, Zikai
Pubblicazione: (2024)
Order Doesn't Matter, But Reasoning Does: Training LLMs with Order-Centric Augmentation
di: He, Qianxi, et al.
Pubblicazione: (2025)
di: He, Qianxi, et al.
Pubblicazione: (2025)
Zebra-Llama: Towards Extremely Efficient Hybrid Models
di: Yang, Mingyu, et al.
Pubblicazione: (2025)
di: Yang, Mingyu, et al.
Pubblicazione: (2025)
Resonance RoPE: Improving Context Length Generalization of Large Language Models
di: Wang, Suyuchen, et al.
Pubblicazione: (2024)
di: Wang, Suyuchen, et al.
Pubblicazione: (2024)
LABO: Towards Learning Optimal Label Regularization via Bi-level Optimization
di: Lu, Peng, et al.
Pubblicazione: (2023)
di: Lu, Peng, et al.
Pubblicazione: (2023)
ProSA: Assessing and Understanding the Prompt Sensitivity of LLMs
di: Zhuo, Jingming, et al.
Pubblicazione: (2024)
di: Zhuo, Jingming, et al.
Pubblicazione: (2024)
RoToR: Towards More Reliable Responses for Order-Invariant Inputs
di: Yoon, Soyoung, et al.
Pubblicazione: (2025)
di: Yoon, Soyoung, et al.
Pubblicazione: (2025)
From 0-Order Selection to 2-Order Judgment: Combinatorial Hardening Exposes Compositional Failures in Frontier LLMs
di: Liu, Hanmeng, et al.
Pubblicazione: (2026)
di: Liu, Hanmeng, et al.
Pubblicazione: (2026)
Addressing Order Sensitivity of In-Context Demonstration Examples in Causal Language Models
di: Xiang, Yanzheng, et al.
Pubblicazione: (2024)
di: Xiang, Yanzheng, et al.
Pubblicazione: (2024)
iAgentBench: Benchmarking Sensemaking Capabilities of Information-Seeking Agents on High-Traffic Topics
di: Dammu, Preetam Prabhu Srikar, et al.
Pubblicazione: (2026)
di: Dammu, Preetam Prabhu Srikar, et al.
Pubblicazione: (2026)
OTTAWA: Optimal TransporT Adaptive Word Aligner for Hallucination and Omission Translation Errors Detection
di: Huang, Chenyang, et al.
Pubblicazione: (2024)
di: Huang, Chenyang, et al.
Pubblicazione: (2024)
CHARP: Conversation History AwaReness Probing for Knowledge-grounded Dialogue Systems
di: Ghaddar, Abbas, et al.
Pubblicazione: (2024)
di: Ghaddar, Abbas, et al.
Pubblicazione: (2024)
Recognition Without Authorization: LLMs and the Moral Order of Online Advice
di: van Nuenen, Tom
Pubblicazione: (2026)
di: van Nuenen, Tom
Pubblicazione: (2026)
Evaluating Prompting Strategies with MedGemma for Medical Order Extraction
di: Balachandran, Abhinand, et al.
Pubblicazione: (2025)
di: Balachandran, Abhinand, et al.
Pubblicazione: (2025)
Intuitive or Dependent? Investigating LLMs' Behavior Style to Conflicting Prompts
di: Ying, Jiahao, et al.
Pubblicazione: (2023)
di: Ying, Jiahao, et al.
Pubblicazione: (2023)
Beyond the Limits: A Survey of Techniques to Extend the Context Length in Large Language Models
di: Wang, Xindi, et al.
Pubblicazione: (2024)
di: Wang, Xindi, et al.
Pubblicazione: (2024)
Sorted LLaMA: Unlocking the Potential of Intermediate Layers of Large Language Models for Dynamic Inference
di: Kavehzadeh, Parsa, et al.
Pubblicazione: (2023)
di: Kavehzadeh, Parsa, et al.
Pubblicazione: (2023)
Higher-Order Asynchronous Effects
di: Ahman, Danel, et al.
Pubblicazione: (2023)
di: Ahman, Danel, et al.
Pubblicazione: (2023)
Unveiling the Lexical Sensitivity of LLMs: Combinatorial Optimization for Prompt Enhancement
di: Zhan, Pengwei, et al.
Pubblicazione: (2024)
di: Zhan, Pengwei, et al.
Pubblicazione: (2024)
RealWebAssist: A Benchmark for Long-Horizon Web Assistance with Real-World Users
di: Ye, Suyu, et al.
Pubblicazione: (2025)
di: Ye, Suyu, et al.
Pubblicazione: (2025)
When Domains Interact: Asymmetric and Order-Sensitive Cross-Domain Effects in Reinforcement Learning for Reasoning
di: Yang, Wang, et al.
Pubblicazione: (2026)
di: Yang, Wang, et al.
Pubblicazione: (2026)
Hearing the Order: Investigating Position Bias in Large Audio-Language Models
di: Lin, Yu-Xiang, et al.
Pubblicazione: (2025)
di: Lin, Yu-Xiang, et al.
Pubblicazione: (2025)
A Typologically Grounded Evaluation Framework for Word Order and Morphology Sensitivity in Multilingual Masked LMs
di: Feldman, Anna, et al.
Pubblicazione: (2026)
di: Feldman, Anna, et al.
Pubblicazione: (2026)
Understanding the Prompt Sensitivity
di: Liu, Yang, et al.
Pubblicazione: (2026)
di: Liu, Yang, et al.
Pubblicazione: (2026)
Exploring the Sensitivity of LLMs' Decision-Making Capabilities: Insights from Prompt Variation and Hyperparameters
di: Loya, Manikanta, et al.
Pubblicazione: (2023)
di: Loya, Manikanta, et al.
Pubblicazione: (2023)
Documenti analoghi
-
How Well Do LLMs Represent Values Across Cultures? Empirical Analysis of LLM Responses Based on Hofstede Cultural Dimensions
di: Kharchenko, Julia, et al.
Pubblicazione: (2024) -
How Order-Sensitive Are LLMs? OrderProbe for Deterministic Structural Reconstruction
di: He, Yingjie, et al.
Pubblicazione: (2026) -
Towards Practical Tool Usage for Continually Learning LLMs
di: Huang, Jerry, et al.
Pubblicazione: (2024) -
Do Robot Snakes Dream like Electric Sheep? Investigating the Effects of Architectural Inductive Biases on Hallucination
di: Huang, Jerry, et al.
Pubblicazione: (2024) -
I Think, Therefore I Am Under-Qualified? A Benchmark for Evaluating Linguistic Shibboleth Detection in LLM Hiring Evaluations
di: Kharchenko, Julia, et al.
Pubblicazione: (2025)