Why Prompt Optimization Works, and Why It Sometimes Doesn't: A Causal-Inspired Edit-Level Analysis
Fuente:
arXiv
Saved in:
| Main Authors: | Gong, Shuzhi, Wen, Hechuan |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Straight to Zero: Why Linearly Decaying the Learning Rate to Zero Works Best for LLMs
by: Bergsma, Shane, et al.
Published: (2025)
by: Bergsma, Shane, et al.
Published: (2025)
Why Fine-Tuning Encourages Hallucinations and How to Fix It
by: Kaplan, Guy, et al.
Published: (2026)
by: Kaplan, Guy, et al.
Published: (2026)
Language Models Learn Universal Representations of Numbers and Here's Why You Should Care
by: Štefánik, Michal, et al.
Published: (2025)
by: Štefánik, Michal, et al.
Published: (2025)
Why Flow Matching is Particle Swarm Optimization?
by: Ouyang, Kaichen
Published: (2025)
by: Ouyang, Kaichen
Published: (2025)
CAPO: Cost-Aware Prompt Optimization
by: Zehle, Tom, et al.
Published: (2025)
by: Zehle, Tom, et al.
Published: (2025)
Indian Wedding System Optimization (IWSO): A Novel Socially Inspired Metaheuristic with Operational Design and Analysis
by: Saxena, Deepika, et al.
Published: (2026)
by: Saxena, Deepika, et al.
Published: (2026)
An enhanced Teaching-Learning-Based Optimization (TLBO) with Grey Wolf Optimizer (GWO) for text feature selection and clustering
by: Azarshab, Mahsa, et al.
Published: (2024)
by: Azarshab, Mahsa, et al.
Published: (2024)
Evolutionary Multi-Objective Optimization of Large Language Model Prompts for Balancing Sentiments
by: Baumann, Jill, et al.
Published: (2024)
by: Baumann, Jill, et al.
Published: (2024)
Neural Exploratory Landscape Analysis for Meta-Black-Box-Optimization
by: Ma, Zeyuan, et al.
Published: (2024)
by: Ma, Zeyuan, et al.
Published: (2024)
Dataset Optimization for Chronic Disease Prediction with Bio-Inspired Feature Selection
by: Dyoub, Abeer, et al.
Published: (2023)
by: Dyoub, Abeer, et al.
Published: (2023)
Large Language Models Suffer From Their Own Output: An Analysis of the Self-Consuming Training Loop
by: Briesch, Martin, et al.
Published: (2023)
by: Briesch, Martin, et al.
Published: (2023)
What's the Magic Word? A Control Theory of LLM Prompting
by: Bhargava, Aman, et al.
Published: (2023)
by: Bhargava, Aman, et al.
Published: (2023)
Financial Decision Making using Reinforcement Learning with Dirichlet Priors and Quantum-Inspired Genetic Optimization
by: Nandy, Prasun, et al.
Published: (2025)
by: Nandy, Prasun, et al.
Published: (2025)
A Review of Neuroscience-Inspired Machine Learning
by: Ororbia, Alexander, et al.
Published: (2024)
by: Ororbia, Alexander, et al.
Published: (2024)
A Hormone-inspired Emotion Layer for Transformer language models (HELT)
by: Reda, Eslam, et al.
Published: (2026)
by: Reda, Eslam, et al.
Published: (2026)
Sorbet: A Neuromorphic Hardware-Compatible Transformer-Based Spiking Language Model
by: Tang, Kaiwen, et al.
Published: (2024)
by: Tang, Kaiwen, et al.
Published: (2024)
Bandit-Based Prompt Design Strategy Selection Improves Prompt Optimizers
by: Ashizawa, Rin, et al.
Published: (2025)
by: Ashizawa, Rin, et al.
Published: (2025)
Model Uncertainty in Evolutionary Optimization and Bayesian Optimization: A Comparative Analysis
by: Hao, Hao, et al.
Published: (2024)
by: Hao, Hao, et al.
Published: (2024)
Detect and Act: Automated Dynamic Optimizer through Meta-Black-Box Optimization
by: Gao, Zijian, et al.
Published: (2026)
by: Gao, Zijian, et al.
Published: (2026)
Empirical Analysis of Nature-Inspired Algorithms for Autism Spectrum Disorder Detection Using 3D Video Dataset
by: Panchal, Aneesh, et al.
Published: (2025)
by: Panchal, Aneesh, et al.
Published: (2025)
RUMC: A Rule-based Classifier Inspired by Evolutionary Methods
by: Mokhtari, Melvin
Published: (2024)
by: Mokhtari, Melvin
Published: (2024)
D-Score: A Synapse-Inspired Approach for Filter Pruning
by: Park, Doyoung, et al.
Published: (2023)
by: Park, Doyoung, et al.
Published: (2023)
Surrogate Learning in Meta-Black-Box Optimization: A Preliminary Study
by: Ma, Zeyuan, et al.
Published: (2025)
by: Ma, Zeyuan, et al.
Published: (2025)
Why "classic" Transformers are shallow and how to make them go deep
by: Yu, Yueyao, et al.
Published: (2023)
by: Yu, Yueyao, et al.
Published: (2023)
Decomposing Evolutionary Mixture-of-LoRA Architectures: The Routing Lever, the Lifecycle Penalty, and a Substrate-Conditional Boundary
by: Kumaresan, Ramchand
Published: (2026)
by: Kumaresan, Ramchand
Published: (2026)
EvoX: Meta-Evolution for Automated Discovery
by: Liu, Shu, et al.
Published: (2026)
by: Liu, Shu, et al.
Published: (2026)
EvolKV: Evolutionary KV Cache Compression for LLM Inference
by: Yu, Bohan, et al.
Published: (2025)
by: Yu, Bohan, et al.
Published: (2025)
An In-depth Walkthrough on Evolution of Neural Machine Translation
by: Jagtap, Rohan, et al.
Published: (2020)
by: Jagtap, Rohan, et al.
Published: (2020)
Pruner-Zero: Evolving Symbolic Pruning Metric from scratch for Large Language Models
by: Dong, Peijie, et al.
Published: (2024)
by: Dong, Peijie, et al.
Published: (2024)
SpikeLM: Towards General Spike-Driven Language Modeling via Elastic Bi-Spiking Mechanisms
by: Xing, Xingrun, et al.
Published: (2024)
by: Xing, Xingrun, et al.
Published: (2024)
Hysteresis Activation Function for Efficient Inference
by: Kimhi, Moshe, et al.
Published: (2024)
by: Kimhi, Moshe, et al.
Published: (2024)
Genetic Instruct: Scaling up Synthetic Generation of Coding Instructions for Large Language Models
by: Majumdar, Somshubra, et al.
Published: (2024)
by: Majumdar, Somshubra, et al.
Published: (2024)
Pre-trained Language Models Learn Remarkably Accurate Representations of Numbers
by: Kadlčík, Marek, et al.
Published: (2025)
by: Kadlčík, Marek, et al.
Published: (2025)
Large Language Models for Tuning Evolution Strategies
by: Kramer, Oliver
Published: (2024)
by: Kramer, Oliver
Published: (2024)
AP-BMM: Approximating Capability-Cost Pareto Sets of LLMs via Asynchronous Prior-Guided Bayesian Model Merging
by: Chen, Kesheng, et al.
Published: (2025)
by: Chen, Kesheng, et al.
Published: (2025)
On the Power of Convolution Augmented Transformer
by: Li, Mingchen, et al.
Published: (2024)
by: Li, Mingchen, et al.
Published: (2024)
HAT: Hardware-Aware Transformers for Efficient Natural Language Processing
by: Wang, Hanrui, et al.
Published: (2020)
by: Wang, Hanrui, et al.
Published: (2020)
SpikingSSMs: Learning Long Sequences with Sparse and Parallel Spiking State Space Models
by: Shen, Shuaijie, et al.
Published: (2024)
by: Shen, Shuaijie, et al.
Published: (2024)
SwitchHead: Accelerating Transformers with Mixture-of-Experts Attention
by: Csordás, Róbert, et al.
Published: (2023)
by: Csordás, Róbert, et al.
Published: (2023)
SpikeLLM: Scaling up Spiking Neural Network to Large Language Models via Saliency-based Spiking
by: Xing, Xingrun, et al.
Published: (2024)
by: Xing, Xingrun, et al.
Published: (2024)
Similar Items
-
Straight to Zero: Why Linearly Decaying the Learning Rate to Zero Works Best for LLMs
by: Bergsma, Shane, et al.
Published: (2025) -
Why Fine-Tuning Encourages Hallucinations and How to Fix It
by: Kaplan, Guy, et al.
Published: (2026) -
Language Models Learn Universal Representations of Numbers and Here's Why You Should Care
by: Štefánik, Michal, et al.
Published: (2025) -
Why Flow Matching is Particle Swarm Optimization?
by: Ouyang, Kaichen
Published: (2025) -
CAPO: Cost-Aware Prompt Optimization
by: Zehle, Tom, et al.
Published: (2025)