Rethinking Prompt Optimization: Reinforcement, Diversification, and Migration in Blackbox LLMs
Fuente:
arXiv
Salvato in:
| Autori principali: | Davari, MohammadReza, Garg, Utkarsh, Cai, Weixin, Belilovsky, Eugene |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Model Breadcrumbs: Scaling Multi-Task Model Merging with Sparse Masks
di: Davari, MohammadReza, et al.
Pubblicazione: (2023)
di: Davari, MohammadReza, et al.
Pubblicazione: (2023)
FairDropout: Using Example-Tied Dropout to Enhance Generalization of Minority Groups
di: Nanfack, Geraldin, et al.
Pubblicazione: (2025)
di: Nanfack, Geraldin, et al.
Pubblicazione: (2025)
Understanding and Exploiting Weight Update Sparsity for Communication-Efficient Distributed RL
di: Miahi, Erfan, et al.
Pubblicazione: (2026)
di: Miahi, Erfan, et al.
Pubblicazione: (2026)
Celo2: Towards Learned Optimization Free Lunch
di: Moudgil, Abhinav, et al.
Pubblicazione: (2026)
di: Moudgil, Abhinav, et al.
Pubblicazione: (2026)
Improving Structural Diversity of Blackbox LLMs via Chain-of-Specification Prompting
di: Young, Halley, et al.
Pubblicazione: (2024)
di: Young, Halley, et al.
Pubblicazione: (2024)
Celo: Training Versatile Learned Optimizers on a Compute Diet
di: Moudgil, Abhinav, et al.
Pubblicazione: (2025)
di: Moudgil, Abhinav, et al.
Pubblicazione: (2025)
Heterogeneous Low-Bandwidth Pre-Training of LLMs
di: Obeidi, Yazan, et al.
Pubblicazione: (2026)
di: Obeidi, Yazan, et al.
Pubblicazione: (2026)
When Data Falls Short: Grokking Below the Critical Threshold
di: Singh, Vaibhav, et al.
Pubblicazione: (2025)
di: Singh, Vaibhav, et al.
Pubblicazione: (2025)
Stabilizing Native Low-Rank LLM Pretraining
di: Janson, Paul, et al.
Pubblicazione: (2026)
di: Janson, Paul, et al.
Pubblicazione: (2026)
Dual-Phase Continual Learning: Supervised Adaptation Meets Unsupervised Retention
di: Singh, Vaibhav, et al.
Pubblicazione: (2024)
di: Singh, Vaibhav, et al.
Pubblicazione: (2024)
Optimizing Rank-based Metrics with Blackbox Differentiation
di: Rolínek, Michal, et al.
Pubblicazione: (2019)
di: Rolínek, Michal, et al.
Pubblicazione: (2019)
Meta-learning Optimizers for Communication-Efficient Learning
di: Joseph, Charles-Étienne, et al.
Pubblicazione: (2023)
di: Joseph, Charles-Étienne, et al.
Pubblicazione: (2023)
Whispering to a Blackbox: Bootstrapping Frozen OCR with Visual Prompts
di: Samandarov, Samandar, et al.
Pubblicazione: (2026)
di: Samandarov, Samandar, et al.
Pubblicazione: (2026)
Warming Up for Zeroth-Order Federated Pre-Training with Low Resource Clients
di: Legate, Gwen, et al.
Pubblicazione: (2025)
di: Legate, Gwen, et al.
Pubblicazione: (2025)
Efficient Refusal Ablation in LLM through Optimal Transport
di: Nanfack, Geraldin, et al.
Pubblicazione: (2026)
di: Nanfack, Geraldin, et al.
Pubblicazione: (2026)
Preference-Optimized Pareto Set Learning for Blackbox Optimization
di: Haishan, Zhang, et al.
Pubblicazione: (2024)
di: Haishan, Zhang, et al.
Pubblicazione: (2024)
AVATAR: Adversarial Autoencoders with Autoregressive Refinement for Time Series Generation
di: EskandariNasab, MohammadReza, et al.
Pubblicazione: (2025)
di: EskandariNasab, MohammadReza, et al.
Pubblicazione: (2025)
SeriesGAN: Time Series Generation via Adversarial and Autoregressive Learning
di: EskandariNasab, MohammadReza, et al.
Pubblicazione: (2024)
di: EskandariNasab, MohammadReza, et al.
Pubblicazione: (2024)
ChronoGAN: Supervised and Embedded Generative Adversarial Networks for Time Series Generation
di: EskandariNasab, MohammadReza, et al.
Pubblicazione: (2024)
di: EskandariNasab, MohammadReza, et al.
Pubblicazione: (2024)
PyLO: Towards Accessible Learned Optimizers in PyTorch
di: Janson, Paul, et al.
Pubblicazione: (2025)
di: Janson, Paul, et al.
Pubblicazione: (2025)
Communication Efficient LLM Pre-training with SparseLoCo
di: Sarfi, Amir, et al.
Pubblicazione: (2025)
di: Sarfi, Amir, et al.
Pubblicazione: (2025)
Differentiation of Blackbox Combinatorial Solvers
di: Vlastelica, Marin, et al.
Pubblicazione: (2019)
di: Vlastelica, Marin, et al.
Pubblicazione: (2019)
Not Only the Last-Layer Features for Spurious Correlations: All Layer Deep Feature Reweighting
di: Hameed, Humza Wajid, et al.
Pubblicazione: (2024)
di: Hameed, Humza Wajid, et al.
Pubblicazione: (2024)
Incentivizing Permissionless Distributed Learning of LLMs
di: Lidin, Joel, et al.
Pubblicazione: (2025)
di: Lidin, Joel, et al.
Pubblicazione: (2025)
AdaFisher: Adaptive Second Order Optimization via Fisher Information
di: Gomes, Damien Martins, et al.
Pubblicazione: (2024)
di: Gomes, Damien Martins, et al.
Pubblicazione: (2024)
$μ$LO: Compute-Efficient Meta-Generalization of Learned Optimizers
di: Thérien, Benjamin, et al.
Pubblicazione: (2024)
di: Thérien, Benjamin, et al.
Pubblicazione: (2024)
DragD3D: Realistic Mesh Editing with Rigidity Control Driven by 2D Diffusion Priors
di: Xie, Tianhao, et al.
Pubblicazione: (2023)
di: Xie, Tianhao, et al.
Pubblicazione: (2023)
TIMED: Adversarial and Autoregressive Refinement of Diffusion-Based Time Series Generation
di: EskandariNasab, MohammadReza, et al.
Pubblicazione: (2025)
di: EskandariNasab, MohammadReza, et al.
Pubblicazione: (2025)
From Feature Visualization to Visual Circuits: Effect of Adversarial Model Manipulation
di: Nanfack, Geraldin, et al.
Pubblicazione: (2024)
di: Nanfack, Geraldin, et al.
Pubblicazione: (2024)
Enhancing Multivariate Time Series-based Solar Flare Prediction with Multifaceted Preprocessing and Contrastive Learning
di: EskandariNasab, MohammadReza, et al.
Pubblicazione: (2024)
di: EskandariNasab, MohammadReza, et al.
Pubblicazione: (2024)
Rethinking Hallucinations: Correctness, Consistency, and Prompt Multiplicity
di: Ganesh, Prakhar, et al.
Pubblicazione: (2026)
di: Ganesh, Prakhar, et al.
Pubblicazione: (2026)
MuLoCo: Muon is a practical inner optimizer for DiLoCo
di: Thérien, Benjamin, et al.
Pubblicazione: (2025)
di: Thérien, Benjamin, et al.
Pubblicazione: (2025)
Less is More: Undertraining Experts Improves Model Upcycling
di: Horoi, Stefan, et al.
Pubblicazione: (2025)
di: Horoi, Stefan, et al.
Pubblicazione: (2025)
Model Parallelism With Subnetwork Data Parallelism
di: Singh, Vaibhav, et al.
Pubblicazione: (2025)
di: Singh, Vaibhav, et al.
Pubblicazione: (2025)
Non-Uniform Parameter-Wise Model Merging
di: Camacho, Albert Manuel Orozco, et al.
Pubblicazione: (2024)
di: Camacho, Albert Manuel Orozco, et al.
Pubblicazione: (2024)
Deep Graph Matching via Blackbox Differentiation of Combinatorial Solvers
di: Rolínek, Michal, et al.
Pubblicazione: (2020)
di: Rolínek, Michal, et al.
Pubblicazione: (2020)
Outline of an Independent Systematic Blackbox Test for ML-based Systems
di: Wiesbrock, Hans-Werner, et al.
Pubblicazione: (2024)
di: Wiesbrock, Hans-Werner, et al.
Pubblicazione: (2024)
PETRA: Parallel End-to-end Training with Reversible Architectures
di: Rivaud, Stéphane, et al.
Pubblicazione: (2024)
di: Rivaud, Stéphane, et al.
Pubblicazione: (2024)
SolarGPT-QA: A Domain-Adaptive Large Language Model for Educational Question Answering in Space Weather and Heliophysics
di: Chapagain, Santosh, et al.
Pubblicazione: (2026)
di: Chapagain, Santosh, et al.
Pubblicazione: (2026)
Blackbox Model Provenance via Palimpsestic Membership Inference
di: Kuditipudi, Rohith, et al.
Pubblicazione: (2025)
di: Kuditipudi, Rohith, et al.
Pubblicazione: (2025)
Documenti analoghi
-
Model Breadcrumbs: Scaling Multi-Task Model Merging with Sparse Masks
di: Davari, MohammadReza, et al.
Pubblicazione: (2023) -
FairDropout: Using Example-Tied Dropout to Enhance Generalization of Minority Groups
di: Nanfack, Geraldin, et al.
Pubblicazione: (2025) -
Understanding and Exploiting Weight Update Sparsity for Communication-Efficient Distributed RL
di: Miahi, Erfan, et al.
Pubblicazione: (2026) -
Celo2: Towards Learned Optimization Free Lunch
di: Moudgil, Abhinav, et al.
Pubblicazione: (2026) -
Improving Structural Diversity of Blackbox LLMs via Chain-of-Specification Prompting
di: Young, Halley, et al.
Pubblicazione: (2024)