FlowBot: Inducing LLM Workflows with Bilevel Optimization and Textual Gradients
Fuente:
arXiv
Salvato in:
| Autori principali: | Yu, Hongyeon, Kim, Young-Bum, Kim, Yoon |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2026
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Textual Gradients are a Flawed Metaphor for Automatic Prompt Optimization
di: Melcer, Daniel, et al.
Pubblicazione: (2025)
di: Melcer, Daniel, et al.
Pubblicazione: (2025)
On the Duality between Gradient Transformations and Adapters
di: Torroba-Hennigen, Lucas, et al.
Pubblicazione: (2025)
di: Torroba-Hennigen, Lucas, et al.
Pubblicazione: (2025)
LCSB: Layer-Cyclic Selective Backpropagation for Memory-Efficient On-Device LLM Fine-Tuning
di: Park, Juneyoung, et al.
Pubblicazione: (2026)
di: Park, Juneyoung, et al.
Pubblicazione: (2026)
DataFlow: An LLM-Driven Framework for Unified Data Preparation and Workflow Automation in the Era of Data-Centric AI
di: Liang, Hao, et al.
Pubblicazione: (2025)
di: Liang, Hao, et al.
Pubblicazione: (2025)
R-Bot: An LLM-based Query Rewrite System
di: Sun, Zhaoyan, et al.
Pubblicazione: (2024)
di: Sun, Zhaoyan, et al.
Pubblicazione: (2024)
Joint Unsupervised and Supervised Training for Automatic Speech Recognition via Bilevel Optimization
di: Saif, A F M, et al.
Pubblicazione: (2024)
di: Saif, A F M, et al.
Pubblicazione: (2024)
References Indeed Matter? Reference-Free Preference Optimization for Conversational Query Reformulation
di: Kim, Doyoung, et al.
Pubblicazione: (2025)
di: Kim, Doyoung, et al.
Pubblicazione: (2025)
Not All Options Are Created Equal: Textual Option Weighting for Token-Efficient LLM-Based Knowledge Tracing
di: Kim, JongWoo, et al.
Pubblicazione: (2024)
di: Kim, JongWoo, et al.
Pubblicazione: (2024)
SEAL: Safety-enhanced Aligned LLM Fine-tuning via Bilevel Data Selection
di: Shen, Han, et al.
Pubblicazione: (2024)
di: Shen, Han, et al.
Pubblicazione: (2024)
Exploring Gradient-Guided Masked Language Model to Detect Textual Adversarial Attacks
di: Zhang, Xiaomei, et al.
Pubblicazione: (2025)
di: Zhang, Xiaomei, et al.
Pubblicazione: (2025)
Zero2Text: Zero-Training Cross-Domain Inversion Attacks on Textual Embeddings
di: Kim, Doohyun, et al.
Pubblicazione: (2026)
di: Kim, Doohyun, et al.
Pubblicazione: (2026)
Why Is RLHF Alignment Shallow? A Gradient Analysis
di: Young, Robin
Pubblicazione: (2026)
di: Young, Robin
Pubblicazione: (2026)
Flow-of-Options: Diversified and Improved LLM Reasoning by Thinking Through Options
di: Nair, Lakshmi, et al.
Pubblicazione: (2025)
di: Nair, Lakshmi, et al.
Pubblicazione: (2025)
TRPrompt: Bootstrapping Query-Aware Prompt Optimization from Textual Rewards
di: Nica, Andreea, et al.
Pubblicazione: (2025)
di: Nica, Andreea, et al.
Pubblicazione: (2025)
DiscoverLLM: From Executing Intents to Discovering Them
di: Kim, Tae Soo, et al.
Pubblicazione: (2026)
di: Kim, Tae Soo, et al.
Pubblicazione: (2026)
Retrieval Collapses When AI Pollutes the Web
di: Yu, Hongyeon, et al.
Pubblicazione: (2026)
di: Yu, Hongyeon, et al.
Pubblicazione: (2026)
Learn-to-learn on Arbitrary Textual Conditioning: A Hypernetwork-Driven Meta-Gated LLM
di: Ji, Luo, et al.
Pubblicazione: (2026)
di: Ji, Luo, et al.
Pubblicazione: (2026)
Multi-Response Preference Optimization with Augmented Ranking Dataset
di: Gwon, Hansle, et al.
Pubblicazione: (2024)
di: Gwon, Hansle, et al.
Pubblicazione: (2024)
Stable Preference Optimization: A Bilevel Approach to Catastrophic Preference Shift
di: Jian, Chengtao, et al.
Pubblicazione: (2025)
di: Jian, Chengtao, et al.
Pubblicazione: (2025)
Contextually Entangled Gradient Mapping for Optimized LLM Comprehension
di: Sisate, Colin, et al.
Pubblicazione: (2025)
di: Sisate, Colin, et al.
Pubblicazione: (2025)
Parallelizing Linear Transformers with the Delta Rule over Sequence Length
di: Yang, Songlin, et al.
Pubblicazione: (2024)
di: Yang, Songlin, et al.
Pubblicazione: (2024)
ConfPO: Exploiting Policy Model Confidence for Critical Token Selection in Preference Optimization
di: Yoon, Hee Suk, et al.
Pubblicazione: (2025)
di: Yoon, Hee Suk, et al.
Pubblicazione: (2025)
Hyperloop Transformers
di: Zeitoun, Abbas, et al.
Pubblicazione: (2026)
di: Zeitoun, Abbas, et al.
Pubblicazione: (2026)
MAGE: All-[MASK] Block Already Knows Where to Look in Diffusion LLM
di: Kwon, Omin, et al.
Pubblicazione: (2026)
di: Kwon, Omin, et al.
Pubblicazione: (2026)
LAB: Large-Scale Alignment for ChatBots
di: Sudalairaj, Shivchander, et al.
Pubblicazione: (2024)
di: Sudalairaj, Shivchander, et al.
Pubblicazione: (2024)
DataDreamer: A Tool for Synthetic Data Generation and Reproducible LLM Workflows
di: Patel, Ajay, et al.
Pubblicazione: (2024)
di: Patel, Ajay, et al.
Pubblicazione: (2024)
Do MLLMs Capture How Interfaces Guide User Behavior? A Benchmark for Multimodal UI/UX Design Understanding
di: Jeon, Jaehyun, et al.
Pubblicazione: (2025)
di: Jeon, Jaehyun, et al.
Pubblicazione: (2025)
$C^2$: Scalable Auto-Feedback for LLM-based Chart Generation
di: Koh, Woosung, et al.
Pubblicazione: (2024)
di: Koh, Woosung, et al.
Pubblicazione: (2024)
DistiLLM: Towards Streamlined Distillation for Large Language Models
di: Ko, Jongwoo, et al.
Pubblicazione: (2024)
di: Ko, Jongwoo, et al.
Pubblicazione: (2024)
From Static Benchmarks to Dynamic Protocol: Agent-Centric Text Anomaly Detection for Evaluating LLM Reasoning
di: Yoa, Seungdong, et al.
Pubblicazione: (2026)
di: Yoa, Seungdong, et al.
Pubblicazione: (2026)
ELF: Embedded Language Flows
di: Hu, Keya, et al.
Pubblicazione: (2026)
di: Hu, Keya, et al.
Pubblicazione: (2026)
In-Context Language Learning: Architectures and Algorithms
di: Akyürek, Ekin, et al.
Pubblicazione: (2024)
di: Akyürek, Ekin, et al.
Pubblicazione: (2024)
GFlowPO: Generative Flow Network as a Language Model Prompt Optimizer
di: Cho, Junmo, et al.
Pubblicazione: (2026)
di: Cho, Junmo, et al.
Pubblicazione: (2026)
DiZiNER: Disagreement-guided Instruction Refinement via Pilot Annotation Simulation for Zero-shot Named Entity Recognition
di: Kim, Siun, et al.
Pubblicazione: (2026)
di: Kim, Siun, et al.
Pubblicazione: (2026)
Translate Policy to Language: Flow Matching Generated Rewards for LLM Explanations
di: Yang, Xinyi, et al.
Pubblicazione: (2025)
di: Yang, Xinyi, et al.
Pubblicazione: (2025)
ERD: A Framework for Improving LLM Reasoning for Cognitive Distortion Classification
di: Lim, Sehee, et al.
Pubblicazione: (2024)
di: Lim, Sehee, et al.
Pubblicazione: (2024)
LQ-LoRA: Low-rank Plus Quantized Matrix Decomposition for Efficient Language Model Finetuning
di: Guo, Han, et al.
Pubblicazione: (2023)
di: Guo, Han, et al.
Pubblicazione: (2023)
EdiText: Controllable Coarse-to-Fine Text Editing with Diffusion Language Models
di: Lee, Che Hyun, et al.
Pubblicazione: (2025)
di: Lee, Che Hyun, et al.
Pubblicazione: (2025)
Harnessing Uncertainty: Entropy-Modulated Policy Gradients for Long-Horizon LLM Agents
di: Wang, Jiawei, et al.
Pubblicazione: (2025)
di: Wang, Jiawei, et al.
Pubblicazione: (2025)
Label Smoothing Improves Gradient Ascent in LLM Unlearning
di: Pang, Zirui, et al.
Pubblicazione: (2025)
di: Pang, Zirui, et al.
Pubblicazione: (2025)
Documenti analoghi
-
Textual Gradients are a Flawed Metaphor for Automatic Prompt Optimization
di: Melcer, Daniel, et al.
Pubblicazione: (2025) -
On the Duality between Gradient Transformations and Adapters
di: Torroba-Hennigen, Lucas, et al.
Pubblicazione: (2025) -
LCSB: Layer-Cyclic Selective Backpropagation for Memory-Efficient On-Device LLM Fine-Tuning
di: Park, Juneyoung, et al.
Pubblicazione: (2026) -
DataFlow: An LLM-Driven Framework for Unified Data Preparation and Workflow Automation in the Era of Data-Centric AI
di: Liang, Hao, et al.
Pubblicazione: (2025) -
R-Bot: An LLM-based Query Rewrite System
di: Sun, Zhaoyan, et al.
Pubblicazione: (2024)