Rational Tuning of LLM Cascades via Probabilistic Modeling
Fuente:
arXiv
Saved in:
| Main Authors: | Zellinger, Michael J., Thomson, Matt |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Efficiently Deploying LLMs with Controlled Risk
by: Zellinger, Michael J., et al.
Published: (2024)
by: Zellinger, Michael J., et al.
Published: (2024)
Fail Fast, or Ask: Mitigating the Deficiencies of Reasoning LLMs with Human-in-the-Loop Systems Engineering
by: Zellinger, Michael J., et al.
Published: (2025)
by: Zellinger, Michael J., et al.
Published: (2025)
Cost-Saving LLM Cascades with Early Abstention
by: Zellinger, Michael J., et al.
Published: (2025)
by: Zellinger, Michael J., et al.
Published: (2025)
Economic Evaluation of LLMs
by: Zellinger, Michael J., et al.
Published: (2025)
by: Zellinger, Michael J., et al.
Published: (2025)
Learning with Noisy Labels by Adaptive Gradient-Based Outlier Removal
by: Sedova, Anastasiia, et al.
Published: (2023)
by: Sedova, Anastasiia, et al.
Published: (2023)
Counterfactual Reasoning with Knowledge Graph Embeddings
by: Zellinger, Lena, et al.
Published: (2024)
by: Zellinger, Lena, et al.
Published: (2024)
Generalised Probabilistic Modelling and Improved Uncertainty Estimation in Comparative LLM-as-a-judge
by: Fathullah, Yassir, et al.
Published: (2025)
by: Fathullah, Yassir, et al.
Published: (2025)
Context-Aware Probabilistic Modeling with LLM for Multimodal Time Series Forecasting
by: Yao, Yueyang, et al.
Published: (2025)
by: Yao, Yueyang, et al.
Published: (2025)
Probabilistic Federated Prompt-Tuning with Non-IID and Imbalanced Data
by: Weng, Pei-Yau, et al.
Published: (2025)
by: Weng, Pei-Yau, et al.
Published: (2025)
Label-Free Reinforcement Learning via Cross-Model Entropy
by: Gorbett, Matt, et al.
Published: (2026)
by: Gorbett, Matt, et al.
Published: (2026)
What's the Magic Word? A Control Theory of LLM Prompting
by: Bhargava, Aman, et al.
Published: (2023)
by: Bhargava, Aman, et al.
Published: (2023)
Probabilistic ML Verification via Weighted Model Integration
by: Morettin, Paolo, et al.
Published: (2024)
by: Morettin, Paolo, et al.
Published: (2024)
Graffe: Graph Representation Learning via Diffusion Probabilistic Models
by: Chen, Dingshuo, et al.
Published: (2025)
by: Chen, Dingshuo, et al.
Published: (2025)
Likelihood Training of Cascaded Diffusion Models via Hierarchical Volume-preserving Maps
by: Li, Henry, et al.
Published: (2025)
by: Li, Henry, et al.
Published: (2025)
CasCast: Skillful High-resolution Precipitation Nowcasting via Cascaded Modelling
by: Gong, Junchao, et al.
Published: (2024)
by: Gong, Junchao, et al.
Published: (2024)
Towards Realistic Guarantees: A Probabilistic Certificate for SmoothLLM
by: Kumarappan, Adarsh, et al.
Published: (2025)
by: Kumarappan, Adarsh, et al.
Published: (2025)
Alignment Dynamics in LLM Fine-Tuning
by: Huang, Yuhan, et al.
Published: (2026)
by: Huang, Yuhan, et al.
Published: (2026)
CATS: Cascaded Adaptive Tree Speculation for Memory-Limited LLM Inference Acceleration
by: Han, Yuning, et al.
Published: (2026)
by: Han, Yuning, et al.
Published: (2026)
GILT: An LLM-Free, Tuning-Free Graph Foundational Model for In-Context Learning
by: Ma, Weishuo, et al.
Published: (2025)
by: Ma, Weishuo, et al.
Published: (2025)
Thompson Sampling via Fine-Tuning of LLMs
by: Menet, Nicolas, et al.
Published: (2025)
by: Menet, Nicolas, et al.
Published: (2025)
DipLLM: Fine-Tuning LLM for Strategic Decision-making in Diplomacy
by: Xu, Kaixuan, et al.
Published: (2025)
by: Xu, Kaixuan, et al.
Published: (2025)
Probabilistic Dreaming for World Models
by: Wong, Gavin
Published: (2026)
by: Wong, Gavin
Published: (2026)
Wireless Federated Multi-Task LLM Fine-Tuning via Sparse-and-Orthogonal LoRA
by: Yang, Nuocheng, et al.
Published: (2026)
by: Yang, Nuocheng, et al.
Published: (2026)
ECLIPTICA -- A Framework for Switchable LLM Alignment via CITA - Contrastive Instruction-Tuned Alignment
by: Wanaskar, Kapil, et al.
Published: (2026)
by: Wanaskar, Kapil, et al.
Published: (2026)
Robust and Efficient Zeroth-Order LLM Fine-Tuning via Adaptive Bayesian Subspace Optimizer
by: Feng, Jian, et al.
Published: (2026)
by: Feng, Jian, et al.
Published: (2026)
SafeTuneBed: A Toolkit for Benchmarking LLM Safety Alignment in Fine-Tuning
by: Hossain, Saad, et al.
Published: (2025)
by: Hossain, Saad, et al.
Published: (2025)
Hyperparameter Optimization via Interacting with Probabilistic Circuits
by: Seng, Jonas, et al.
Published: (2025)
by: Seng, Jonas, et al.
Published: (2025)
Probabilistic Circuits with Constraints via Convex Optimization
by: Ghandi, Soroush, et al.
Published: (2024)
by: Ghandi, Soroush, et al.
Published: (2024)
Probabilistic Graph Circuits: Deep Generative Models for Tractable Probabilistic Inference over Graphs
by: Papež, Milan, et al.
Published: (2025)
by: Papež, Milan, et al.
Published: (2025)
Probabilistic Abduction for Visual Abstract Reasoning via Learning Rules in Vector-symbolic Architectures
by: Hersche, Michael, et al.
Published: (2024)
by: Hersche, Michael, et al.
Published: (2024)
On the Relationship between Bayesian Networks and Probabilistic Structural Causal Models
by: Lucas, Peter J. F., et al.
Published: (2026)
by: Lucas, Peter J. F., et al.
Published: (2026)
Continuous Mixtures of Tractable Probabilistic Models
by: Correia, Alvaro H. C., et al.
Published: (2022)
by: Correia, Alvaro H. C., et al.
Published: (2022)
Observation-Guided Diffusion Probabilistic Models
by: Kang, Junoh, et al.
Published: (2023)
by: Kang, Junoh, et al.
Published: (2023)
A Rational Model of Dimension-reduced Human Categorization
by: Hong, Yifan, et al.
Published: (2023)
by: Hong, Yifan, et al.
Published: (2023)
LENSLLM: Unveiling Fine-Tuning Dynamics for LLM Selection
by: Zeng, Xinyue, et al.
Published: (2025)
by: Zeng, Xinyue, et al.
Published: (2025)
Internalizing Curriculum Judgment for LLM Reinforcement Fine-Tuning
by: Zheng, Han, et al.
Published: (2026)
by: Zheng, Han, et al.
Published: (2026)
Overcoming Forgetting in LLM Fine-Tuning with Evolution Strategies
by: Schweighofer, Kajetan, et al.
Published: (2026)
by: Schweighofer, Kajetan, et al.
Published: (2026)
Goal-Conditioned Supervised Learning for LLM Fine-Tuning
by: Li, Shijun, et al.
Published: (2026)
by: Li, Shijun, et al.
Published: (2026)
Fine-Tuning Diffusion Models via Intermediate Distribution Shaping
by: Anil, Gautham Govind, et al.
Published: (2025)
by: Anil, Gautham Govind, et al.
Published: (2025)
Geometry-Aware Probabilistic Circuits via Voronoi Tessellations
by: Sidheekh, Sahil, et al.
Published: (2026)
by: Sidheekh, Sahil, et al.
Published: (2026)
Similar Items
-
Efficiently Deploying LLMs with Controlled Risk
by: Zellinger, Michael J., et al.
Published: (2024) -
Fail Fast, or Ask: Mitigating the Deficiencies of Reasoning LLMs with Human-in-the-Loop Systems Engineering
by: Zellinger, Michael J., et al.
Published: (2025) -
Cost-Saving LLM Cascades with Early Abstention
by: Zellinger, Michael J., et al.
Published: (2025) -
Economic Evaluation of LLMs
by: Zellinger, Michael J., et al.
Published: (2025) -
Learning with Noisy Labels by Adaptive Gradient-Based Outlier Removal
by: Sedova, Anastasiia, et al.
Published: (2023)