Bi-directional Model Cascading with Proxy Confidence
Fuente:
arXiv
Saved in:
| Main Authors: | Warren, David, Dras, Mark |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Cascading and Proxy Membership Inference Attacks
by: Du, Yuntao, et al.
Published: (2025)
by: Du, Yuntao, et al.
Published: (2025)
Gatekeeper: Improving Model Cascades Through Confidence Tuning
by: Rabanser, Stephan, et al.
Published: (2025)
by: Rabanser, Stephan, et al.
Published: (2025)
Empirical Calibration and Metric Differential Privacy in Language Models
by: Faustini, Pedro, et al.
Published: (2025)
by: Faustini, Pedro, et al.
Published: (2025)
When Does Confidence-Based Cascade Deferral Suffice?
by: Jitkrittum, Wittawat, et al.
Published: (2023)
by: Jitkrittum, Wittawat, et al.
Published: (2023)
Graded Suspiciousness of Adversarial Texts to Human
by: Tonni, Shakila Mahjabin, et al.
Published: (2024)
by: Tonni, Shakila Mahjabin, et al.
Published: (2024)
HybridServe: Efficient Serving of Large AI Models with Confidence-Based Cascade Routing
by: Xue, Leyang, et al.
Published: (2025)
by: Xue, Leyang, et al.
Published: (2025)
VITAL: A New Dataset for Benchmarking Pluralistic Alignment in Healthcare
by: Shetty, Anudeex, et al.
Published: (2025)
by: Shetty, Anudeex, et al.
Published: (2025)
Seeing the Forest through the Trees: Data Leakage from Partial Transformer Gradients
by: Li, Weijun, et al.
Published: (2024)
by: Li, Weijun, et al.
Published: (2024)
BiJEPA: Bi-directional Joint Embedding Predictive Architecture for Symmetric Representation Learning
by: Huang, Yongchao
Published: (2026)
by: Huang, Yongchao
Published: (2026)
JanusDNA: A Powerful Bi-directional Hybrid DNA Foundation Model
by: Duan, Qihao, et al.
Published: (2025)
by: Duan, Qihao, et al.
Published: (2025)
Proxy Compression for Language Modeling
by: Zheng, Lin, et al.
Published: (2026)
by: Zheng, Lin, et al.
Published: (2026)
ProxyKV: Cross-Model Proxy Pruning for Efficient Long-Context LLM Inference
by: Li, Junjie, et al.
Published: (2026)
by: Li, Junjie, et al.
Published: (2026)
Proxy-RLHF: Decoupling Generation and Alignment in Large Language Model with Proxy
by: Zhu, Yu, et al.
Published: (2024)
by: Zhu, Yu, et al.
Published: (2024)
IDT: Dual-Task Adversarial Attacks for Privacy Protection
by: Faustini, Pedro, et al.
Published: (2024)
by: Faustini, Pedro, et al.
Published: (2024)
Prior Distribution and Model Confidence
by: Kazanskii, Maksim, et al.
Published: (2025)
by: Kazanskii, Maksim, et al.
Published: (2025)
PAPM: A Physics-aware Proxy Model for Process Systems
by: Liu, Pengwei, et al.
Published: (2024)
by: Liu, Pengwei, et al.
Published: (2024)
Proxy-Guided Measurement Calibration
by: Vishnubhatla, Saketh, et al.
Published: (2026)
by: Vishnubhatla, Saketh, et al.
Published: (2026)
Proxy Methods for Domain Adaptation
by: Tsai, Katherine, et al.
Published: (2024)
by: Tsai, Katherine, et al.
Published: (2024)
Bi-directional Recurrence Improves Transformer in Partially Observable Markov Decision Processes
by: Arora, Ashok, et al.
Published: (2025)
by: Arora, Ashok, et al.
Published: (2025)
TIMBA: Time series Imputation with Bi-directional Mamba Blocks and Diffusion models
by: Solís-García, Javier, et al.
Published: (2024)
by: Solís-García, Javier, et al.
Published: (2024)
Human-AI Collaborative Autonomous Experimentation With Proxy Modeling for Comparative Observation
by: Biswas, Arpan, et al.
Published: (2026)
by: Biswas, Arpan, et al.
Published: (2026)
Tighter Confidence Bounds for Sequential Kernel Regression
by: Flynn, Hamish, et al.
Published: (2024)
by: Flynn, Hamish, et al.
Published: (2024)
C3PO: Optimized Large Language Model Cascades with Probabilistic Cost Constraints for Reasoning
by: Valkanas, Antonios, et al.
Published: (2025)
by: Valkanas, Antonios, et al.
Published: (2025)
Multiaccuracy and Multicalibration via Proxy Groups
by: Bharti, Beepul, et al.
Published: (2025)
by: Bharti, Beepul, et al.
Published: (2025)
Predicting LLM Reasoning Performance with Small Proxy Model
by: Koh, Woosung, et al.
Published: (2025)
by: Koh, Woosung, et al.
Published: (2025)
Accelerated Preference Elicitation with LLM-Based Proxies
by: Huang, David, et al.
Published: (2025)
by: Huang, David, et al.
Published: (2025)
BBS: Bi-directional Bit-level Sparsity for Deep Learning Acceleration
by: Chen, Yuzong, et al.
Published: (2024)
by: Chen, Yuzong, et al.
Published: (2024)
Personalized Steering of Large Language Models: Versatile Steering Vectors Through Bi-directional Preference Optimization
by: Cao, Yuanpu, et al.
Published: (2024)
by: Cao, Yuanpu, et al.
Published: (2024)
CATTO: Balancing Preferences and Confidence in Language Models
by: Parikh, Nisarg, et al.
Published: (2026)
by: Parikh, Nisarg, et al.
Published: (2026)
Confidence-Aware Multi-Field Model Calibration
by: Zhao, Yuang, et al.
Published: (2024)
by: Zhao, Yuang, et al.
Published: (2024)
On the Runway Cascade of Transformers for Language Modeling
by: Lee, Hunjae, et al.
Published: (2026)
by: Lee, Hunjae, et al.
Published: (2026)
OptiProxy-NAS: Optimization Proxy based End-to-End Neural Architecture Search
by: Lyu, Bo, et al.
Published: (2025)
by: Lyu, Bo, et al.
Published: (2025)
Kernel Single Proxy Control for Deterministic Confounding
by: Xu, Liyuan, et al.
Published: (2023)
by: Xu, Liyuan, et al.
Published: (2023)
Balanced Filtering via Disclosure-Controlled Proxies
by: Deng, Siqi, et al.
Published: (2023)
by: Deng, Siqi, et al.
Published: (2023)
Uncertainty Distillation: Teaching Language Models to Express Semantic Confidence
by: Hager, Sophia, et al.
Published: (2025)
by: Hager, Sophia, et al.
Published: (2025)
Speed is Confidence
by: Dillon, Joshua V.
Published: (2026)
by: Dillon, Joshua V.
Published: (2026)
Confidence Intervals and Simultaneous Confidence Bands Based on Deep Learning
by: Arie, Asaf Ben, et al.
Published: (2024)
by: Arie, Asaf Ben, et al.
Published: (2024)
FedVeca: Federated Vectorized Averaging on Non-IID Data with Adaptive Bi-directional Global Objective
by: Luo, Ping, et al.
Published: (2022)
by: Luo, Ping, et al.
Published: (2022)
FedPFT: Federated Proxy Fine-Tuning of Foundation Models
by: Peng, Zhaopeng, et al.
Published: (2024)
by: Peng, Zhaopeng, et al.
Published: (2024)
FedProxy: Federated Fine-Tuning of LLMs via Proxy SLMs and Heterogeneity-Aware Fusion
by: Fan, Tao, et al.
Published: (2026)
by: Fan, Tao, et al.
Published: (2026)
Similar Items
-
Cascading and Proxy Membership Inference Attacks
by: Du, Yuntao, et al.
Published: (2025) -
Gatekeeper: Improving Model Cascades Through Confidence Tuning
by: Rabanser, Stephan, et al.
Published: (2025) -
Empirical Calibration and Metric Differential Privacy in Language Models
by: Faustini, Pedro, et al.
Published: (2025) -
When Does Confidence-Based Cascade Deferral Suffice?
by: Jitkrittum, Wittawat, et al.
Published: (2023) -
Graded Suspiciousness of Adversarial Texts to Human
by: Tonni, Shakila Mahjabin, et al.
Published: (2024)