Model Equality Testing: Which Model Is This API Serving?
Fuente:
arXiv
Saved in:
| Main Authors: | Gao, Irena, Liang, Percy, Guestrin, Carlos |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Discovering Implicit Large Language Model Alignment Objectives
by: Chen, Edward, et al.
Published: (2026)
by: Chen, Edward, et al.
Published: (2026)
Independence Tests for Language Models
by: Zhu, Sally, et al.
Published: (2025)
by: Zhu, Sally, et al.
Published: (2025)
Post-Hoc Reversal: Are We Selecting Models Prematurely?
by: Ranjan, Rishabh, et al.
Published: (2024)
by: Ranjan, Rishabh, et al.
Published: (2024)
Out-of-Domain Robustness via Targeted Augmentations
by: Gao, Irena, et al.
Published: (2023)
by: Gao, Irena, et al.
Published: (2023)
Self-Verified Distillation: Your Language Model Is Secretly Its Own Synthetic Data Pipeline
by: Lee, Tony, et al.
Published: (2026)
by: Lee, Tony, et al.
Published: (2026)
Learning to (Learn at Test Time)
by: Sun, Yu, et al.
Published: (2023)
by: Sun, Yu, et al.
Published: (2023)
AlpacaFarm: A Simulation Framework for Methods that Learn from Human Feedback
by: Dubois, Yann, et al.
Published: (2023)
by: Dubois, Yann, et al.
Published: (2023)
On the Entropy Calibration of Language Models
by: Cao, Steven, et al.
Published: (2025)
by: Cao, Steven, et al.
Published: (2025)
Deep Generative Models with Hard Linear Equality Constraints
by: Li, Ruoyan, et al.
Published: (2025)
by: Li, Ruoyan, et al.
Published: (2025)
The Morgan-Pitman Test of Equality of Variances and its Application to Machine Learning Model Evaluation and Selection
by: Arratia, Argimiro, et al.
Published: (2025)
by: Arratia, Argimiro, et al.
Published: (2025)
Which Model to Transfer? A Survey on Transferability Estimation
by: Ding, Yuhe, et al.
Published: (2024)
by: Ding, Yuhe, et al.
Published: (2024)
Modeling Multi-Objective Tradeoffs with Monotonic Utility Functions
by: Chen, Edward, et al.
Published: (2024)
by: Chen, Edward, et al.
Published: (2024)
HybridServe: Efficient Serving of Large AI Models with Confidence-Based Cascade Routing
by: Xue, Leyang, et al.
Published: (2025)
by: Xue, Leyang, et al.
Published: (2025)
PluRel: Synthetic Data unlocks Scaling Laws for Relational Foundation Models
by: Kothapalli, Vignesh, et al.
Published: (2026)
by: Kothapalli, Vignesh, et al.
Published: (2026)
Interactive Multi-Objective Probabilistic Preference Learning with Soft and Hard Bounds
by: Chen, Edward, et al.
Published: (2025)
by: Chen, Edward, et al.
Published: (2025)
Automating REST API Postman Test Cases Using LLM
by: Sri, S Deepika, et al.
Published: (2024)
by: Sri, S Deepika, et al.
Published: (2024)
On the Cost of Model-Serving Frameworks: An Experimental Evaluation
by: De Rosa, Pasquale, et al.
Published: (2024)
by: De Rosa, Pasquale, et al.
Published: (2024)
Learning to Discover at Test Time
by: Yuksekgonul, Mert, et al.
Published: (2026)
by: Yuksekgonul, Mert, et al.
Published: (2026)
Replaying pre-training data improves fine-tuning
by: Kotha, Suhas, et al.
Published: (2026)
by: Kotha, Suhas, et al.
Published: (2026)
Robust Distortion-free Watermarks for Language Models
by: Kuditipudi, Rohith, et al.
Published: (2023)
by: Kuditipudi, Rohith, et al.
Published: (2023)
Apt-Serve: Adaptive Request Scheduling on Hybrid Cache for Scalable LLM Inference Serving
by: Gao, Shihong, et al.
Published: (2025)
by: Gao, Shihong, et al.
Published: (2025)
CascadeServe: Unlocking Model Cascades for Inference Serving
by: Kossmann, Ferdi, et al.
Published: (2024)
by: Kossmann, Ferdi, et al.
Published: (2024)
On the Learnability of Watermarks for Language Models
by: Gu, Chenchen, et al.
Published: (2023)
by: Gu, Chenchen, et al.
Published: (2023)
Cache & Distil: Optimising API Calls to Large Language Models
by: Ramírez, Guillem, et al.
Published: (2023)
by: Ramírez, Guillem, et al.
Published: (2023)
The CAP Principle for LLM Serving: A Survey of Long-Context Large Language Model Serving
by: Zeng, Pai, et al.
Published: (2024)
by: Zeng, Pai, et al.
Published: (2024)
Quick-Tune: Quickly Learning Which Pretrained Model to Finetune and How
by: Arango, Sebastian Pineda, et al.
Published: (2023)
by: Arango, Sebastian Pineda, et al.
Published: (2023)
Rethinking Key-Value Cache Compression Techniques for Large Language Model Serving
by: Gao, Wei, et al.
Published: (2025)
by: Gao, Wei, et al.
Published: (2025)
DuetServe: Harmonizing Prefill and Decode for LLM Serving via Adaptive GPU Multiplexing
by: Gao, Lei, et al.
Published: (2025)
by: Gao, Lei, et al.
Published: (2025)
End-to-End Test-Time Training for Long Context
by: Tandon, Arnuv, et al.
Published: (2025)
by: Tandon, Arnuv, et al.
Published: (2025)
Efficient Diffusion Models under Nonconvex Equality and Inequality constraints via Landing
by: Jeon, Kijung, et al.
Published: (2026)
by: Jeon, Kijung, et al.
Published: (2026)
Learning with Statistical Equality Constraints
by: Barthakur, Aneesh, et al.
Published: (2025)
by: Barthakur, Aneesh, et al.
Published: (2025)
Identified-Set Geometry of Distributional Model Extraction under Top-$K$ Censored API Access
by: Nie, Wenhua, et al.
Published: (2026)
by: Nie, Wenhua, et al.
Published: (2026)
Fairness in Serving Large Language Models
by: Sheng, Ying, et al.
Published: (2023)
by: Sheng, Ying, et al.
Published: (2023)
When RL Meets Adaptive Speculative Training: A Unified Training-Serving System
by: Wang, Junxiong, et al.
Published: (2026)
by: Wang, Junxiong, et al.
Published: (2026)
An Interpretable Latency Model for Speculative Decoding in LLM Serving
by: Kong, Linghao, et al.
Published: (2026)
by: Kong, Linghao, et al.
Published: (2026)
EdgeServe: A Streaming System for Decentralized Model Serving
by: Shaowang, Ted, et al.
Published: (2023)
by: Shaowang, Ted, et al.
Published: (2023)
Blackbox Model Provenance via Palimpsestic Membership Inference
by: Kuditipudi, Rohith, et al.
Published: (2025)
by: Kuditipudi, Rohith, et al.
Published: (2025)
Large Language Models as Analogical Reasoners
by: Yasunaga, Michihiro, et al.
Published: (2023)
by: Yasunaga, Michihiro, et al.
Published: (2023)
RW-TTT: Batched Serving for Request-Owned Test-Time Training State
by: Yang, Jian, et al.
Published: (2026)
by: Yang, Jian, et al.
Published: (2026)
A Paired Testing Protocol for Batch-Conditioned Refusal Robustness in LLM Serving
by: Kadadekar, Sahil
Published: (2026)
by: Kadadekar, Sahil
Published: (2026)
Similar Items
-
Discovering Implicit Large Language Model Alignment Objectives
by: Chen, Edward, et al.
Published: (2026) -
Independence Tests for Language Models
by: Zhu, Sally, et al.
Published: (2025) -
Post-Hoc Reversal: Are We Selecting Models Prematurely?
by: Ranjan, Rishabh, et al.
Published: (2024) -
Out-of-Domain Robustness via Targeted Augmentations
by: Gao, Irena, et al.
Published: (2023) -
Self-Verified Distillation: Your Language Model Is Secretly Its Own Synthetic Data Pipeline
by: Lee, Tony, et al.
Published: (2026)