Auditing Pay-Per-Token in Large Language Models

Fuente: arXiv
Saved in:
Bibliographic Details
Main Authors: Velasco, Ander Artola, Tsirtsis, Stratis, Gomez-Rodriguez, Manuel
Format: Preprint
Published: 2025
Subjects:
Online Access:
Tags: Add Tag
No Tags, Be the first to tag this record!
_version_ 1866911536673456128
author Velasco, Ander Artola
Tsirtsis, Stratis
Gomez-Rodriguez, Manuel
author_facet Velasco, Ander Artola
Tsirtsis, Stratis
Gomez-Rodriguez, Manuel
contents Millions of users rely on a market of cloud-based services to obtain access to state-of-the-art large language models. However, it has been very recently shown that the de facto pay-per-token pricing mechanism used by providers creates a financial incentive for them to strategize and misreport the (number of) tokens a model used to generate an output. In this paper, we develop an auditing framework based on martingale theory that enables a trusted third-party auditor who sequentially queries a provider to detect token misreporting. Crucially, we show that our framework is guaranteed to always detect token misreporting, regardless of the provider's (mis-)reporting policy, and not falsely flag a faithful provider as unfaithful with high probability. To validate our auditing framework, we conduct experiments across a wide range of (mis-)reporting policies using several large language models from the $\texttt{Llama}$, $\texttt{Gemma}$ and $\texttt{Ministral}$ families, and input prompts from a popular crowdsourced benchmarking platform. The results show that our framework detects an unfaithful provider after observing fewer than $\sim 70$ reported outputs, while maintaining the probability of falsely flagging a faithful provider below $α= 0.05$.
format Preprint
id arxiv_https___arxiv_org_abs_2510_05181
institution arXiv
publishDate 2025
record_format arxiv
spellingShingle Auditing Pay-Per-Token in Large Language Models
Velasco, Ander Artola
Tsirtsis, Stratis
Gomez-Rodriguez, Manuel
Cryptography and Security
Artificial Intelligence
Computers and Society
Millions of users rely on a market of cloud-based services to obtain access to state-of-the-art large language models. However, it has been very recently shown that the de facto pay-per-token pricing mechanism used by providers creates a financial incentive for them to strategize and misreport the (number of) tokens a model used to generate an output. In this paper, we develop an auditing framework based on martingale theory that enables a trusted third-party auditor who sequentially queries a provider to detect token misreporting. Crucially, we show that our framework is guaranteed to always detect token misreporting, regardless of the provider's (mis-)reporting policy, and not falsely flag a faithful provider as unfaithful with high probability. To validate our auditing framework, we conduct experiments across a wide range of (mis-)reporting policies using several large language models from the $\texttt{Llama}$, $\texttt{Gemma}$ and $\texttt{Ministral}$ families, and input prompts from a popular crowdsourced benchmarking platform. The results show that our framework detects an unfaithful provider after observing fewer than $\sim 70$ reported outputs, while maintaining the probability of falsely flagging a faithful provider below $α= 0.05$.
title Auditing Pay-Per-Token in Large Language Models
topic Cryptography and Security
Artificial Intelligence
Computers and Society
url https://arxiv.org/abs/2510.05181