Skip to content

Topic

LLM Token Consumption

The input and output tokens a model call consumes, which determines spend and how quickly a deployment reaches its rate and context limits.

Current clusters