Skip to content

Topic

Client-side LLM inference

The pattern where model calls originate in a browser or mobile client and reach a hosted model directly, instead of passing through a backend the developer runs.

Current clusters