Skip to content

Topic

Inference offloading

Running a device's AI models on nearby edge servers or cloud GPUs instead of on hardware carried by the device itself.

Current clusters