Skip to content

Topic

Quantized open-weight models

Open-weight language models compressed to lower numeric precision so they fit in modest memory and run at usable speed on CPUs.

Current clusters