Skip to content

other

LLM.int8()

A 2022 paper and method for 8-bit Transformer inference that handles large-magnitude activation features separately from the rest of a tensor.

Known aliases

  • LLM.int8

Relationships

No evidence-backed relationships are recorded.

Current clusters