Skip to content

other

Tuned lens

Variant of logit lens that learns a per-layer affine transformation, initialised at the identity and trained to minimise KL divergence against the model's final-layer output distribution.

Current clusters