Every Component Is a Lookup: One Linear Graph for Interaction, Composition and Attribution

arXiv:2605.23393v3 Announce Type: replace-cross Abstract: Interpretability methods for transformers are typically built around separate questions: which components interact, how information routes to the output, and which input tokens contribute. Because these methods rely on different assumptions…

aiscience

Sources