Transformer internals are sparser than they look
A new analysis finds that most neurons in large AI models depend on only a small subset of earlier neurons, revealing hidden structure.
A new analysis finds that most neurons in large AI models depend on only a small subset of earlier neurons, revealing hidden structure.