A Tiny Network Packed Five Features into Two Dimensions — and Drew a Pentagon
Reproducing Anthropic’s toy model of superposition in pure NumPy with hand-derived gradients shows how a small neural network can encode more features than its dimensionality by arranging them at angles, producing a pentagon when compressing five sparse features into two dimensions.



