Toy-model experiments on how training configurations affect feature-entanglement proxies in neural representations.
-
Updated
Sep 10, 2026 - Python
Toy-model experiments on how training configurations affect feature-entanglement proxies in neural representations.
Reproducible causal-subspace experiments in GPT-2-small: interventions, matched-span controls, and audit-ready evidence.
To associate your repository with the polysemanticity topic, visit your repo's landing page and select "manage topics."