Turn opaque embeddings into 104 human-readable dimensions — Hardness, Love, Divinity… — then steer a live LLM with them. Causally verified dials, honest scorecard included.
pytorch embeddings llama interpretability llm mechanistic-interpretability semantic-space representation-engineering activation-steering
-
Updated
Jun 12, 2026 - Python