Skip to content
#

interpretability-ai

Here are 2 public repositories matching this topic...

Located the number "five" inside Qwen2.5-32B and DeepSeek-V4-Flash (304B) and replaced it with "four" at 128 neurons/layer. The model still reads 5 but computes with it as 4, and in every test we ran it never noticed. Full records, reproducible on low-end hardware.

  • Updated Oct 5, 2026
  • Python

Add this topic to your repo

To associate your repository with the interpretability-ai topic, visit your repo's landing page and select "manage topics."

Learn more