interpretability instructions

6 tagged interpretability, measured the same way as everything else here.

hijohnnylin/neuronpedia

Instructions file GitHub Copilot

Instructions for hijohnnylin/neuronpedia: This repository keeps all agent instructions in AGENTS.md at the root, so that no rule is visible to one coding agent and invisible to another. Read it before making changes.

1.1k 3d ago A 121 tokens original Apache-2.0

hijohnnylin/neuronpedia

Instructions file CodexOpenCode

Instructions for hijohnnylin/neuronpedia, covering neuronpedia development guide, repository layout, committed files that are build outputs, configuration and .env files and plans/.

1.1k 3d ago A 7,202 tokens original Apache-2.0

hijohnnylin/neuronpedia

Instructions file

Instructions for hijohnnylin/neuronpedia, a project described as: open source interpretability platform 🧠.

1.1k 3d ago A 5 tokens copy Β· 100% Apache-2.0

nnsight CLAUDE.md

04

ndif-team/nnsight

Instructions file

Instructions for ndif-team/nnsight, covering nnsight β€” agent guide, how to use this file, by task, "i want multi-token / autoregressive generation" and "i want to run multiple prompts at once".

1.1k 4d ago A 3,828 tokens original MIT

dreadnode/agent-lens

Instructions file

Instructions for dreadnode/agent-lens, covering agentlens, project structure, running experiments, config format (yaml) and engines.

114 2mo ago A 2,762 tokens original MIT

ndif CLAUDE.md

06

ndif-team/ndif

Instructions file

Instructions for ndif-team/ndif, covering claude.md β€” ndif agent guide, what ndif is, top-level layout, architecture at a glance and security sandbox (the highest-stakes area).

51 4d ago A 4,885 tokens original MIT