transformer-lens-interpretability

Provides guidance for mechanistic interpretability research using TransformerLens to inspect and manipulate transformer internals via HookPoints and activation caching. Use when reverse-engineering model algorithms, studying attention patterns, or performing activation patching experiments.

By orchestra-research · 763 installs

npx skills add orchestra-research/ai-research-skills --skill transformer-lens-interpretability

Source repository · Upstream listing