Show HN: Nanointerpret – LLM Interpretability Playground(nanointerpret.pages.dev) I find LLM interpretability extremely interesting and wanted to create a minimal repo for:
- SAE training
- Automatic feature interpretation
- Visualizing features and running interventions through a GUI You can try it here: https://nanointerpret.pages.dev/ Or check the repo: https://github.com/Belluxx/nanointerpret |