Mapping the Mind of a Neural Net: Goodfire’s Eric Ho on the Future of Interpretability
Source
youtube.com
Author
Sequoia Capital
Date
Why it matters
Interpretability tooling aims to let you inspect, steer and debug model behavior directly, a possible alternative to prompting and fine-tuningTaking a trained model and training it a bit more on your own examples so it gets better at one specific job.Full definition → for reliability.
Terms in this piece · Glossary
fine-tuning — Taking a trained model and training it a bit more on your own examples so it gets better at one specific job.