Loading the catalog…
Loading the catalog…
We use GPT-4 to automatically write explanations for the behavior of neurons in large language models and to score those explanations. We release a dataset of these (imperfect) explanations and scores for every neuron in GPT-2.
What RADAR observed and classified to build this opportunity. It is what the source published, not a verification that the offer is still active.
Language models can explain neurons in language models. We use GPT-4 to automatically write explanations for the behavior of neurons in large language models and to score those explanations. We release a dataset of these (imperfect) explanations and scores for every neuron in GPT-2.
Open sourceOpens an external website. Availability and terms may change.