Cite
Notes
Only stored in your browser.
Attribution
Not All Language Model Features Are Linear
arXiv 2024
Opening the AI black box: program synthesis via mechanistic interpretability
from 2 papers
Eric J. Michaud
Max Tegmark
professor
Anish Mudide
Chloe Loughridge
Joshua Engels
Mateja Vukelić
Tara Rezaei Kheirkhah
Vedang Lad
Wes Gurnee
Zifan Carl Guo