Neural Networks, Secretly Symbolic
A new paper from McCoy, Soulos, Linzen, and Smolensky shows that neural network representations implicitly realize symbolic structures — precisely enough to replace the network's entire representation process with a closed-form equation, and to intervene on LLM behavior by directly editing the identified structures.
Read more →
