inceptionv1

Tag

Cards List
#inceptionv1

Mechanistic interpretability: a first paper on disentangling a convolutional neuron [R]

Reddit r/MachineLearning · 2026-07-15

This paper introduces a technique to disentangle a single convolutional neuron in Inceptionv1 by analyzing Hadamard products, revealing clean monosemantic clusters (cars, cats, dogs) and also low-valued clusters (letters, faces) with distributed weights as evidence of gradient descent behavior.

0 favorites 0 likes
← Back to home

Submit Feedback