x
Can we teach a model to encode a semantic feature on a chosen manifold in just three channels? — LessWrong