11:44
2026-07-27
lesswrong.com
ai-safety
Can we teach a model to encode a semantic feature on a chosen manifold in just three channels?
A researcher received an Honorable Mention in BlueDot's Technical AI Safety Puzzle #1 for training a small MLP to encode a country feature on a chosen nonlinear manifold in three reserved channels, deβ¦