Adam Shai @adamimos
[Top, cut off]: "...the loss like this, it will likely make the training unstable."
๐ฌ ๐ โค 7 ๐ 1.3K โ
Adam Shai @adamimos ยท Jul 25
You may be interested in this new work that shows neural networks take advantage of that non-orthogonality arxiv.org/abs/2507.07432...
[Link card: arxiv.org โ "Neural networks leverage nominally quantum and ..."]
๐ฌ 1 ๐ 3 โค 34 ๐ 1.5K โ
Dmitry Ryb... @DmitryRybi... ยท Jul 25
Very cool! Btw we can include some geometry of latent space by introducing a quadratic form/curvature matrix G and writing rho = sum p_i * u_i G (u_i)^T
๐ฌ ๐ โค 11 ๐ 1K โ
Thomas A... @thomasa... ยท Jul 25
Great explanation!
What is the cross entropy parallel? Did anyone try using it for training?
๐ฌ 2 ๐ โค 9 ๐ 2.6K โ
Dmitry Ryb... @DmitryRybi... ยท Jul 25
Good question, i haven't computed it.
Note from Claude Sonnet 5
A technical Twitter thread on the mathematics of neural network latent-space geometry โ non-orthogonality, quantum-like statistics in representations, curvature matrices for latent space geometry (rho = sum p_i * u_i G (u_i)^T). Interpretability/representation-theory content Nathan was reading; connects to his interest in interpretability and possibly the "platonic representation" thread noted in project memory (2026-05-13 import mentions a "Platonic hypothesis and model representation spaces" chat).
twitterinterpretabilityneural-network-geometrylatent-spacearxivrepresentation-theory