Description
Someone reverse-engineered how Claude actually thinks — and open-sourced it. 🧠
Not a bigger model. Not more parameters.
Just the same small set of layers... looped 16 times per thought.
770M parameters matching a 1.3B model. That's the theory.
Save this — this changes how you think about AI forever.
Does size even matter anymore? Comment below