this is getting to be a theme. I figure out something amazing . cant get traction ..a year later some academics write about it like its brand new .. i post receipts showing we already had the engine .....not a thing happens and the cycle is repeated .
"Hyperloop Transformers"
This paper propose a memory-efficient LLM via looped Transformers.
They basically reuse the middle block across depth, then add hyper-connections only between loops.
Key result is that this restores flexibility lost from weight sharing, letting the model beat depth-matched Transformers with ~50% fewer parameters. The result still holds after INT4 quantization too.