🚀 Qwythos-27B-v1 is here! The 27B you've been waiting for.
The bigger sibling to Qwythos-9B. Native MTP head intact, full vision tower, still uncensored, still 1M context. Apache-2.0.
huggingface.co/empero-ai/Qwy…
📄 Does Recurrence Pay? Our RLT paper is here!
The Recurrent Looped Transformer shipped without experiments. So we ran them.
Same data. Same order. Same recipe. 500M tokens.
At 140M params RLT is the worst model we trained. And it needed ~20× the GPU-hours. 🧵
We are at the dawn of Superintelligence.
Introducing the Recurrent Looped Transformer (RLT),
We now have Transformers with Infinite Reasoning depth.
From now on, we should pace progress at the Open Frontier of Superintelligence,
Until Safe Superintelligence is achieved.
github.com/yifanzhang-pro/re…
We wrote exact CUDA-graph and Triton kernels for the recurrence (4.4× faster, math untouched) to give RLT a fair shot.
Still 8.4K tok/s vs 124K for a Transformer on a 5090.
Recurrence doesn't pay. Not at this scale.
alphaxiv.org/abs/2609.recurr…
🚀 autocode is here! The simplest coding agent we could build.
One file. One tool: a shell. Stdlib only. It writes its own tools and rewrites its own source as it works. Any OpenAI-compatible endpoint. Apache-2.0.
pip install empero-autocode
github.com/empero-org/autoco…
How it works: autocode copies runner.py into your project and runs it. That file is the whole agent.
When it edits itself, it reloads mid-task. If the edit doesn't compile, the old version keeps running. Every project ends up with its own agent 🫶
empero.org/writing/autocode
Currently our servers are very overloaded so thank you for everyone using the API! 🚀
We are happy to say we have processed over 10B token during the last week! We got to learn a lot and ran some experiments in inference optimization.
The API will stay online and we will do our best to offer more models and capacity soon. We are working on some cool releases to be announced soon aswell! Stay tuned
Thank you to everyone that has used the Qwen endpoints! 🫶
We currently do not have the compute capacity to keep serving Qwen Flash due to training but will open another endpoint soon. Model to come!
In the meantime there is free GLM 5.3 Flash!
free.empero.org/v1 any key
Thank you to everyone that has used the Qwen endpoints! 🫶
We currently do not have the compute capacity to keep serving Qwen Flash due to training but will open another endpoint soon. Model to come!
In the meantime there is free GLM 5.3 Flash!
free.empero.org/v1 any key