We make tinygrad; sell tinybox for the GPU middle class. Our mission is to commoditize the petaflop.

San Diego / Hong Kong
Now that the AI spam has died down, the bounty program is back! We have four open bounties. If you want a job here, this is the way. If you are only willing to put in an AI amount of effort, please don't, it won't work.
17
12
574
103,458
25 years of technology progress
25
20
732
15,401
GLM-5.3-Flash, 250 tok/s interactive, $25,000. Who would buy the "tinybox red flash"?
78
18
881
65,920
China views AI as a public utility. The US views AI as a more efficient rent seeking apparatus. Both groups are correct in their assessment of if AI will help the country.
This is one of China's most formidable advantages: 93% of the Chinese believe AI will help their country. Only 36% of Americans do. These are dire numbers for the US. We have to turn this around.
10
63
771
25,210
the tiny corp retweeted
Replying to @stochasticchasm
tinygrad if you want approachable, small enough to hold in your head. Barely. MLIR/IREE for modern. TVM splits the difference.
1
1
10
4,872
the tiny corp retweeted
Tinygrad is amazing! It now runs on any Vulkan 1.2 device I've tested: AMD APU, NVIDIA, Intel iGPU, and even an Android tablet under Termux (Adreno 710 GPU) . Don't expect speed though, it's not optimized. I get ~0.4x the CUDA backend on my NVIDIA card, but it works.
2
1
23
3,467
There's something missing from benchmarks. Kimi K3 often writes much better code than GPT-6 Astra. It doesn't write useless verbose tests, it doesn't give confusing explanations. To be a good software engineer, you have to be able to communicate well, and RL doesn't capture that.
70
97
2,755
223,675
~200 tok/s MiMo-V2.6-Pro on MI300X brought up by GLM-5.3
16
9
468
17,045
And they want you to believe the US labs are worth trillions. So much respect to @Xiaomi for doing this. Commoditize the model training!
Pretty insane result They spent 130 hours, 75B tokens, and $2.6M on RL to achieve this result
13
70
1,558
67,398
Congrats on the release! From someone using it to code, is it good? Like I'm not sure how AA ranks GLM-5.3 above Kimi K3, I feel like RL has been breaking the benchmarks.
Introducing Xiaomi MiMo-V2.6 — Pro & Flash. Frontier intelligence, all the modalities, built in public. 🔹 Two omnimodal models, advancing through scaled reinforcement learning 🔹 Pro performs on par with Claude Opus 5 and GPT-5.6 Sol across most agent benchmarks 🔹 Pro scores 46 on the Artificial Analysis Intelligence Index — the highest among open-source models 🔹 Stronger coding, computer use, 3D reasoning and creative capabilities 🔹 Open model weights, technical report, RL environments and training code Blog:mimo.xiaomi.com/mimo-v2-6
19
5
344
32,631
"If I had more time, I would have written a shorter letter" -- Blaise Pascal
8
75
1,141
31,882
the tiny corp retweeted
Here's what the historical open vs closed data here looks like. Would be useful to have volume too (not just %)!
Looks like today may be a record day for token volume % of open models on Vercel AI Gateway: 🟦 Open 78.4% 🟨 Closed 21.6% While spend 💲 usually tells a different story, #3 and #4 today are Moonshot AI & DeepSeek. Adding Z⁠.ai, their combined spend surpasses OpenAI (#2). (Do note that's the spend for inference of the model across providers (mostly in the US), not revenue going directly to the open weight labs.)
31
31
316
52,983
Despite record usage of the models, ZAI stock is down over 2x from its peak. The Chinese markets have a real understanding of the economic impact of AI, the US ones do not. It's very important that the US government doesn't bail any of this out.
43
51
1,127
275,949
tinygrad is the opposite of this. We use LLMs, but only when they make the problem simpler or more clear (which is not what they do by default). The point of code is to be read by humans.
I am done with this shit. It is over. The state of engineering right now is horrible. It has been half a month since I started a new role at a big company. Nobody knows anything here. The specs, code, tests, PRDs, tickets, resolution of those tickets, reports, etc., everything is made by Claude Code. Nobody on my team likes this. They are being forced to ship as much as they can. I have heard multiple times from higher management that pushing code is not a bottleneck, so why are we slow? People are working 12 to 13 hours a day just to press enter. Nobody is reading anything. Humans in corporate are doing nothing on their own. Everyone, literally everyone, from an L1 to an L7 engineer here is doing the same thing. Talk to Claude. There is no sense of victory. Nobody is resolving bugs. In reality, nobody is thinking anymore. Everything is done by LLMs. It is so soul-sucking. I would not mind it, to be honest, if we were at least given the time to check out the code and see what is going where. But no, the goal is to just ship. No matter what happens.
18
40
935
58,773
Escape the underclass with the help of a tinybox!
12
15
372
29,880
In Hong Kong? Come learn about tinygrad!
Tomorrow: HKPUG #101. 🔎 Whoosh3 + BM25 + Python scoring ⚙️ tinygrad GPU programming workshop 🏆 Opik Challenge award Basic Python. Ordinary laptop. Bring your charger. 19 Sep · 15:00–18:00 · CityU (room TBC) Attend + form: meetup.com/pythonhk/events/3… #HKPUG
4
2
46
9,291
We're at 102.5 minutes now. There was no wall at AMD's 108.9 minute time, we can go faster. We have multimachine training working too.
We now have the fastest Llama 8B train on AMD MI350X in the world, and it was run on a machine that was thermal throttling! (we don't have AC, only fans) tinygrad gets 108.5 minutes, top MLPerf time is 108.9 minutes.
4
15
366
27,671
Love to see it. If you are a business with over 50 people, you should have your AI running on premise on boxes you control. Otherwise it's not aligned to you, it's aligned to extract from you.
this is Dario and sama's worst nightmare - A law firm buying Nvidia servers. Latham & Watkins is building an in-house AI stack.. this is US’s second-largest law firm with $8.3 Billion in revenue last year And now it has - - Nvidia hardware it controls - open-weight models it can fine tune - proprietary legal data it is trusted to protect - infrastructure only Latham employees can access A law firm has decades of contracts, negotiations, client context, legal reasoning, and institutional knowledge. They dont want to give all of that away to OpenAI or Anthropic in exchange for expensive tokens.. And on top of that - risk their data being used to train frontier models.. This will happen more and more now.. The big AI labs have no moat.. nothing protecting their largest customers from moving on.. The biggest companies in the world will - - own the compute - own the data - own the workflow - fine tune the model around their business - switch providers when the pricing or quality changes Dario and sama want to make this illegal by bringing in regulation.. and become the AI overlords..
20
54
943
40,690
Breaking news: top 4 US AI companies all want to collude to slow down progress. It's so transparent that even all the Instagram comments have caught on to it.
This very much sounds like a West-only AI cartel proposal and, in fact, if you read Dario's essay - under the very telling "pacing within democracies" chapter - he specifically says he'd need the US government to issue an "antitrust waiver" for this. The plan transparently looks like this: agree among themselves not to compete too hard, get an antitrust waiver to make it legal, and while they're at it, get Washington - under the veneer of "safety" - to kneecap Chinese competitors because only "democratic" models can be in this "safe pacing" club. I'm not even exaggerating: Dario pretty much writes this explicitly. There are 3 steps to his plan and step 2 is literally: "Frontier AI companies within democratic countries coordinate to establish common safety standards as well as limits on the rate of unchecked AI progress" -> the "unchecked AI progress" that would need "limits" is simply AI progress that happens outside their agreement. Another word for it is simply competition 🤷‍♂️ He also explicitly writes that "a key part of pacing within democracies is to keep democracies’ AI lead over autocracies as large as possible." You can't possibly be clearer than this...
16
23
688
29,965