MTS @ Wafer, Stanford, Prev Co-Founder at Willow.

why r startup ppl always fighting sm on twitter cant we just all be frens :(
5
11
490
so sad that the year worlds is in NA i have a j*b fml
5
229
carnivore snax is a top 2 human invention
1
2
5
658
drank water today instead of spamming energy drinks + zyns and felt great. is ts the new meta
7
2
19
1,448
growing up is realizing that tps is overrated
14
3
42
6,379
when are the low proletariat mts gonna rise up against the high bourgeoisie mts circles
4
4
15
4,167
cpu lowkey more important in inference than gpu
34
13
408
32,624
Ian Ye retweeted
We’re excited to announce that we’ve raised a $40M Series A! co-led by @MarathonMP and @chemistry, with participation from @Wing_VC, @AMD Ventures, @outsetcap, @fiftyyears, and @ycombinator, and our existing investors doubling down on @wafer_ai. we are also joined by an incredible list of angels, including @JeffDean (CEO, @DiscoLoopAI), @rauchg (CEO, @vercel), @andyfang (CTO, @DoorDash), @kvogt (CEO, Bot), @akothari (COO, @NotionHQ), @eastdakota (CEO, @Cloudflare), @deepgramscott (CEO, @DeepgramAI), and more! Most inference optimization today is manual, service-heavy, and done one-time before deployment. Wafer’s vision is AI that optimizes AI. Wafer continuously learns from your workload’s traffic patterns and performance constraints, then finds the optimal deployment across model, engine, kernels, and hardware. This capital helps us accelerate automating more of the inference optimization loop, giving every deployment the leverage of an expert inference-performance team that continuously finds ways to improve performance per dollar. Thank you to the customers who trusted us with their workloads, the partners who built alongside us, the investors who believed in our mission, and the Wafer team members who have worked tirelessly to turn this vision into reality. 🫶🧇
42
21
310
65,580
Ian Ye retweeted
i'm sad to announce that our recent fundraising round did not go according to plan... we initially planned to raise $18m. we didn't end up getting that number. we ended up raising a $40m series a instead. co-led by @MarathonMP and @chemistry, with participation from @Wing_VC, @AMD Ventures, @outsetcap, @fiftyyears, and @ycombinator, and our existing investors doubling down on @wafer_ai. we are also joined by an incredible list of angels, including @JeffDean (CEO, @DiscoLoopAI), @rauchg (CEO, @vercel), @andyfang (CTO, @DoorDash), @kvogt (CEO, Bot), @akothari (COO, @NotionHQ), @eastdakota (CEO, @Cloudflare), @deepgramscott (CEO, @DeepgramAI), and more. back to the kernel mines
257
56
1,260
184,273
sometimes i be facetiming myself. just to hear what a real one gotta say.
4
4
34
2,683
github actions down as soon as glm53 flash weights out 🤔 #nooticing #china #sabotage #theytrynakeepusdown
1
14
1,355
heard you can go goat milking in the bay, didn't know bron was visiting
2
2
13
1,032
getting back on hinge. mama didn’t raise no quitter
2
12
1,153
mfs be like 'yeah i play smash bros' then pick kirby
1
1
11
1,030
all da modern day goats come from toronto drake, sga, karpathy, me who else
2
20
1,700
Ian Ye retweeted
🚨 BREAKING: these engineers figured out how to serve Kimi K3 on @AMD MI355X at 952 tok/s/node and 118 tok/s single stream! this crushes B200 by 3.8x in aggregate throughput/node and 1.3x in single stream decode + beats B300 on performance per dollar (48 vs 33 tok/s/$) See how in the thread.
24
83
889
444,172
my mexican boss has discovered heytea and the office has not been the same since
1
1
15
3,661
my least favorite part of kimi k3 is that it fits on an mi355 node
1
15
1,997
kimi k3 kinda too big just gonna wait for k4
12
1,296
Ian Ye retweeted
This is a killer stack I just started using Wafer to serve my qwen3.6-27b custom fine tuned llm and it's excellent
Replying to @jsawadd
Potential stack of something like: Hermes from @NousResearch @joinmassive from @jsongrad (web search and more 👀) Gbrain from @garrytan (second brain) @obsdmd (multi-purpose) @ZeroEntropy_AI from @ghita__ha (specialized models) Wafer from @gpuemi & @gpusteve for open source Inference? Delegation ability to Claude Code/Codex ^ some interchangeable, some can be consolidated
24
31
489
105,941