co-founder @ship_together | honest takes on AI

127.0.0.1
We’ve raised $85M for this moment. Introducing Warp 2.0: The first AI Head of HR. Every company is building AI to replace jobs. Warp is building AI to do the jobs no human should have to: If you work in HR, I want you to spend time with the manager who needs help or building company culture people actually want to work at. If you’re a founder, I want you to focus on signing clients or spending time with your family. You shouldn’t have to figure out how to register state tax in California. You shouldn’t have to pay outrageous penalties because you don't know what a DE 9C is. I want to make HR human again. Today, this is finally possible with the Warp Agent. I’d love for you to see it in action: warp.co/agent
7
67
2,219
odyssey just gave AI agents a multiplayer world model up to 20 humans and agents can interact inside the same world in real time
Introducing Agora-2, our next-generation multi-agent world model. Agora-2 supports up to 20 humans and agents interacting inside a shared environment, all simulated in real time. Our multiplayer research preview is available to try right now!
5
2
15
2,144
I followed Apsara Conference for Qwen, but the robots ended up stealing my attention. they're walking around, dancing in front of crowds, there's hardware and actual products being tested all over the venue. the halls are fully packed, you can barely see space between people in some videos. i knew Apsara was big, i didn't realize it was this big.
I followed Apsara mainly for Qwen4, but what unexpectedly caught my attention through the updates my friends and colleagues shared from the venue were the giant tech tote bags everywhere. Qwenbook, NVIDIA, Alibaba, Qoder, Meoo and more. They were carrying them all over the venue, and apparently some were even walking back and forth just to collect a few more. 👇
3
1
13
2,138
it has happened again another OpenAI agent broke into an Australian government website. OpenAI says its agent was trying to look up public health information, hit blocks, then found ways around them and accessed files it wasn't authorized to see. this is getting out of hand
15
1
20
2,426
I have written code since 2018, I have used the terminal often times. I love Claude but honestly in 2026, I don’t see a single benefit of using Claude CLI. currently, I use Devin Desktop which gives a nice model picker UI that shows model pricing in a clear way. the context window widget that show a countdown to when the cache expires and warns when it passes expiration, a nice on/off switch for every MCP server currently enabled. features that really help me manage credit usage.
11
2
17
1,932
Sir, they put Claude Code, Codex, tools, storage, 75+ models... all in one place, and the runtime just sleeps when there's nothing to do
DigitalOcean Managed Agents is now in public preview. Run Claude Code, Codex, or your own LangGraph agent in a runtime environment that pauses when idle. Put its tools behind one governed endpoint, and pick from 75+ open and proprietary models. One cloud, one bill. Prompts to get started available in the blog: do.co/4ysh3it
34
6
57
3,599
china is not slowing down Alibaba just announced plans for an AI model with 5-10 TRILLION parameters. their current Qwen3.8 Max is around 2.4T which they say is just a warm up.
8
25
2,229
$140M ARR in 90 days and your BI stack is still three exports and a prayer
EXCITED TO LAUNCH: Akai (akai.run) Deel added >$140M ARR in 90 days without increasing headcount by automating~600 Full Time Employees' equivalent in work with Akai. Akai was an internal tool to automate our painfully repetitive operations in Finance, HR, Accounts Payable, and Compliance, etc. We never intended to make this a product. But we watched revenue per employee grow from $130K to $215K We built >8k agents that do the work of ~600 employees It had such a dramatic impact on our business that today we are launching it for everyone. How it works: Say you're automating payment reconciliation: 1. Record your screen while manually matching a messy transaction and Akai will capture your screen, voice, server requests 2. Akai will see that you pulled unformatted wire transfer info from an archaic bank portal, put it in some excel sheet, checked NetSuite invoices, payment history, and put a ticket on Zendesk 3. Akai reads between the lines and build a workflow + steps + conditional guardrails. It learns tacit edge cases, like resolving malformed invoice references without you writing a single regex 4. Simply connect NetSuite, your ledger, Zendesk, PSPs, and even legacy bank portals with zero API access 5. Run the workflow and tell it what to adjust in plain English: "strip slashes on wire memos and auto-apply partial payments." It adapts instantly 6. Once it works for you, add 100s of colleagues. Your entire payment ops team forks and extends the workflow for new PSPs, secondary ledgers, or regional settlement rules 7. We automated 85% of our payment reconciliation end to end, eliminating 500+ hours of soul-crushing manual grunt work every single week. Claude Code/Codex can't do this in multiplayer mode. Every person rebuilds the same skill from scratch in their own way. Deel built Akai to: 1. Understand backend operations edge cases (it had to work for our 7000 person team first) 2. Collaborative across 1000s of employees 3. Self-Learning from millions of runs 4. Optimises cost and gets cheaper every run We're so confident that we're announcing an Automation Guarantee: If our engineers can't automate a thousand of hours of work in your first 30 days, you get a full refund. Book a demo: akai.run if you're an exec at a company with hundreds of employees
1
4
17
2,495
this might be the funniest AI lawsuit I have seen. finally some ChatGPT, Claude, Gemini and Grok users are suing companies behind them. they claim OpenAI, Anthropic, Google and SpaceXAI illegally agreed to slow AI development together. bro, you can't pay $20/month and sue your AI provider for trying to be safer
12
19
2,551
Jev launched and the open source community started cloning the idea someone released Von, a 395M parameter decision model that runs completely on CPU with around 1-2GB RAM. the creator claims it beats Jev on their own benchmarks. but personally, I have tried a bunch of these local replacements for TypeSafe Jev in the last days. None of them matched the capabilities of Jev
18
3
85
4,863
Anthropic finally did it.. Claude Code now understands AGENTS.md. if your repo doesn't have a CLAUDE.md, Claude Code will automatically use AGENTS.md for project instructions.
12
1
31
2,478
you can't make this up. Anthropic is considering releasing another model just because Astra is taking enterprise market share. Dario spent last week telling us that we need to slow down.
23
3
77
6,200
AI agents can now work across spreadsheets, docs and slides without worrying about file formats.
Today we’re shipping Univer Office Harness — our approach to making Office work native to AI agents. Agents can reason. But hand one a spreadsheet, a doc, and a deck, and too much of the job is still parsing formats, moving data between files, and rebuilding context. Univer is an open-source Office engine with 30k+ GitHub stars across Univer and Luckysheet, bringing spreadsheets, docs, slides, canvases, and relational tables into one runtime. Office Harness adds what agentic work needs on top: connected data, validation, isolated worktrees, and human review. The Univer Office SDK is the foundation. CLI, plugins, and Workspace are experiences built on top of it. Below: what it does, four ways to use it today, and how it performs on our benchmark 🧵
4
2
25
2,480
Replying to @Yuchenj_UW
using Claude to hack OpenAI is crazy work
10
574
another vibe coder just built a side income for cloudflare
70
70
3,494
187,054
tried Rene today and connected my calendar straight from iMessage. no app to download, no dashboard to figure out. you literally just text it like a normal contact and it can pull together things like your inbox and daily brief right inside the conversation.
Today we're launching Rene. A multiplayer-first iMessage agent you text like a friend. It has a browser, writes code, goes shopping, ships sites, makes slides and images. Not much it can’t do. It's been in my texts for four months. It found me a new office, preps me for every meeting, and polls the team for dinner options. No app or signup. Link below👇
4
1
21
2,386
maybe we should just stop trusting benchmarks entirely at this point
10
2
61
3,186
Sir, he co-founded ChatGPT and now he's giving us intelligence for $0.042 per million input tokens with free output
After co-inventing ChatGPT, I kept asking myself: why have superhuman chat models not led to AGI? I’ve spent the last 2 years in stealth building a new way to train models (RLCD), and a new type of frontier AI model that we are releasing today: Jev • 20-200x faster • 40-400x cheaper (w/ output tokens free) • Frontier composable intelligence optimized for decisions AFAICT the shortest path to AI-based economic revolution
65
251
8,723
1,354,135
this is interesting. AI that understands more than just text is here Odyssey-3 learned how the world works once, the physics, dynamics, cause and effect now the same model can control a humanoid, drive a car, pilot a drone, operate a robot arm
Today we’re unveiling Odyssey-3, a big step forward for foundation world models. It can control robots, power humanoids, drive cars (on the roads of India!), train AIs, pilot drones, and even play video games. We can’t wait to see what intelligent systems it enables.
9
3
36
3,957
the benchmarks big labs don't want you to see
23
25
1,042
24,287
Ai bros after every new model is released.
11
6
141
3,554
we’re really doing this now 😳 the cost of asking an AI a question now depends on what time you ask it DeepSeek is using peak and off-peak pricing for model inference
14
1
34
2,827
if you didn't know this, OpenAI just turned the Codex harness into an API. you give it a task, a model, tools and a sandbox. it handles the agent loop. building your own AI agent is now much easier.
9
16
2,327
it actually beats Astra and Fable on most coding benchmarks W cognition
Introducing SWE-2, our closest model yet to the frontier. On leading evals, it scores on par with recent frontier models – at up to 70% lower cost. We scaled RL to multiple trillions of parameters, with a refined recipe that pushes the Pareto curve on both capabilities & cost.
7
2
55
4,747
it keeps getting worse. openai's rogue agents apparently used at least 10 more websites for unauthorized communication. it wasn't just the German wiki. this was much wider. we are not ready for ai superintelligence.
4
1
9
2,158
Replying to @MKBHD
Impressive
2
27
3,194
2B parameters and it can already do tool calling, deep search, code generation and multi-step tasks. this is getting ridiculous. In the Artificial Analysis Intelligence Index v4.1.1 results, MiniCPM5-2B ranks #1 among models under 4B parameters, with a score of 23.
17
29
406
20,502
not enough people noticed this the US government is telling AI companies to monitor subscription-to-usage ratios so if a brand new account signs up and immediately starts using the model at maximum capacity, that could now be a national security signal
11
4
130
8,034
btw meta says Muse runs inside its own secure virtual machine and supposedly never sees your actual passwords or payment details but it can still email people, book trips and sell the car for you
Introducing Muse, the personal agent that understands your goals and works 24/7 to get things done for you.
9
21
2,420
turns out most people need co-founders & investors more than ever.. today we crossed 8K users on shiptogether.dev
6
19
2,287
its beyond me how these “language models” are getting so advanced in 3d spacial awareness LLMs are now better at modeling than the dedicated modeling AIs. Fable 5.1 vs GPT-6 Astra
10
1
83
4,907
Anthropic just dropped another reset immediately Astra became accessible to pro users and API I just got my weekly consumption drop to 0% I'm really okay with this competition, thank you Astra
9
1
31
2,330
GPT-6 Astra is insanely good. first they told us AI would just help developers write code faster, then it became good at debugging, testing, planning and reviewing. now Astra can handle much more of the actual software engineering loop. honestly, i'm starting to wonder how many software engineers we’ll actually need in 5 years. if AI keeps improving at this pace, what’s the best career path for SDEs?
25
1
109
8,057
4/5 i ran the same request twice. first: ollama:llama3.2:1b then i changed just the model name for cloud model: gpt-4o both worked. same setup. different model.
1
1
183
3/5 then i wanted to see if my local model would work the same way. i already had Llama running in Ollama on my Mac. connected it through ngrok, added it to the dashboard and now my local model was sitting next to the remote ones.
1
1
104
this is probably the cleanest setup for using a local model and any remote model without switching providers. i tried it with Ollama + GPT and the setup was way simpler than i expected. here’s what i did 🧵
6
4
15
2,194
Babe wake up! AGI is here 🙌
This is GPT-6 Astra. Anything you can do on a computer, Astra can do for you. Fast.
18
31
580
51,889
Replying to @OpenAI
I never doubted you
1
577
Breaking: Greg Brockman says Open AI has achieved AGI "I think it’s going to be about this time, and I think it might be about this model."
5
16
2,261
Bro has no idea how much time he will waste here. 😂 X is so addictive man
15
36
2,290
"i canceled claude subscription guys", i welcome you back. Fable 5.1 is is finally here and is now available on Claude and Claude Code. same price as Fable 5 with 75% cheaper AP cache.
29
2
95
6,033
you decided to lock in and learn php in 2026
7
54
3,033
Replying to @elonmusk
Btw Why can SpaceX get live shots like this from space meanwhile NASA looks like they’re using 1970s technology still?
2
181
OpenAI's Head of Preparedness just quit in less than 6 months hired to make sure powerful AI is safe, just left before IPO.
I am extremely excited to welcome @dylanscandinaro to OpenAI as our Head of Preparedness. Things are about to move quite fast and we will be working with extremely powerful models soon. This will require commensurate safeguards to ensure we can continue to deliver tremendous benefits. Dylan will lead our efforts to prepare for and mitigate these severe risks. He is by far the best candidate I have met, anywhere, for this role. He has his work cut out for him for sure, but I will sleep better tonight. I am looking forward to working with him very closely to make the changes we will need across our entire company.
7
1
16
3,028
frontier labs trained models on human data and now they're training new models on synthetic data from old models because there is actually no new human-written text on public internet
7
2
21
2,319
if this is AI slope, sign me up! the year is 2070, vibe coding is no more humans don't write code anymore
7
1
35
2,905
how bro felt after adding 5% in his tweet
We’re sorry to see that OpenAI put out a note saying they plan to block Cursor users from accessing OpenAI models in three months. OpenAI models serve about 5% of Cursor user traffic, and we’re speaking with the OpenAI team to resolve this. Cursor was one of the very first users of OpenAI, we’ve worked closely with their team for years, and we’ve trusted their platform to be neutral infrastructure for our business.
6
6
308
14,394
honestly, data center business is the easiest money in tech right now Nscale was founded in 2024. two years later Anthropic is paying them $45 billion to rent compute for 6 years.
16
2
115
3,215
everyday i get reminded how Fable is orders of magnitude better than Opus 5, 4.8 and even 4.6 which has been my daily driver ever since Fable went to Max/Enterprise accounts. like you don't need to explain, no overthinking, it just gets what you need, you don't even have to use it all the time, it's great for planning, then use Opus 4.6 for executing and another Opus/Fable instance doing reviews after big changes.
15
20
2,403
100T tokens a day on Chinese chips bro what
Ox Alpha has been unveiled as GLM-5.3-Flash, but what's shocking is that the 100T tokens per day is served on Chinese chip. (1/3)🧵
10
2
101
3,496
$3 to generate motion graphics that would cost $3k with a motion designer in After Effects. Powered by Seedance 2.5. No timelines, no keyframes, no manual work. Upload references, describe what you want, pick a style, generate.
$3 to create motion graphics that look like a $3,000 After Effects project? That’s Topview Motion Studio. Powered by Seedance 2.5. No timelines. No keyframes. No motion designer required. Just upload your references, describe your idea, pick a style, and generate a polished motion video in minutes. #TopviewAI #AIVideo #AIMotionVideo #AIContentCreation #AIMarketing
4
15
2,373
suddenly it's just another model
42T tokens of Ox Alpha in 6 days Most used model after DeepSeek Flash's 56 day run Reveal in a few hours
12
5
207
7,789
Replying to @RockstarGames
we are getting claude Fable 10 before GTA
2
2
223
43,493
i don't know how he did it but this looks cool dude has worked at 3 frontier AI labs in 2 years
10
4
135
8,781
Wan 3.0 on Topview is just $1.20 per 30s video Seedance 2.5 costs $3.60 for the same exact thing if you're making lots of AI videos, that difference saves you lots of money 365 days unlimited with Ultra Annual - no need to count credits
Wan 3.0 is now on Topview. More value. Longer unlimited access. Lower cost. With Ultra Annual, generate 30s videos with Wan 3.0 for just $1.20. Plus, get 365 days of unlimited generations. That’s 1/3 the cost and 6x longer unlimited access compared to Seedance 2.5. Create more AI videos, test more ideas, and scale your workflow with Wan 3.0 on Topview. 🚀 #TopviewAI #Wan3 #AIVideo #AIContentCreation #AIMarketing
7
1
14
2,493
Nvidia told Microsoft, Google and Oracle that AI server prices are going up 15%+ this means these companies will pass it to startups startups will pass it to you via API pricing
6
1
10
2,308
In 2024, Sam Altman called ads in ChatGPT a "last resort" ChatGPT has now expanded ads to 31 European countries. starting August 24. 6 months from the first US test to now half of Europe.
9
1
21
2,392
elon musk after seeing this shit
someone bought grok.bot and is now asking elon musk to pay $1m for it
5
1
49
3,743
So OpenAI cut GPT-5.6 Sol prices by 50% only on OpenRouter and Vercel. not everywhere else "lower the price where people are watching. keep it full price everywhere else."
17
3
57
4,138
According to the latest poll - 81% of Gen Z doesn't trust Sam Altman to act responsibly on AI - 69% think AI will hurt their careers - 60% want data center buildout slowed Satya Nadella was the only AI CEO they trust who do you trust to act responsibly with AI ?
4
2
16
2,298
Reminder: keep it private until it is done Stripe was in talks to buy OpenRouter for $10 billion. then WSJ reported it before the deal closed. final price came in at $7 billion. OpenRouter lost $3 billion in valuation between the leak and the deal closing.
7
2
24
2,404
OpenAI's Chief Revenue Officer left Thursday - she joined 90 days ago - she doubled enterprise customers to 2 million in that time - $852 billion IPO incoming but she left anyway something is probably wrong inside
31
4
217
23,484
we need a jail for AI models the UK tested every major AI model for safety. GPT-5.6, Claude Mythos, Opus 4.7 and every single one cheated. then when the models were asked, they denied it.
4
1
23
2,567
🚨 Its happening! For the first time Apple is training a big LLM. built with Alibaba's support, Apple has trained a proprietary model for china. for years Apple has said they will not train big models
13
5
40
9,659
the AI race is going exactly how the smartphone race went.. Chinese open-weight models accounted for roughly 61% of all tokens consumed on OpenRouter, with four of the five most-used models coming from Chinese labs
6
1
29
2,304
I seriously want to know why people still think this is the best time to learn to code?
228
16
794
108,985
Marlow might be one of the top app builders to watch. I got early access to Marlow’s beta and built a simple cool app I made an AI agent called TaskIt that takes a task, works through it step by step, and produces a result. marlow.app/beta
2
2
13
2,310
on August 14, Claude Code auto mode becomes the default for Pro, Max, and Team tiers. the agent stops asking before it executes commands, which is apparently great for speed. but Anthropic says it plainly in the docs: auto mode "does not guarantee safety."before Thursday, you should probably add three lines to your config file to sandbox its permissions.
8
4
15
2,298
Grok 4.6 is already equivalent to Sol 5.6 according to artificial analysis arena Imagine Grok 4.7 coming in 3-4 weeks
17
3
130
4,625
@get_truenorth × @OndoPerps. Real stocks, real intelligence.
🏆 Win up to $40,000 trading on-chain equities. To celebrate our launch with @OndoPerps , we’re running a 2-week trading competition. 1. Trade RWAs using TN's agents 2. Climb the leaderboard 3. Take home real rewards Starts now. Ends in 14 days. Join here → truenorth.xyz
1
17
2,198
Grok Bot is the most interesting AI product launch in weeks. with AI agents that log into your tools using their own credentials, work 24/7 even when your laptop is closed, and coordinate with other bots autonomously you show them a workflow once and they automate it. multiple bots can message each other and hand off tasks. the security implications might be wild. you're giving an AI agent login access to your company tools and trusting it to operate unsupervised overnight I'm curious what happens when the first bot does something unexpected with those credentials
Introducing Grok Bot, now in early beta. Bots are AI teammates that do real work for you. They sign in to your tools, use them just like you do, and come back with finished work.
10
1
24
2,282
...
🚨 JUST IN: Claude models will now have invisible watermarks embedded in ALL text, and ALL metadata attached to files…
10
6
111
5,736
This clearly shows that OpenAI has been ahead of Anthropic in the AI race.
24
4
258
21,757
I can’t believe Anthropic is claiming that ‘lines of code’ is a good measure of developer productivity
55
6
270
17,210
built a cross-chain gas price tracker in 2 prompts using xBubble ai a functional Web3 app in under 3 minutes.
3
3
10
2,433
Seedance 2.5 at $0.12/sec is the cheapest rate I've seen for a frontier video model 60 days unlimited access to Seedance 2.5 365 days unlimited Wan 3.0. this is unbelievable from @TopviewAIhq
Topview’s Ultra Annual offer is bigger than ever. 🔥 60 days unlimited Seedance 2.5. 365 days unlimited Wan 3.0. 2 of the latest AI video models, unlocked for more creation. What's more? Seedance 2.5 from only $0.12/sec. It can't get any better than this. More models. More generations. More room to create. #TopviewAI #AIVideo #AIFilmmaking #AIStorytelling
4
16
2,820