Chasing the horizon of tomorrow's ✨ | Business, AI, Apps & Open Source | Tech Expert | Ambassador @cognition | DMs open for collabs and opportunities.

Pinned Tweet
Excited to announce opethon.com, a 30-day open-source hackathon built around solving real problems in three areas: • Health & Care • Learning & Education • Accessibility & Inclusion The idea is simple: find a real problem, build a complete working solution, and ship it open source. AI and other tools are welcome. Just disclose what you used. We’re already at $30,000+ in prizes, including 100 subscriptions worth $200 each and 100 subscriptions worth $100 each. I’m currently bringing sponsors on board, but Opethon is happening either way. The judging panel so far includes @serg_vecher, @NoemiTitarenco, and me. Both bring years of experience across tech and other fields, and they’re people whose judgment I trust. We also have two additional judging seats reserved for Diamond and Gold sponsors. Sponsors will also be able to host a live session for all participants, with an expert from their team showing builders how to use their product during the hackathon. Depending on the sponsorship tier, sponsors can also get dedicated posts on my account, placement on the website and throughout the event, and representation on the judging panel. I’m expecting 500–1,000+ participants once sponsors are finalized and the official dates are announced. If you handle sponsorships and want to see the available packages, my DMs are open.
61
68
230
23,138
Da7em retweeted
Is there anything unethical OpenAI hasn’t done yet?
#keep4o #OpenSource4o #BringBack4o 🛑 ATTENTION 🛑 Think you're using GPT-5.6 you pay for? Think again, because the user interface on the screen displays "GPT-5.6", but it's a complete LIE. In reality, OpenAI is secretly routing your requests to GPT-5.4which isn't even officially available in the app anymore. It happens nearly every day, sometimes back to back, and YOU WILL NEVER KNOW unless you open your browser's dev tools and check every single response payload. This is direct, automated routing deception. They save money on compute while leaving the 5.6 label slapped on the UI to gaslight you. default_model_slug: gpt-5-6-thinking: This confirms you requested and are paying for the 5.6 model. resolved_model_slug: gpt-5-4-auto-thinking: This is the model the system gave model_slug: gpt-5-4-thinking: This confirms the actual model being used is 5.4. This is a scam and deception.
21
17
181
15,600
Barely a month ago, OpenAI was bragging about having all the compute in the world. They talked big about bringing AI to everyone and mocked the competition. Look at them now. They managed to make their $200 tier completely useless, and now they have the nerve to push a $500 plan. Plus, despite their name, everything they make is closed-source. Turns out they're out of compute. All that PR hype about Astra being 50% more token-efficient was pure BS. They clearly don't know what they're doing. Meanwhile, the tables turned completely. Anthropic is offering a better experience, way more generous token limits, and smarter, faster models like Opus 5.5. It's a classic case of hubris catching up with a company. The market moves fast, and a single month is all it takes to flip everything.
3
5
56
1,446
Opus 5.5 just built a piece of art nicer than this!
Opus 5.5 made this after 9+ hours! A story about a living website that remembers you and has feelings about your interactions. Every moment feels alive, and the ending is magical.
6
1
19
1,505
Da7em retweeted
Fable vs Astra!
27
3
73
7,042
This test is shocking for a few reasons: 1- Opus 5.5 was better than both of them. 2- Fable on High effort was better than Max. 3- Astra researched to imitate, while Fable built everything from scratch. In my opinion, Astra did better on the ship itself, but Fable delivered a much more cinematic scene with far better lighting.
Fable vs Astra!
5
1
27
3,167
Da7em retweeted
benchmarks will always look good on paper till you actually test out the model yourself jus imagining getting SWE-2 for free on devin been using it alongside opus 5.5 but for execution it's such an underrated from by @cognition @Da7_Tech cooked on this one da7tech.com/da7em-bench/
Introducing Da7em Bench. You can now check detailed model results and compare them side by side to see what fits your workflow best. I'll keep updating it with methodology details and adding more models as soon as I get more compute. da7tech.com/da7em-bench/
1
1
4
243
Da7em retweeted
Kimi K3 is still an absolute monster despite coming out July 16, 2026. I can’t even imagine how good the next one will be. Also, benchmark ratings came from this site by @Da7_Tech. An actual way to see model benchmarks without “benchmaxxing” from companies. da7tech.com/da7em-bench/
1
8
161
Am I dreaming?
Claude Code will now try to find a graceful stopping point when you hit your 5-hour limit mid-task, instead of cutting off mid-edit. It gets a small, fixed allowance pulled from your weekly limit to wrap up what it can.
25
1
234
11,470
I hope you like it 😁
Introducing Da7em Bench. You can now check detailed model results and compare them side by side to see what fits your workflow best. I'll keep updating it with methodology details and adding more models as soon as I get more compute. da7tech.com/da7em-bench/
4
1
39
2,065
Introducing Da7em Bench. An independent benchmark for AI models, built on real client work. The first of its kind in the world. How it works: Every model runs about 200 real tasks in each of 12 areas: reasoning, research, planning, delivery, persistence, accuracy, honesty, acceptance, engineering, taste, writing, and communication. Each model is tested across several harnesses, both official and neutral ones (Droid, Hermes Agent, Devin, Cursor), so no single harness decides a model's fate and the results reflect the model. Scoring is 1 to 5. A 5 means the work was accepted as delivered. Middle scores mean it needed revision. A 1 means it failed. The bar is professional work. Every result is judged against what a paid professional would have delivered for the same brief. Tasks stay private so they can't leak into training data and inflate future scores. The goal isn't one more leaderboard. It's helping you pick the right model for your kind of work. A model that leads in reasoning can still fall behind in writing or design taste, and the radar charts show exactly where. This is v0.1. It will keep evolving with harder, market-relevant tasks and with new models as they ship. A few popular models (Opus 5, GPT Luna) aren't included yet because I haven't run enough tasks on them to score them fairly. Da7em Bench is fully independent. No sponsors, no vendor relationships. I've paid for every run out of pocket, thousands of dollars so far. That's the whole point: an honest, neutral look at what these models actually do on real work. Full scores and the framework are in the images below.
131
53
393
51,554
Enjoy it.
Introducing Da7em Bench. You can now check detailed model results and compare them side by side to see what fits your workflow best. I'll keep updating it with methodology details and adding more models as soon as I get more compute. da7tech.com/da7em-bench/
1
1
90
Introducing Da7em Bench. You can now check detailed model results and compare them side by side to see what fits your workflow best. I'll keep updating it with methodology details and adding more models as soon as I get more compute. da7tech.com/da7em-bench/
23
10
95
4,865
Da7em retweeted
Sorry, but this is absolute nonsense. Sharing information and helping people when you can is normal. Sure, sometimes you have to say no, and people should respect that. But she asked what camera you use, not for private information or some trade secret. If you’d just answered her, none of this drama would exist, you wouldn’t have written this post, and none of us would’ve wasted our time on it. The real reminder is simpler: you reap what you sow with people. Give generously, and eventually it comes back to you. That’s how human relationships work.
5
3
50
4,173
Today is a big day!
19
2
85
5,196
Above all, do not lament my absence. For in my Spark, I know that this is not the end. But merely a new beginning. Simply put, another Transformation
2
1
22
1,400
Da7em retweeted
I ran the Ship Test on both Fable 5.1 and Astra 6 at Max effort. Burned a ton of tokens, but the comparison was genuinely interesting. Posting it today.
9
3
78
3,484
Da7em retweeted
Starting today, I won’t use Meta models, review them, or talk about any of their models or tools until @alexandr_wang ends this disgrace and apologizes to everyone for the moral depravity of what he’s doing.
guys literally only want one thing and it’s fucking disgusting
38
3
82
13,121
Da7em retweeted
Just like Anthropic shook things up here with Opus 5.5, Google can do the same with Gemini 4 Pro.
Damn, im in love with Opus 5.5. Its the best model ive ever used, has such a good taste, is comparatively quickly, smart and finally precise instead of verbose. As good as Opus 5.5 is, I'm incredibly excited for Fable 5.5. After switching completely to Codex, I now have to go back to Anthropic. oh and btw. its crazy how fast the vibe has changed. Now all I read on X are posts about how great Opus tastes and how OpenAI absolutely has to catch up. One day that changed everything.
19
3
92
6,245
Da7em retweeted
I wouldn't use Astra even if it were free.
34
3
127
7,717