Product & Engineering @BetterStackHQ Your agent should be doing dangerous things

Prague
Simon Let retweeted
I don't think writing a spec and splitting it into tickets is how you should use agents since Sol 5.6/Opus 5. Just tell the agent what you want to do and it will do it. If the task is large, discuss with the agent how to split it in several shippable PRs and do it one at a time.
mattpocock/skills v1.3 is out! - /pr (new) writes easy-to-read PR bodies, showing hard evidence that the changes work and assessing merge risk - /implement-spec (new) takes in a spec and tickets, and implements them with subagents - CONTEXT.md renamed to GLOSSARY.md - /retro (new) reviews recent coding agent sessions and suggests repo improvements /retro, especially, feels like a huge upgrade. Enjoy!
104
15
556
102,787
Simon Let retweeted
We unfortunately have decided that we cannot continue providing access to our models through Cursor and are ending our partnership. It boils down to trust and we’ve asked that this takes effect on November 12 to give you some time to plan. Many have used the GPT models through Cursor and here are options we know should work in the future: - We will continue to allow using your own OpenAI API key and similarly will continue to provide access through our IDE extensions for Cursor. - We will keep working with the broadest range of tools and harnesses, some of which are OSS, but also many many closed-source ones. We are as committed as ever to continue supporting developers and the flourishing ecosystem of tools, harnesses and products. We will also continue to invest in our own open-source initiatives and believe in broad optionality for developers. You can read more about our decision in the blog: openai.com/index/our-decisio…
We’re ending our partnership with Cursor following its acquisition by SpaceX. Under our proposal, Cursor’s direct access to our models would end on November 12. We know that the people most affected by this decision are the developers who rely on OpenAI models in Cursor. We care about their experience in this transition and we’re ready to go above and beyond to support them. openai.com/index/our-decisio…
1,472
600
10,013
4,301,061
extremely eye opening how much worse every model performs in chat vs in code no agetic harness poor prompts -> recipe for disaster
40
every time i use the chat in claude macos app
53
Simon Let retweeted
I've started a new company: @superlogical! We're going to begin by building a terminal multiplexer. The entire vision is much larger, but the multiplexer is the foundation. Sign up for the newsletter to get beta access and devlogs (product updates only I promise). superlogical.com/
552
521
9,162
1,188,764
opus 5 is hard to trust i'm scared to let it run unattended it's so confident probably smarter than 4.8 but it makes incredibly stupid mistakes
1
1
109
Simon Let retweeted
why does someone still uses @datadoghq in 2026 - there event based pricing is absolutely ridiculous. we got ~500$ bill for a month, where we recorded ~168 million of logs and it was just about 40gbish. I shifted our whole logging infra to @BetterStackHQ , took <1hr, our last month bill was ~45$. and betterstack mcp is such god tier, datadog mcp is also bad.
12
9
35
3,793
Simon Let retweeted
There's no full-window "Chat" mode in the new ChatGPT app for Mac. Don't upgrade.
1
2
5
507
Come chat about observability, eBPF, automated AI incident resolution, and everything in between.
We're heading to WeAreDevelopers World Congress Europe! If you're in Berlin on July 9–10, stop by Hall 2, Booth 27, to see a radically better approach to observability for modern engineering teams. Meet the team, catch a live demo, and grab some swag. #WeAreDevelopers #wearedevs #observability
1
66
MCP's still carry SO much friction i want to 1) ask my agent to "integrate tool X" 2) hand it a token -> start using X no session restart no /mcp command that my agent can't use directly this is why i just run with API's 99% of the timef
1
40
RT @GuidedHacking: 🏗️ Reverse Engineering DirectX Turn an IDA Pro database into a functional project. We skip generic advice and show how…
12
your product doesn't send enough emails Linkedin will send you 60 emails in the first month when you create a fresh account
1
46
Simon Let retweeted
The lovely folks at @JuicedataInc wrote a guide on monitoring JuiceFS with @BetterStackHQ
1
2
4
441
Simon Let retweeted
Under-discussed: how good agents are at debugging issues in production. Everybody's constantly talking about how well they write code, but give an agent access to the gcloud CLI and a screenshot of a graph and, good god, will it go.
48
8
350
48,071
Simon Let retweeted
We have moved from incident.io to @BetterStackHQ for our incident reporting! • status.scholarxiv.com The status page also reports more monitors; platform, API, AI-Chat, Sandboxes and Subscriptions with status history and chart. We've also made our health check API public for transparency! scholarxiv.com/api/health will respond with a comprehensive result! We've made this change to enable for more in-depth monitoring and incident reporting for our users and our on-call team!
1
4
23
886
Simon Let retweeted
Replying to @stevenssup
claude is so good i wouldnt call it vibe coding anymore
9
1
65
4,804
ah it's the "your mac will now attempt to restart every night" season
every morning, half my apps closed and ghostty sitting there with a dialog saying "nah not today, mac"
65
what browser are yall using? @claudeai keeps closing my chrome so I need something else
23
unacceptable 2 frames of broken clock app icon every time I open the app where is apple's attention to detail 👀
1
46