The Lure of the Underground by Alfred Leete (1927)
1
2
86
Lukas Levert retweeted
It'll soon be seen as deeply irresponsible to have critical code reviewed by human eyes alone.
280
245
4,586
215,683
Lukas Levert retweeted
the rain had settled in, the afternoon had nowhere particular to be, naturally you set the @p0 web agent w/ opus 5.5 on the trail of a garden painted in 1874. off it went, through a catalogue raisonné or two, a few museum archives, and several increasingly excellent footnotes. whether lunch at le grand véfour was involved remains between it and the maître d’.
2
7
411
Lukas Levert retweeted
Replying to @p0
@p0 natively supported in @LangChain ✨
Introducing Managed Deep Agents 0.8 ✅ User-owned credentials + memory let agents work off correct permissions & user-specific context. ✅ HTTP channels for bringing agents into internal tools ✅ Native @slackhq file transfers ✅ Built-in web search via @p0 langchain.com/blog/langsmith…
3
15
2,438
Excited to partner with @p0 on Managed Deep Agents! You now have parallel built into your agents so your agents search the web fast and efficiently! @travers00 @hwchase17 @VictorMoreira16
6
18
43
19,366
Lukas Levert retweeted
It is becoming table stakes to offer private data alongside search. To me, this marks the beginning of a new era of the web, where the most valuable information no longer sits siloed in a database, but rather sits on equal footing with web data, both discoverable and paid.
Today we're launching Data Connectors in Parallel. Your agents can now work with specialized third-party data through Parallel's best-in-class agentic web research APIs. Our first partners are @AlliumLabs, @useapolloio, @Baselayerhq, @CarbonArcAI, @crunchbase, Faraday AI, @harmonic_ai, @MiddeskHQ, @Similarweb , @particle_news, @pensa_systems, and @Polymarket. We're also launching free, opt-in access to biomedical sources including PubMed, ClinicalTrials.gov, and ChEMBL.
2
7
66
12,258
Lukas Levert retweeted
Today we're launching Data Connectors in Parallel. Your agents can now work with specialized third-party data through Parallel's best-in-class agentic web research APIs. Our first partners are @AlliumLabs, @useapolloio, @Baselayerhq, @CarbonArcAI, @crunchbase, Faraday AI, @harmonic_ai, @MiddeskHQ, @Similarweb , @particle_news, @pensa_systems, and @Polymarket. We're also launching free, opt-in access to biomedical sources including PubMed, ClinicalTrials.gov, and ChEMBL.
9
22
101
52,443
Opus 5.5 is so good - especially at writing. It's crazy how fast our perception of "good" changes. Let us know what you think of the model leaderboard. We're always trying to improve it.
Claude Opus 5.5 takes the #1 spot on Parallel's Search Capability Leaderboard, with a +4.3 gain over the second-highest score from Opus 5.
1
10
560
Everyone on this app fails MentalHealthBench
We’re demonstrating how frontier models have continued to improve in realistic mental health conversations with MentalHealthBench. This new open benchmark was built with input from more than 80 mental health clinicians. We’re releasing it openly so other researchers can examine the methods, run their own evaluations, and build on the work. openai.com/index/introducing…
4
315
New: Opus 5.5 on the parallel search capability leaderboard. More evals are in progress and coming 🔜 for 🌞 and 🌚
Claude Opus 5.5 takes the #1 spot on Parallel's Search Capability Leaderboard, with a +4.3 gain over the second-highest score from Opus 5.
1
4
380
Per your toddler
honestly not to waste your time. I saw you're running product comms at Stripe and being a business man per your toddler.
1
199
Sf is crazy because when you talk about “fulfilment” people think you’re talking about supply chain logistics
2
77
In what world is this breaking news
BREAKING: Pod 6 is live. The sixth generation of our intelligent sleep system, redesigned for every kind of bed. ✅ The new Hub is small enough to fit under your bed frame ✅ More powerful than ever ✅ With 9x more sensors for enhanced accuracy ✅ In new Solo sizes, at our lowest starting price yet This is our most significant hardware launch to date, and our most accessible. Available now at eightsleep.com
1
234
Lukas Levert retweeted
What an insane day in AI. The frontier models just became substantially cheaper, with the Opus 5.5 price cuts, and now with GPT-6 Sol and Luna dropping token prices by 50%. The rate at which the cost per task (on a like-for-like basis) drops in AI is unlike any other type of technology in history. And every time the cost of AI drops, the use-cases you can deploy agents against dramatically increase. This is Jevons paradox applied to agents. These improvements will directly lead to broader diffusion of AI in the economy as we can use agents to process all of our data, scan our code for security issues, read through all log data to make decisions, have agent swarms in workflows, and much more. The cost of tokens is directly correlated to these use-cases being opened up at scale.
Please welcome GPT-6 Sol and GPT-6 Luna to the GPT-6 universe. GPT-6 Sol and Luna build on the advances behind GPT-6 Astra, bringing much of its strengths into faster and more affordable models to support work at scale. We’ve also made caching and inference more efficient, and we’re passing the savings directly to you: 50% lower API prices for Sol and Luna compared with GPT‑5.6 promotional pricing.
107
84
749
141,155
NEW MODEL DAY
2
90
Berd desperately needs a mobile app if it wants to compete with other agents @jack
1
118
Lukas Levert retweeted
The life of a button... 🔊
119
1,412
12,961
1,492,489
Lukas Levert retweeted
Pranay Reddy Samala helped build @p0's deep research system, the one that set the state of the art on BrowseComp. His Fully Connected talk is about why the benchmark isn't the point. Sept 29–Oct 1, SF. Register: utm.io/usfdT
1
3
20
7,738