it's so funny watching other people use LLMs, they'll be writing paragraphs directing the LLM to solve the bug they're working on while i'll paste a log file with the single word "bruh" attached
Introducing Gemini 4 Argon – our new frontier model.
It’s built for complex workflows across coding, enterprise knowledge work, and cybersecurity defense – rolling out today to a set of trusted testers through our Fairwind Program.
stuckfunds.eth.limo finds crypto you bridged but never claimed.
$135M+ sits unclaimed across 84 bridges on over 500k wallets (Arbitrum, Optimism, Base, zkSync, Wormhole, CCTP...).
I got scammed out of $5K a year ago
I DM'd the guy every week for 7 months straight
Today he got so annoyed and finally decided to refund me
Never give up
Did Anthropic nerf Claude Opus 5.5?
The first NerfBench results are live.
We retested Opus 5.5 and GPT 6 Astra against their own launch scores.
Claude Opus 5.5: 99.2% (-0.8% vs launch)
GPT 6 Astra: 102.8% (+2.8% vs launch)
Verdict: No nerf detected.
Opus 5.5's small dip and GPT 6's small increase is within normal variance.
More models and more frequent retests are coming.
NerfBench only gets better as we collect more data.
Which models should we retest next?