Something super weird is happening with Codex! My Luna reserve usage just crashed from 60-70% down to under 2% in two messages.
Recent updates made it totally unusable. Time to switch back to Claude!?
chatgpt.com/share/6ab59d89-4โฆ
Hehe, so now they update the endpoint without updating the UI?
@wandb released DeepSeek-V4.1-Flash a while ago!
Don't know the pricing or Context, but I know its there (even though they didn't update the UI).
made these visuals with Opus 5.5 for people who get a headache comparing all the benchmarks. (between 6 Sol and 5.5 Opus)
It seems like Opus 5.5 is the clear winner in terms of benchmarks, but 6 Sol is more Cost efficient.
I wanted sharing controls that get out of the way while you read.
Share, copy link and email fold into a reading-progress ring as you scroll. Tap the percentage to bring them back.
Inspired by Appleโs new design.
Hehe, so now they update the endpoint without updating the UI?
@wandb released DeepSeek-V4.1-Flash a while ago!
Don't know the pricing or Context, but I know its there (even though they didn't update the UI).
I beat the @wandb intern to it again!
Z.AI GLM 5.3 Flash is now live on serverless inference on WANDB. 1M Context, vision capable, 0.50$ per million output.
Its also very fast.
@CoreWeave
Sometimes your app needs a label, not a paragraph.
JEV is TypeSafe's model for that. Send it text or JSON plus questions; get back choices, scores and probabilities your code can use.
I'm not affiliated with TypeSafe or JEV.
ALT Abstract graphic featuring a dark, binary-coded sphere with an orange detail and the text "JEV Decisions, not prose."
Think of a support ticket:
Choice: which team should handle it?
Score: how urgent is it on your defined scale?
Noul: does it ask for a refund?
Noul returns 0โ1. A value of 0.91 means the model assigns a 91% probability to yes.
ALT A textured poster outlines three decision-making methods: Choice, Score, and Noul, each with brief descriptions and examples.
The useful part: one request can ask several independent questions about the same data, in parallel.
Your code decides what happens next.
Typed output isn't a guarantee of a correct answer.
Docs: docs.typesafe.ai/prim
itives
ALT Graphic illustrates parallel evaluation of software, highlighting multiple questions processed through a central system, JEV.
Origin, our code hosting platform, is now live.
It's fast, easy to use, and deeply integrated with Cursor.
Get started by syncing your repos from GitHub.
Was given early access to MiMo-X-Pro-Preview by @Xiaomi and its surprisingly a good model!
Testing it some more and will update my finding in this thread.
I beat the @wandb intern to it again!
Z.AI GLM 5.3 Flash is now live on serverless inference on WANDB. 1M Context, vision capable, 0.50$ per million output.
Its also very fast.
@CoreWeave