I'm happy that the 6-series improved on cost, but saying it's a flatly better model isn't completely accurate
Here's an example of a benchmark where it underperforms compared to the prior versions
Apple needs to hire more app reviewers fast. On web you can deploy in minutes, on iOS it takes multiple days, depending on when you get your review.
Increasingly, one bottleneck to shipping quality updates becomes review time..
would love for you to try Ness, it's an all-in-one AI Health coach with nutrition built in. You can text it photos too if you'd like
happy to send over a lifetime code so you can try it out!
Found a fun way to incorporate Jev into Ness!
Whenever you send a message, Jev determines intent and will select an emotion for Ness to display
@typesafeai
Would love to try this out, how is the cost/query so low? Is search data being trained on/sold?
ps. the signup/login screen doesn't seem to render properly on Helium
Competitive pricing for the lower-end gemini models (3.1 flash-lite, 3.5 flash-lite) would be a plus
currently get outclassed by 5.6 luna and V4 flash 0731 on everything except vision and long context recall
that said, Fable and Kimi are both relatively token hungry
ideal token efficiency would be 5.6 sol or grok, those are very efficient for their capability
you should try Performant, refreshes take <1s
we do a history computation step which runs in the background and does take some time, but once that completes, refreshes are blazingly fast!
Testing out a fully local version of Performant Chat, running Gemma 4-E2B. One thing that's concerning is moderate/high hallucination rate and tool call error rates.
Should we ship this as an option to Performant?