I'm worried in their push to catch up on the SOTA front, it's going to lose that natural sounding touch it currently has.
Meanwhile Claude and Astra like to couch all their agreements with caveats and provisos.
Sometimes that's what being smart sounds like.
Excited to try this out! Shame on Google for not releasing Gemini 3.8 for Google AI Plus users yet, though.
As opposed to what, them not having it and burning money that isn't their instead like openai and anthropic? At least Google is feeding itself instead of having to create a bubble to stay alive
All audio generated by our AI products is watermarked with SynthID. This imperceptible watermark is woven directly into the audio output, ensuring AI-generated content remains detectable to help prevent misinformation. For details on our approach to safety and responsibility, review the model card.
As PrimeTime said; these are the guys that invented the 'T' in 'GPT', that deployed their first TPU in 2015, that is using billions on AI - and they are beaten by 300 people startup named Moonshot AI even. People are going to write books about this complete fumble.
Given the very high margins on inference, once volume is large enough the other can also start printing enough money.
My advice is to listen less to brainrot 'influencers' that optimise for engagement through sensationalism.
They have "unlimited" resources and has researched AI since the very beginning - PageRank is a form of AI even. And still, Gemini is behind Claude, GPT, Grok, Muse, GLM, Kimi and is maybe on par with DeepSeek?
As I said, it is embarrassing.
However, the "Extended Thinking" should be renamed to "Slightly Extended Thinking". Considering that it's the maximum thinking option for Gemini Flash in the chat UI, it doesn't actually think a whole lot, leading to an uncomfortably high number of incorrect/poor replies.
Nothing but constant errors with cryptic messages.
Let me tell you unlike every other mentioned model Gemini 3.8 Flash trial had to be reverted the same day. Instead of simply delegating tasks it would invent additional requirements and implementation details it knew nothing about and no amount of convincing not to do it would work. That's the first time a model failed on me so spectacularly despite having practically same Artificial Analysis Intelligence Index as another model that just worked (and higher than working DS Flash).
The reason I think it is relevant is: Live is likely even stupider model in every way possible (except hearing better than separate STT). So beware using it for agentic scenarios.
Lately it became load-bearingly-reality-difficult to not only read, but to comprehend the Claude output
On my TODO is try and run all of the analysis pipeline in dense "machine speak" to save on tokens and just let Gemini sort it out at the end.
Big rich companies take on debt for reasons that are sometimes inscrutable from the outside. Recently, they have been borrowing for ~5%, about a half point above what the US government gets for 10-year Treasuries.
Apple has been financing operations with debt for a number of years as part of a complex optimization plan.
No, Google is not broke.
I can't think of a better general purpose model than 3.8 flash right now. It also writes more naturally than the other big models too.
However, if I want it to DO something then Gemini is in absolute last place. I don't trust it for anything more than renaming files that I don't care about very much or extracting data (though it's too expensive for data extraction at scale).
No one is behind grok. It literally has "be funny and irreverent when appropriate" (whatever the hell "when appropriate" means for them) baked into the system prompt. To me, that is all you need to know about how useful it is.
No serious people use it and the numbers bear it out tbh. It has the smallest market share of the "big companies" for a reason - and it's by a very, very large margin (~2.5% last I checked).
And they're making money doing it.
Perhaps they don't have the best coding model right now (although 3.8 flash is arguably SOTA at some benchmarks), but is that the be-all and end-all of AI? Only coding matters?
varispeed•1h ago