I said for a few years to many a downvote on HN, everyone wants AI, nobody wants to pay the true costs, the AI race will turn into a "race to the bottom" that is, who can give you the most compute for the lowest cost, and still remain profitable?
Personally, In SWE, i think the industry has made a grave mistake with the agents and we're just one big Catastrophe waiting to happen. I do think that there is very real value when software engineers use these tools as something akin to exoskeletons that allow the human to do more, rather than just fully replacing them. However, I'm finding more and more that companies are slop shops and just attempting to automate all of their software engineering. That will certainly end terribly. I hope we are not cannon fodder.
For what? There are some things I want AI for because it does it well. There are some things I don't want AI for because it just makes a mess (hallucinations). Maybe the next AI will be different and we will have the conversation again.
Indeed, it kills our planet, our culture and our economy extremely well. Oh yes, and some code monkeys enjoy that it can make computer code on the side.
> I said for a few years to many a downvote on HN, everyone wants AI, nobody wants to pay the true costs, the AI race will turn into a "race to the bottom" that is, who can give you the most compute for the lowest cost, and still remain profitable?
I keep seeing this but this line of thinking doesn't make any sense. What does it really mean?
There are expensive models that increase the probability of you doing your task under a lower cost. That means you can't use Gemma for coding your new compiler - it would just be overall costlier.
Heavier models are cheaper at more complicated tasks because they use fewer turns and fewer mistakes.
Cheaper models are more likely to be cheap at less complicated tasks. Like if you just ask Gemma "Hi" it would probably be cheaper than asking Opus.
So what does this statement really mean? People don't want to pay the extra for a more costly model? Why wouldn't you? It reduces your overall cost!
Because real Fable usage starts at $20/month, and has oppressive usage limits even at that (ridiculous) monthly price.
Compared to my $3/month GLM-5.2 subscription, I have never felt like I was leaving capabilities on the table by refusing to cough up $20 for 15 minutes of Fable use per day.
This is the wrong way to look at it. If you have a complicated task , you can solve it for cheaper if you used Fable. It will use fewer turns to achieve the same result.
You can solve it for cheaper if you use GLM but if you are involved in it more, but that defeats the purpose.
The point is that there aren't many complex tasks were fable delivers a significant value increase over cheaper models.
Single prompting a very complex tasks is rare even on frontier models, because it can be done successfully only for specific situations (e.g. you have a very strong verification step the model can iterate on).
Most of my everyday usage is for smaller takes, were you don't really get the benefit of the most expensive models, and my guess is that is the case for the most users
Again this is a resolution problem. Your tasks are small enough that fit into a nice $3 quota. If you are an enterprise or a power user, the right-sizing argument doesn't work.
I'm talking about API prices - subscription is a different game.
> And gemma downloads also can be from auto CI pipelines etc. Nothing concrete
I have always found NPM download numbers truly suspect. Is no one caching? Are they estimating true number of downloads base on some estimate of cache hits?
Who knew the real revenue unlock wouldn’t be based on how much paranoid red-teaming the model underwent to resist users jailbreaking its ‘alignment’, and instead more on whether the model is post-trained to use ‘sed’ and ‘git’? Poor Gemini
Scientist-heavy orgs that want to solve everything in token space may overtook tool use; meanwhile Anthropic has been super focused on MCP, Claude Code etc for over a year
Out of the three main US AI companies' models, Gemini is obviously the less aligned (read: censored). So I really don't know what you're talking about.
If you just talk to it over API (no web search) the Gemini models are extremely resistant to thinking the user may be living in a universe outside their training data. Try to discuss any news etc and they assume it’s fake or fiction
> Lastly, after an incredible 27-year run, Jeff Dean is at a moment where he wants to try something new, and we’re excited to support him in that. Jeff and Google Senior Fellow Sanjay Ghemawat are launching an independent public benefit corporation to accelerate discoveries in ML, science, and engineering.
Oof. Good for Jeff and Sanjay (who just joined Twitter), bu that is a big loss for Google. Google stock is down 5%. It might not be much of an exaggeration to say these two are worth ~$200 billion.
For at least a year now the standard reflexive reply to “Google seems way behind OAI and Anthropic” has been “it’ll be ok, they’ve got Dean and Hassabis.” And now they don’t. What reason is there to be bullish about Google now?
As much as everyone wants to pay accomplished celebrities, all of these companies have young nameless geniuses that are about to make one for themselves. A guard passing torch can be opportunity.
Now, whether Google is the right environment to nurture, that’s its own quandary.
Couple of things come to mind. One is that Web sites tend to actively fight AI companies' crawlers while actively courting Google's crawler.
Another is that they hold the key Transformer architecture patent. If it is still relevant (which I'm not personally clueful about) and if they start enforcing it, then we may see a reprise of the situation where Microsoft made money for years every time an Android phone was sold. Regardless of that particular patent, it's probably safe to say they'll be better-positioned than anyone else if AI companies start lobbing patent nukes at each other.
A third factor that shouldn't be discounted is that Google has access to warehouses of training data that other companies don't. Google Books alone is an Alexandria-scale archive that the courts forced them to keep to themselves. Those restrictive copyright decisions may turn out to be a blessing in disguise for Google because no one else will have been able to scrape the data.
Google's weakness (well one of) is its total lack of cohesion. If the Google Books team could, they would sell access to that data in a heartbeat to boost their metrics.
Whoa. So is Jeff effectively leaving Google to work on this new venture full time? Or is the venture a side project? It sounds like the former. I’m sure he’ll still have internal access as an advisor of sorts. But this feels like a seismic change. Much larger than I initially realized?
It just tells you about their intended audience: Investors rather than consumers.
Nowadays, every company—even huge ones—prefers to be seen as a "growth opportunity", so they are trying to play up their ability to create new products which will somehow be so incredible that the line keeps moving up forever.
That said, there is definitely a correlation between companies chasing new products and leaving old ones to become crap.
Hasn't that always been the case with Google? Outside of a few that are mostly good and have stuck around, I've always seen Google has having great tech but being quite bad and making products out of it.
I'd bet my mortgage that if the first sentence was "We've got amazing products, amazing talent, and world-class compute", you'd be in the comments complaining that they didn't put talent first.
If all the other AI companies are promising AGI by Q4 of next year, what else can you do to satisfy shareholders than also jump on that same bandwagon?
The labs are all very interested in bio right now. Demis is working on Isomorphic rn (Google's version of that) but I could imagine a lab tempting him away to work on their equivalent if it had stronger momentum
This is a promotion for Demis and this could be a path for Demis to be CEO of Alphabet in the future in the AI era.
Google already invested in Discovery Loop (Jeff Dean, Orol Vinyals, Quoc Le and Sanjay Ghemawat's company), so what is happening is still a win in investment terms.
The question is about Sundar's future at Google, he is a mobile era CEO at Alphabet and I would hazard a guess he will probably step down in less than 3 years.
I don't see Demis becoming CEO. He's a scientist and researcher. He wouldn't want to be bogged down by minutiae of corporate politics, org structure, government relations, mobile hardware, etc. Chair lets him have authority to explore any path of interest without overhead of operations.
My impression, inside and out of G, was that Sundar (and Ruth) were about scaling down R&D expenses (as a fraction of revenue), and focusing on exploiting the monopolies.
Perhaps now there could be a shift back to investing in R&D to get fresh monopolies.
1) I think about this a lot. The AI overview has already destroyed the need to even look further down the page for a lot of people. And if you look further down the page, there is often a whole page of AI generated blog spam, which is in turn being regurgitated by the AI overview up-top. The AI overview often regurgitates a completely false reddit comment from two days ago as well. That AI generated blog spam is probably reading the AI overview.
That's the fun part. We don't. Could have happened 10 years ago without us noticing. I don't expect my skin cells to be able to recognize "me" anymore than I expect we will be able to recognize a superintelligent AGI. If it arrived 10 years ago, then the last 10 years could simply be its PR campaign. We wouldn't even be able to tell if it was successfully achieving its "goals" or not, assuming an AGI even has "goals"
Can we please stop saying anything about "AGI"? I remember when people would be like "AGI in 3 months" / "AGI in 2024" / "AGI is confirmed in 2025" like stop, you don't know if it's even a thing, let alone if it's coming/imminent.
> Google DeepMind: We are building strong momentum: Flash is in high demand, our Cyber model is live, and Gemma models have surpassed 900M+ downloads
Considering these are the best stats they could find, gemini usage+general situation must be really, really bleak.
High demand means nothing. A model being live is nothing to brag about. And gemma downloads also can be from auto CI pipelines etc. Nothing concrete
> Flash is in high demand
I said for a few years to many a downvote on HN, everyone wants AI, nobody wants to pay the true costs, the AI race will turn into a "race to the bottom" that is, who can give you the most compute for the lowest cost, and still remain profitable?
That said though, Flash isn't it. The prices on the latest flash models put Sonnet and Terra to shame.
Does everyone want AI?
Single data point; no, I don't. I preferred the pre-AI world. It makes me sad that we'll never see it again.
Personally, In SWE, i think the industry has made a grave mistake with the agents and we're just one big Catastrophe waiting to happen. I do think that there is very real value when software engineers use these tools as something akin to exoskeletons that allow the human to do more, rather than just fully replacing them. However, I'm finding more and more that companies are slop shops and just attempting to automate all of their software engineering. That will certainly end terribly. I hope we are not cannon fodder.
For what? There are some things I want AI for because it does it well. There are some things I don't want AI for because it just makes a mess (hallucinations). Maybe the next AI will be different and we will have the conversation again.
Indeed, it kills our planet, our culture and our economy extremely well. Oh yes, and some code monkeys enjoy that it can make computer code on the side.
I guess it should be said, of everyone who wants AI, they don't want to pay the expense for it to the level they want to use it.
No.
But people want what AI does for them. It's a tragedy of the commons situation.
> I said for a few years to many a downvote on HN, everyone wants AI, nobody wants to pay the true costs, the AI race will turn into a "race to the bottom" that is, who can give you the most compute for the lowest cost, and still remain profitable?
I keep seeing this but this line of thinking doesn't make any sense. What does it really mean?
There are expensive models that increase the probability of you doing your task under a lower cost. That means you can't use Gemma for coding your new compiler - it would just be overall costlier.
Heavier models are cheaper at more complicated tasks because they use fewer turns and fewer mistakes.
Cheaper models are more likely to be cheap at less complicated tasks. Like if you just ask Gemma "Hi" it would probably be cheaper than asking Opus.
So what does this statement really mean? People don't want to pay the extra for a more costly model? Why wouldn't you? It reduces your overall cost!
> Why wouldn't you? It reduces your overall cost!
Because real Fable usage starts at $20/month, and has oppressive usage limits even at that (ridiculous) monthly price.
Compared to my $3/month GLM-5.2 subscription, I have never felt like I was leaving capabilities on the table by refusing to cough up $20 for 15 minutes of Fable use per day.
where are you subbing to GLM-5.2? i've been meaning to try it out and for $3 it's a no-brainer to just load it up and give it a shot.
This is the wrong way to look at it. If you have a complicated task , you can solve it for cheaper if you used Fable. It will use fewer turns to achieve the same result.
You can solve it for cheaper if you use GLM but if you are involved in it more, but that defeats the purpose.
The point is that there aren't many complex tasks were fable delivers a significant value increase over cheaper models.
Single prompting a very complex tasks is rare even on frontier models, because it can be done successfully only for specific situations (e.g. you have a very strong verification step the model can iterate on).
Most of my everyday usage is for smaller takes, were you don't really get the benefit of the most expensive models, and my guess is that is the case for the most users
> The point is that there aren't many complex tasks were fable delivers a significant value increase over cheaper models.
Strong disagree on this. Any decently complicated task like a refactor is going to be more likely to be solved by Fable than by Gemma 3B or whatever.
I have personally tried to use Sonnet over Opus for tasks and Sonnet gets things right sometimes and at other times I wish I had just paid higher.
This is the standard pattern I keep seeing and I can have a bet with you that it would stay like this.
It's not cheaper if the price of admission is $20 for the first taste. And it's definitely not cheaper to pay per-token versus using my GLM-5.2 quota.
Again this is a resolution problem. Your tasks are small enough that fit into a nice $3 quota. If you are an enterprise or a power user, the right-sizing argument doesn't work.
I'm talking about API prices - subscription is a different game.
> And gemma downloads also can be from auto CI pipelines etc. Nothing concrete
I have always found NPM download numbers truly suspect. Is no one caching? Are they estimating true number of downloads base on some estimate of cache hits?
Who knew the real revenue unlock wouldn’t be based on how much paranoid red-teaming the model underwent to resist users jailbreaking its ‘alignment’, and instead more on whether the model is post-trained to use ‘sed’ and ‘git’? Poor Gemini
Scientist-heavy orgs that want to solve everything in token space may overtook tool use; meanwhile Anthropic has been super focused on MCP, Claude Code etc for over a year
Out of the three main US AI companies' models, Gemini is obviously the less aligned (read: censored). So I really don't know what you're talking about.
Gemini isn't as heavily aligned as OAI / Ant models ..?
If you just talk to it over API (no web search) the Gemini models are extremely resistant to thinking the user may be living in a universe outside their training data. Try to discuss any news etc and they assume it’s fake or fiction
> Lastly, after an incredible 27-year run, Jeff Dean is at a moment where he wants to try something new, and we’re excited to support him in that. Jeff and Google Senior Fellow Sanjay Ghemawat are launching an independent public benefit corporation to accelerate discoveries in ML, science, and engineering.
Oof. Good for Jeff and Sanjay (who just joined Twitter), bu that is a big loss for Google. Google stock is down 5%. It might not be much of an exaggeration to say these two are worth ~$200 billion.
For at least a year now the standard reflexive reply to “Google seems way behind OAI and Anthropic” has been “it’ll be ok, they’ve got Dean and Hassabis.” And now they don’t. What reason is there to be bullish about Google now?
[delayed]
Data centers, TPUs, customers, data streams, more money than god, most mature crawler system
As much as everyone wants to pay accomplished celebrities, all of these companies have young nameless geniuses that are about to make one for themselves. A guard passing torch can be opportunity.
Now, whether Google is the right environment to nurture, that’s its own quandary.
Couple of things come to mind. One is that Web sites tend to actively fight AI companies' crawlers while actively courting Google's crawler.
Another is that they hold the key Transformer architecture patent. If it is still relevant (which I'm not personally clueful about) and if they start enforcing it, then we may see a reprise of the situation where Microsoft made money for years every time an Android phone was sold. Regardless of that particular patent, it's probably safe to say they'll be better-positioned than anyone else if AI companies start lobbing patent nukes at each other.
A third factor that shouldn't be discounted is that Google has access to warehouses of training data that other companies don't. Google Books alone is an Alexandria-scale archive that the courts forced them to keep to themselves. Those restrictive copyright decisions may turn out to be a blessing in disguise for Google because no one else will have been able to scrape the data.
Google's weakness (well one of) is its total lack of cohesion. If the Google Books team could, they would sell access to that data in a heartbeat to boost their metrics.
Whoa. So is Jeff effectively leaving Google to work on this new venture full time? Or is the venture a side project? It sounds like the former. I’m sure he’ll still have internal access as an advisor of sorts. But this feels like a seismic change. Much larger than I initially realized?
it's a new venture that he can then sell back to google, and continue his ongoing loop
If you lost your only 2 Senior Fellows, you deserve to be replaced yesterday
The first sentence is quite telling
>We’ve got amazing talent, world-class compute and products…
Products are third on the list. Google is an incubator for talent first and foremost. Products are an afterthought
My opinion is that you're reading way too much into this. Talent is obviously first. Without that you have no products worth speaking of.
It just tells you about their intended audience: Investors rather than consumers.
Nowadays, every company—even huge ones—prefers to be seen as a "growth opportunity", so they are trying to play up their ability to create new products which will somehow be so incredible that the line keeps moving up forever.
That said, there is definitely a correlation between companies chasing new products and leaving old ones to become crap.
Hasn't that always been the case with Google? Outside of a few that are mostly good and have stuck around, I've always seen Google has having great tech but being quite bad and making products out of it.
I'd bet my mortgage that if the first sentence was "We've got amazing products, amazing talent, and world-class compute", you'd be in the comments complaining that they didn't put talent first.
I was about to invest in Google, but your compelling observation has swayed me. Now it's obvious to me they are a dying company.
A company is made up of people, and makes products. Products can come and go for many reasons (or any reason) but the company will always have people.
So then it is similarly telling that compute is second on the list? They have more talent than compute? Layoffs incoming?
This is also how Y Combinator operates.
Founders are first. Ideas are second.
If these products are the place that talent has brought us, of what use was the talent?
If all the other AI companies are promising AGI by Q4 of next year, what else can you do to satisfy shareholders than also jump on that same bandwagon?
Google loves to kick senior execs upstairs.
Whatever happened to Prabhakar Raghavan? Got kicked upstairs and we barely hear from him nowadays.
It's only a matter of time before Demis leaves and joins Anthropic or OpenAI.
He is probably itching to get to Anthropic. After all he is one of the early investors too.
As what? CEO? Demis will not take up the MTS role.
The labs are all very interested in bio right now. Demis is working on Isomorphic rn (Google's version of that) but I could imagine a lab tempting him away to work on their equivalent if it had stronger momentum
He was a protected entity (tambram)
can anyone with corporate background decipher if that is good for demis or not?
It is good for Demis and Google overall.
This is a promotion for Demis and this could be a path for Demis to be CEO of Alphabet in the future in the AI era.
Google already invested in Discovery Loop (Jeff Dean, Orol Vinyals, Quoc Le and Sanjay Ghemawat's company), so what is happening is still a win in investment terms.
The question is about Sundar's future at Google, he is a mobile era CEO at Alphabet and I would hazard a guess he will probably step down in less than 3 years.
I don't see Demis becoming CEO. He's a scientist and researcher. He wouldn't want to be bogged down by minutiae of corporate politics, org structure, government relations, mobile hardware, etc. Chair lets him have authority to explore any path of interest without overhead of operations.
I agree sundar will step down but I’m not sure if this means they’re grooming Demis to takeover. I suspect not, and someone else might take over.
Similar to how Sundar was promoted, isn't it more likely to be a promotion from existing leadership (one of the product leads/SVP).
My impression, inside and out of G, was that Sundar (and Ruth) were about scaling down R&D expenses (as a fraction of revenue), and focusing on exploiting the monopolies.
Perhaps now there could be a shift back to investing in R&D to get fresh monopolies.
> this could be a path for Demis to be CEO in the future in the AI era
His current title is CEO of Google DeepMind. Becoming Chief Scientist of Alphabet seems to be a step away from the path to replacing Sundar, no?
P.S. I think his real interest is Isomorphic, and his new role will offer fewer distractions.
You're crazy of you think Demis is CEO of alphabet material. He's an accomplished AI researcher, Google does about 10000 other things than AI.
Demis is CEO of Alphabet material. He would be exceptional in that role, and Google would be wise to put him there
> mobile era CEO
He was the guy branding Google an AI-first company back when they invented the transformer.
ty!
miss me with the AGI nonsense
Real next chapters:
1) Repair search that had been broken by AI initiatives.
2) Include less invasive AI with ads for those that need to be spoon fed.
3) Pretend to work on AGI and data centers in space.
4) Sell shovels and TPUs to the gold diggers.
1) I think about this a lot. The AI overview has already destroyed the need to even look further down the page for a lot of people. And if you look further down the page, there is often a whole page of AI generated blog spam, which is in turn being regurgitated by the AI overview up-top. The AI overview often regurgitates a completely false reddit comment from two days ago as well. That AI generated blog spam is probably reading the AI overview.
Where do real results live?
Lots of talk of us being at the cusp of AGI, but how will we even know when we get there? When AGI develops a religion for itself?
We'll know that we've reached AGI when a key conservative political belief in the US is that AIs are not people and do not deserve rights.
That's the fun part. We don't. Could have happened 10 years ago without us noticing. I don't expect my skin cells to be able to recognize "me" anymore than I expect we will be able to recognize a superintelligent AGI. If it arrived 10 years ago, then the last 10 years could simply be its PR campaign. We wouldn't even be able to tell if it was successfully achieving its "goals" or not, assuming an AGI even has "goals"
Science fiction.
Can we please stop saying anything about "AGI"? I remember when people would be like "AGI in 3 months" / "AGI in 2024" / "AGI is confirmed in 2025" like stop, you don't know if it's even a thing, let alone if it's coming/imminent.
The momentum is not being conserved here.
"Onwards!"
These guys...
Let's see... they ruined their core search product. What's next after that bold move?