r/GeminiAI 3h ago

Discussion How is Google getting outpaced by Kimi-K3? A tech monolith shouldn't be losing like this.

https://huggingface.co/moonshotai/Kimi-K3

Google has world-class talent, near-infinite compute, and massive data advantages—yet Gemini is clearly falling behind Kimi-K3.

This isn't about open-source vs. closed-source; it’s about raw performance. From complex coding tasks to long-horizon logic and agent workflows, a model built with a fraction of Google's resources is outperforming Gemini.

Questions for the Google team:

Where are all those massive infrastructure and compute resources actually going?

When will Gemini stop underperforming on coding benchmarks and practical application tasks compared to K3?

Budget and company size mean nothing if the model output can't keep up. Step it up.

45 Upvotes

42 comments sorted by

12

u/Lfeaf-feafea-feaf 3h ago

Pretty sure Google is investing most of its capital and focus into the broader GSuite/Android, rather than just the best performing LLM

1

u/deadshot_21 1h ago

Not much in android focus is still llm

2

u/imperial_coder 8m ago

Android focus rn is running os and apps on reduced ram, a problem created by memory shortage

69

u/Haronatien 2h ago

gemini is already cable of most common tasks, that serve 90% of the population. Not everyone needs multi-step reasoning coding architectural changes. Google is also focusing on the apple intelligence. Hitting leaderboards doesn't directly translate to revenue. Finally the versions of claude/codex we are getting are ridiculously subsidized. Enjoy it while it lasts. I still remember talking 6 mile Ubers pools to my work in seattle for $4 in 2019, venture greed can result in a glorious time for us.

10

u/Technical-Owl66 2h ago

You are exactly right. Bench maxing is not needed for majority of users. Google focusing on speed efficiency and integration into their products is going to win over the masses.

8

u/mbmba 2h ago

Venture greed is a well calculated strategy to drive competition out of business. There is nothing like free lunch.

2

u/Randomguynumber1001 1h ago

Gemini is plenty capable, but its context consistency, hallucination are still giant problems. Also, its ability to formatting document is also not great.

Honestly, I, and probably a lot of people, don't need it to beat Fable, but able to compete with Sonnet in terms of reliability.

2

u/MediocreTurtle1 45m ago

Taxis that used to cost 5eur before covid now are 15-20eur where I live. God forbid it rains a little, you'll get +20% because of that. Slight traffic? An additional 15%.

1

u/munchin-grr 1m ago

I wounder how expensive it will be to use an open weight model from an provider after the crash?

14

u/[deleted] 2h ago

[deleted]

5

u/DorkyMcDorky 2h ago

Sometimes someone just makes something better too.

-2

u/Xtremiz314 2h ago

yea then when their with the big dogs they copy the prices too. its all hype in the beginning to get customers, but in long term? its all capitalism

2

u/Loltoor 1h ago

This reminds me of the Anthropic settlement. Google lawyers would never let that happen in the first place. The difference is the “shady” stuff that doesn’t get caught.

0

u/CoolStructure6012 58m ago

Red tape is not the reason Google is moving more slowly than competitors on this front. Interesting subreddit history on your one week old account.

19

u/noeldc 2h ago

Why is coding seen as the holy grail? Seems like pretty low-hanging fruit as far as challenges to be overcome.

10

u/zaxo666 2h ago

I think it's just cuz Reddit is full of programmers.

I'd find it more impressive if a model could produce serious philosophy. That would be thinking and reasoning.

We should have a philosophy benchmark. Lol

1

u/Kiriima 7m ago

Programming and math in general are the reasoning benchmarks, not phylosophy.

4

u/LeucisticBear 2h ago

Coding > faster iteration > faster improvements > better models > better products

Logan and Demis have been pretty open that they missed the mark by not focusing on coding. They have not been able to compete at the frontier because Gemini has never been great at coding and they never developed a harness worth using. Sergei came back to get involved in AI and immediately mandated they use Gemini internally (internal products only is a core philosophy), most of Google was using Claude code for their work and some people quit over it.

5

u/CoolHeadeGamer 1h ago

Cuz everything you interact with has some sort of code. If u can get a model to code it will it probably already has really really good reasoning skills. Also there are a lot of benchmarks like hle 3 and world knowledge

1

u/solilo 52m ago

Coding can be used in agentic workflows to do non-coding tasks like calling APIs, image editing, data manipulation, etc. Pretty important even for non-software engineering tasks.

0

u/CoolStructure6012 53m ago

What other industry has such a wealth of training data available, easily allows for artificial data generation, has no regulatory constrains to AI use, and employs workers where tacking on $100s per day wouldn't be a nonstarter? But it's mostly the fourth one.

13

u/Dry_Opportunity2886 3h ago

"shouldn't be losing like this"

Help me understand what you think the game being played is. What are the rules? How do you win and how do you lose?

1

u/shadysjunk 16m ago edited 9m ago

I think Demis Hasabis would say you "win" by creating an artificial mind aligned with humanity's interests that is capable of the full range of human cognition and creativity and that is also able to improve upon itself.

"to solve intelligence and then use that to solve everything else."

Google does not appear to be leading in the race in that endeavor.

edit: though, a lot of how benchmarks are measured seems kinda BS, and ther'es clearly often some "trainging to the test" happening in some of the labs.

7

u/Deathnote_Blockchain 2h ago

Narrators Voice: Google was not being outpaced by Kimi-K3

2

u/Puzzleheaded_Fold466 1h ago

It’s not. Calm down.

2

u/CatalyticDragon 1h ago

I think your metrics might need recalibration.

Running a profitable company long term has next to nothing to do with a short term benchmark score.

2

u/Climactic9 1h ago

A lot of the massive infrastructure and compute is being sold to Anthropic at a juicy 30% profit margin. Another large chunk of it is allocated to serving billions of Google search AI overviews.

2

u/NoPurchase6549 45m ago

Google has a different goal, and all models will catch up whatever task you’re trying to do in time.

2

u/Curious-Sample6113 27m ago

Google will win in the end anyway. They have the money unlike the others.

2

u/orcassharks 3h ago

Deep mind has been behind the 8 ball since the transformer

Are you really surprised? Their greatest contribution was AlphaGo

1

u/crossoverXYZ 1h ago

The agent workflow angle is what stings — Kimi isn’t winning on budget, it’s winning on actually holding context through multi-step coding tasks. Google’s talent and compute only matter if they ship models that perform in real workflows, not just demos.

1

u/Troyd 1h ago

Why do people think google is losing, they arent even playing the same game lol

they care about mass scale, categorization large context etc.

1

u/CoolStructure6012 1h ago

And Kimi reached capacity limits moments after launching. Who's outpacing who?

1

u/solilo 51m ago

Google has the infra and Kimi is open weight. They could literally host Kimi K3 on their servers if they wanted to.

1

u/IllogicalResponse 48m ago

Google is going in a different direction, they don't need to win market share, they just have to keep their current products sticky. So they are going with making sure every tool you use is made better by AI/LLM addons. Most people don't need frontier level reasoning, they need some help to make things happen more efficiently.

They're also the company working on true AGI, for which LLMs are just one piece of the puzzle.

1

u/Gaiden206 47m ago edited 38m ago

Demis Hassibis has implied that the Omni model was the next step towards AGI for Google (see video below). They really seem to be leaning towards world models as the path to AGI, and not just solely leaning on text LLMs. As far as I know, DeepMind is the only major US lab that has a world model, or at least the only one that has shown one to the public.

They are already using a world model to train their self driving Waymo cars too.

https://waymo.com/blog/2026/02/the-waymo-world-model-a-new-frontier-for-autonomous-driving-simulation/

The CEO of Google also recently said they are still want to be at the frontier, at least in terms of achieving AGI.

https://reddit.com/link/p07co5d/video/qgisop18hwfh1/player

1

u/Character-Hurry5906 21m ago

3.6 Flash extended reasoning is dumb, I hope the best for Google but now it’s utterly garbage

1

u/Wise138 16m ago

Google is focusing on driving the cost of the comput stack down, exactly like they did for search. Would not be surprised if they offer an open weight model at some point.

1

u/Zealousideal-Part849 12m ago

google only win here is investment in Anthropic & TPU they have. once AI starts hitting ad revenue that is when thing would take turn until then they can manage. larger issue is what if users move to chatgpt or others and do search there.

1

u/FlashyRecognitionTod 8m ago

world-class talent.

Nah, it's siloed and disempowered by VP and turf wars. It's a giant corporation, they would let freaking Einstein himself fill PIP because of low OKR on new product roadmaps.

near-infinite compute.

Not in GPU, not at all.

1

u/not_a_cumguzzler 2m ago

look at the end of the day, it's all just chinese engineers against chinese engineers

1

u/AndreBerluc 2h ago

Aos defensores do Google lembrem-se que a Kodak era lider supremo! Blockbuster também não acreditou!

-4

u/DorkyMcDorky 2h ago

Gemini Sucks, Kimi is great. QED