As far as I understand all the inference purpose-build silicon out there is not ...

nightshift1 · 2026-01-02T22:40:08 1767393608

According to this semianalysis article, the Google/Broadcom TPU are being sold to others like Anthropic.

https://newsletter.semianalysis.com/p/tpuv7-google-takes-a-s...

nomel · 2026-01-02T22:21:07 1767392467

> It seems that custom inference silicon is a huge part of the AI game.

Is there any public info about % inference on custom vs GPU, for these companies?

mrinterweb · 2026-01-02T22:57:01 1767394621

Gemini is likely the most widely used gen AI model in the world considering search, Android integration, and countless other integrations into the Google ecosystem. Gemini runs on their custom TPU chips. So I would say a large portion of inference is already using ASIC. https://cloud.google.com/tpu

almostgotcaught · 2026-01-02T22:14:50 1767392090

> soon

When people say things like this I always wonder if they really think they're smarter than all of the people at Nvidia lolol

mrinterweb · 2026-01-02T23:00:48 1767394848

Soon was wrong. I should have said it is already happening. Google Gemini already uses their own TPU chips. Nvidia just dropped $20B to buy the IP for Groq's LPU (custom silicon for inference). $20B says Nvidia sees the writing on the wall for GPU-based inference. https://www.tomshardware.com/tech-industry/semiconductors/nv...

almostgotcaught · 2026-01-03T00:11:45 1767399105

There are so many people on here that are outsiders commenting way out of their depth:

> Google Gemini already uses their own TPU chips

Google has been using TPUs in prod for like a decade.