What's the difference between TPU vs. GPU?
So which is it? The answer is confusingly both. In the Pixel 11 series smartphones, the TPU is basically Google's fancy-schmancy way to say NPU. This version of a TPU handles camera and image processing as well as local AI tasks. But the data center TPU is decidedly different from a smartphone version, a traditional NPU and a GPU, although there are a few similarities. In order to help keep your chipset lingo straight, we're here to formally introduce you to the real TPU and clear up some of the misconceptions.
Great question. A TPU is a chip Google created to deal with cloud-based AI and machine learning tasks in the company's data centers. Known as an AI accelerator, a TPU is optimized to perform millions of mathematical calculations on the fly in AI models, something that NPUs and GPUs do as well.
We've been talking a lot about AI and ML, which for many of us will call into play an NPU. As it should, since we've been beaten over the head with them ad nauseum thanks to Apple, the various PC and smartphone makers and the chip manufacturers. If you're unfamiliar, NPUs are the chips you'll find in all your modern smartphones, Macs and PCs. They do all the heavy lifting for your generative and agentic AI tasks, whether it's erasing an ex from a photo or using ChatGPT to build out a three-day travel itinerary. For the Pixel 11 phones, the TPU replaces the NPU, with Google claiming "up to 3.5 times faster AI processing while using up to 3.5 times less energy."
GPUs are familiar to many of us, especially if you're a gamer or work with 3D models. Before AI, we wanted a powerful GPU to run graphically demanding games like Crysis 3 at the highest settings with blistering frame rates and buttery smooth rendering. Post-AI, GPUs still do that as well as handle AI training and crypto mining , hence why the best ones cost an arm, a leg and a firstborn. A jack-of-all-trades, a GPU can do many things well depending on how powerful it is.
The things that set these chipsets apart are largely use cases and scale. While TPUs are used for cloud-based AI running in Google data centers, NPUs, on-device TPUs and GPUs can be found in your smartphones and laptops.
TPUs have the largest scale by far. Thanks to its specialized architecture, a TPU is employed in accelerating massive computations used in LLM (Large Language Models) and massive deep learning (DL) tasks. TPUs utilize a hardware layout called a systolic array that allows data to go from one unit to the next in a 2D grid of multipliers, which help eliminate potential bottlenecks. That means the calculation output becomes the input for the next unit without having to write back to memory, unlike GPUs, which are stuck juggling data between their compute units and their high-bandwidth memory.
When servicing big AI companies such as Anthropic and Midjourney, which serve billions of AI requests daily, relying on a GPU-based solution can be costly in terms of bottlenecks, latency and power efficiency.
5News aggregated this summary from the outlet’s public feed. The full article, with all the context, is on www.engadget.com — the content belongs to Engadget.