AI COMPUTE•NVIDIA•GPU CLOUD•DATA CENTERS•SEMICONDUCTORS•ENERGY•FUNDING•M&A•AI COMPUTE•NVIDIA•GPU CLOUD•DATA CENTERS•SEMICONDUCTORS•ENERGY•FUNDING•M&A•

Fundamentals

What Is a GPU?

A plain explanation of graphics processors and why they became the engine of modern artificial intelligence.

Updated

LinkedInX

A GPU, or graphics processing unit, is a processor built to run a very large number of similar calculations at the same time.

Where a CPU has a modest number of powerful cores optimised for sequential logic, a GPU has thousands of simpler cores optimised for parallel arithmetic. That design was created for rendering images, but it maps almost perfectly onto the mathematics of neural networks, which is dominated by large matrix multiplications.

Modern data-center GPUs add three things that matter more than raw core counts: very high bandwidth memory close to the chip, specialised matrix units, and fast interconnects that let many GPUs behave like one larger machine. Training a frontier model is largely a networking and memory problem, not only a compute problem.

This is why GPU supply, memory availability and interconnect capacity now shape the economics of the entire AI industry.

GPU Data Hub Daily

Stay Ahead of the AI Infrastructure Economy

The most important GPU, AI, data-center, semiconductor and cloud developments delivered directly to your inbox.

By subscribing you consent to receive the daily briefing. Unsubscribe at any time. See our privacy policy.

Cite this page

“What Is a GPU?.” GPU Data Hub. https://gpudatahub.com/guides/what-is-a-gpu

You are welcome to reference and link to this page. Please link to the URL above.