#12 What Powers AI? A No-Jargon Guide to GPUs, Data Centers & the Compute Behind the Magic

What Powers AI A No-Jargon Guide to GPUs, Data Centers & the Compute Behind the Magic

When you type a question into ChatGPT and an answer appears in seconds, it feels like magic. The words seem to come from nowhere — instant, effortless, almost weightless.

But behind that simple text box sits one of the most massive physical machines humanity has ever built. Somewhere — perhaps in a warehouse the size of several football fields, humming with tens of thousands of specialised chips and cooled by rivers of chilled water — your question is being processed by hardware that costs billions of dollars and consumes as much electricity as a small city.

In our earlier articles, we explained what AI is, how Large Language Models work, and how AI is reshaping jobs. But we kept using words like “chips,” “compute,” and “data centers” without ever opening the hood. Today, we finally do exactly that.

This is your no-jargon guide to the physical engine of artificial intelligence — the GPUs, the data centers, and the enormous amount of power that turns electricity into intelligence.

Section 1: What “Compute” Actually Means

You’ll hear the word compute everywhere in AI. It sounds technical, but the idea is simple.

“Compute” is just shorthand for raw calculating power — the total amount of mathematical work a machine can perform in a given time.

Everything an AI does, at its core, is maths. When an LLM predicts the next word in a sentence (as we explained in our article on Large Language Models), it isn’t “thinking” like a human. It’s performing billions upon billions of tiny multiplication and addition operations, all at once, to calculate which word is most likely to come next.

Think of it like a giant restaurant kitchen. A single home cook (an ordinary computer) can prepare one dish at a time, beautifully. But to feed a stadium of 50,000 people simultaneously, you don’t want one brilliant chef — you want thousands of line cooks all chopping, stirring, and plating in parallel. AI is that stadium. It needs a kitchen built for massive, simultaneous, repetitive work.

That’s where special chips come in.

Section 2: CPUs vs GPUs — Why AI Needs Different Chips

Every laptop and phone you’ve ever used runs on a CPU (Central Processing Unit). The CPU is the “brain” of a normal computer — clever, flexible, and great at doing complicated tasks one after another, very fast.

But AI doesn’t need one genius doing tasks in sequence. It needs an army doing simple tasks in parallel. And for that, the industry turned to an unlikely hero: the GPU (Graphics Processing Unit).

A GPU was originally invented to render video game graphics — drawing millions of pixels on screen at the same time. It turns out that “doing millions of simple calculations simultaneously” is exactly what AI needs too.

Here’s the difference in plain terms:

  • CPU: A few very powerful cores. Brilliant at complex, sequential jobs. Like a handful of expert chefs.
  • GPU: Thousands of smaller cores. Brilliant at doing the same simple job thousands of times at once. Like an army of line cooks.

This is why one company — NVIDIA — became one of the most valuable businesses on the planet. It made the best GPUs for AI at exactly the moment the world needed them. Its chips (with names like Hopper, Blackwell, and the upcoming Rubin) have become the “gold” of the AI era, and each new generation is dramatically faster than the last.

Other players are racing to catch up. AMD makes competing GPUs, and the big tech giants are designing their own custom AI chips — Google’s TPUs, Amazon’s Trainium, and Microsoft’s MAIA — to reduce their dependence on NVIDIA.

Section 3: Training vs Inference — The Two Jobs of AI Hardware

Here’s a distinction that unlocks a huge amount of understanding: AI hardware does two completely different jobs, and they have very different costs.

1. Training — Teaching the Model

Training is the process of building the AI in the first place. The model is shown enormous amounts of text (books, websites, articles) and gradually learns the patterns of language, as we described in our LLM article.

Training a frontier AI model can take weeks or months, using tens of thousands of GPUs running non-stop, and can cost tens or even hundreds of millions of dollars for a single model.

This is the most compute-hungry, expensive, and energy-intensive phase. It’s like educating a student for years before they ever take a job.

2. Inference — Using the Model

Inference is what happens every single time you use the AI. When you ask ChatGPT a question, that’s one inference. It’s far cheaper than training on a per-use basis.

But here’s the catch: inference happens billions of times a day, globally. One trained model serves hundreds of millions of users. So while each individual query is cheap, the total inference load across the world is becoming the biggest driver of AI’s energy demand.

The simple way to remember it:

  • Training = build the brain once (huge upfront cost).
  • Inference = use the brain constantly (small cost × astronomical volume).

Section 4: Inside a Data Center — Where AI Actually Lives

So where do all these GPUs physically sit? Not in a lab. Not in the cloud (the “cloud” isn’t floating in the sky — that’s just marketing). They live in data centers.

A modern AI data center is a purpose-built industrial facility, often as large as several warehouses, containing:

  • Racks of servers, each packed with dozens of GPUs, stacked floor to ceiling in long aisles.
  • Networking equipment — miles of high-speed cabling that lets thousands of GPUs “talk” to each other instantly, so they can work as one giant brain.
  • Massive cooling systems to stop all that hardware from melting.
  • Power infrastructure — transformers, backup generators, and direct connections to the electricity grid.

When people say a data center has “500 megawatts of capacity,” they’re describing how much electricity it can draw — not how much data it stores. In the AI era, power has become the true measure of scale.

These facilities are the modern equivalent of the power plants and factories of the industrial age. And just like factories once clustered near coal and rivers, AI data centers now cluster near cheap, abundant electricity.

This is also why, as we noted in our Investment Spotlight coverage, countries like India are becoming major hubs. Global giants and Indian conglomerates like Reliance and Adani are pouring tens of billions into building data centers here — drawn by lower costs and the ability to power them with renewable energy.

Section 5: The Power Problem — Why AI Is So Hungry

Here’s the fact that increasingly makes headlines: AI’s biggest constraint is no longer chips. It’s electricity.

Training and running large models consumes staggering amounts of power. To put it in perspective:

Global data center electricity demand driven by AI is projected to reach levels comparable to the entire annual electricity consumption of a mid-sized country. A significant share of new AI data center projects now face delays not because of chip shortages — but because the local power grid simply cannot supply enough electricity fast enough.

Why so much? Because:

  • Tens of thousands of GPUs running 24/7 draw enormous continuous power.
  • Every watt of power that goes into computing turns into heat — which then requires even more power to cool.
  • The scale keeps growing, with each new model larger and hungrier than the last.

This has real-world consequences. Tech companies are now signing deals to buy power directly from nuclear plants, building their own solar and wind farms, and even reviving retired power stations — all to feed the AI machine. Energy, not intelligence, has become the bottleneck.

Section 6: Water and the Environmental Footprint

Power isn’t the only resource AI consumes. Water is the quiet second cost.

Many data centers use water-based cooling systems — essentially, water is circulated to absorb the immense heat generated by the chips, much like the radiator in a car. A large facility can consume millions of litres of water per year.

This has sparked genuine debate, especially when data centers are built in water-stressed regions. It’s one of the reasons the environmental impact of AI is becoming a serious topic — and why newer designs are experimenting with more efficient “liquid cooling” and closed-loop systems that recycle water.

For the everyday user, the takeaway is simple but important:

Every AI query has a tiny but real physical cost — a sip of electricity and a drop of water. Multiplied by billions of queries a day, those sips and drops become oceans.

This isn’t a reason to fear or avoid AI. It’s a reason to understand that “digital” doesn’t mean “weightless.” The magic on your screen is anchored to very real physics.

Section 7: The Full Stack — Who Builds What

To tie it all together, here’s the “stack” of players that make AI possible, from the ground up:

  • Chip designers (NVIDIA, AMD, Google, Amazon): Design the GPUs and custom AI chips — the engines.
  • Chip manufacturers (like TSMC in Taiwan): Actually fabricate these incredibly complex chips. Very few companies on Earth can do this.
  • Cloud providers / hyperscalers (Microsoft Azure, Amazon AWS, Google Cloud): Build and operate the giant data centers, then rent out the compute.
  • AI model builders (OpenAI, Google, Anthropic, Meta): Use all that compute to train the models you actually interact with.
  • You, the user: Send a query and receive intelligence — the final step in a chain worth trillions of dollars.

Understanding this stack helps demystify the news. When you read that “OpenAI is spending billions on data centers” or “NVIDIA is now worth $5 trillion,” you now know exactly where in this chain that money and value sits.

Section 8: Why This Matters to You

You might wonder: why should an everyday user care about GPUs and cooling systems? A few practical reasons:

  • It explains the costs. Ever wondered why powerful AI tools have subscription fees or usage limits? Now you know — every response has a real hardware cost behind it.
  • It sharpens your judgment. Understanding that AI is physical, expensive, and constrained helps you see through hype and evaluate claims more realistically.
  • It connects to investing. As we’ve explored in our Investment section, the AI infrastructure buildout is one of the biggest financial stories of the decade. Knowing the hardware helps you understand the “picks and shovels” of the AI gold rush.
  • It grounds the future. Debates about AI’s energy use, water consumption, and environmental impact will only grow louder. Being literate in the basics lets you engage with those debates thoughtfully.

Conclusion: The Machine Behind the Magic

The next time an AI answers your question in a heartbeat, remember what just happened. Your words traveled to a vast, humming facility, were processed by an army of specialised chips working in perfect parallel, cooled by chilled water, and powered by electricity drawn from the grid — all in less than a second, and then sent back to your screen.

AI can feel like magic. But like all great magic, there’s an elaborate machine behind the curtain. Understanding that machine doesn’t make the trick less impressive — it makes it more so.

Think of it this way: for a century, human progress was measured by how well we could turn energy into motion — engines, factories, highways. The AI era is defined by something new: turning energy into intelligence. The data centers rising across the world, including here in India, are the power plants of that new age.

And now, you understand exactly how they work.