Skip to content
  • Facebook
  • X
  • Linkedin
  • WhatsApp
  • YouTube
  • Associate Journalism
  • About Us
  • Privacy Policy
  • 033-46046046
  • editor@artifex.news
Artifex.News

Artifex.News

Stay Connected. Stay Informed.

  • Breaking News
  • World
  • Nation
  • Sports
  • Business
  • Science
  • Entertainment
  • Lifestyle
  • Toggle search form
  • Access Denied Business
  • Karnataka College Students Fight Over Playing Movie Song: Report
    Karnataka College Students Fight Over Playing Movie Song: Report Nation
  • Centre debunks social media claims over Rajnath Singh’s Operation Sindoor speech
    Centre debunks social media claims over Rajnath Singh’s Operation Sindoor speech Nation
  • After Rolls Royce And Porsche, Rs 2.5 Crore Diamond Watch Seized From Tobacco Baron
    After Rolls Royce And Porsche, Rs 2.5 Crore Diamond Watch Seized From Tobacco Baron Nation
  • Access Denied Business
  • PM Modi To German Business Delegation
    PM Modi To German Business Delegation World
  • No Messi in Argentina’s Olympic football squad; Álvarez and Otamendi selected for Paris Games
    No Messi in Argentina’s Olympic football squad; Álvarez and Otamendi selected for Paris Games Sports
  • Govt allows onion exports to Bangladesh, Mauritius, Bahrain, Bhutan
    Govt allows onion exports to Bangladesh, Mauritius, Bahrain, Bhutan Business
What is a GPU? How does it work? | Explained

What is a GPU? How does it work? | Explained

Posted on February 19, 2026 By admin


The story so far: In 1999, California-based Nvidia Corp. marketed a chip called GeForce 256 as “the world’s first GPU”. Its purpose was to make videogames run better and look better. In the 2.5 decades since, GPUs have moved from the discretionary world of games and visual effects to becoming part of the core infrastructure of the digital economy.

What is a GPU?

Very simply speaking, a graphics processing unit (GPU) is an extremely powerful number-cruncher.

Less simply: a GPU is a kind of computer processor built to perform many simple calculations at the same time. The more familiar central processing unit (CPU) is on the other hand built to perform a smaller number of complicated tasks quickly and to switch between tasks well.

To draw a scene on a computer screen, for instance, the computer must decide the colour of millions of pixels several times every second. A 1920 x 1080 screen has 2.07 million pixels per frame. At a frame rate of 60 per second, you will be updating more than 120 million pixels per second. Each pixel’s colour will also depend on lighting, textures, shadows, and the ‘material’ of the object.

This is an example of a task where the same steps are repeated over and over for many pixels — and GPUs are designed to do this better than CPUs.

Imagine you’re a teacher and you need to check the answer papers for an entire school. You can finish it over a few days. But if you have the help of 99 other teachers, each teacher can take a small stack and you can all wrap up in an hour. A GPU is like having hundreds or even thousands of such workers, called cores. While each core won’t be as powerful as a CPU core, the GPU has many of them and can thus complete large repetitive workloads faster.

How does a GPU do what it does?

When a videogame wants to show a scene, it sends the GPU a list of objects described using triangles (most 3D models are broken down into triangles). The GPU then runs a sequence called a rendering pipeline, consisting of four steps.

(i) Vertex processing: The GPU first processes the vertices of each triangle to figure out where they should appear on the screen. This uses maths with matrices (sort of like organised tables of numbers) to rotate objects, move them, and apply the camera’s perspective.

(ii) Rasterisation: After the GPU knows where each triangle lands on the screen, it fills in the triangle by deciding which pixels it covers. This step essentially converts the geometry of triangles into pixel candidates on the screen.

(iii) Fragment or pixel shading: For each pixel-like fragment, the GPU determines the final colour. It could look up a texture (e.g. an image wrapped on the object), calculate the amount of lighting based on the direction of a lamp or the sun, apply shadows, and add effects like reflections.

(iv) Writing to frame buffer: The finished pixel colours are written into an area of memory called the frame buffer. The display system reads the buffer and renders it on the screen.

Small computer programs called shaders perform the calculations required for these steps. The GPU runs the same shader code on many vertices or many pixels in parallel.

Effectively the GPU reads and writes very large amounts of data — including 3D models, textures, and the final image — quickly, which is why many GPUs have their own dedicated memory called VRAM, short for video RAM. VRAM is designed to have high bandwidth, meaning it can move a lot of data in and out per second. Still, to avoid having to fetch the same data, the GPU also contains smaller, faster memory in the form of caches and arrangements for shared memory, with the goal of keeping memory access from becoming a bottleneck.

A mock diagram of a GPU as found in graphics cards.
| Photo Credit:
Public domain

Many tasks outside graphics also involve performing the same type of calculation on large arrays of numbers, including machine learning, image processing, and in simulations (e.g. computer models that simulate rainfall).

Where is the GPU located?

A chip is a flat piece of silicon, called the die, with a fixed surface area measured in square mm.

In a computer, the GPU is not a separate layer that sits below the CPU; instead it is just another chip, or set of chips, mounted on the same motherboard or on a graphics card and wired to the CPU with a high-speed connection.

If your computer has a separate graphics card, the die holding the GPU will be under a flat metal heat sink in the middle of the card, surrounded by several VRAM chips. And the whole card will plug into the motherboard. Alternatively, if your laptop or smartphone has ‘integrated graphics’, it likely means the GPU and the CPU are on the same die.

This is common in modern systems-on-a-chip, which are basically packages containing different chip types that historically used to come in separate packages.

Are GPUs smaller than CPUs?

GPUs are not smaller than CPUs in the sense of using some fundamentally smaller kind of electronics. In fact, both use the same kind of silicon transistors made with similar fabrication nodes, e.g. the 3-5 nm class. GPUs differ in how they use the transistors, i.e. they have a different microarchitecture, including how many computing units there are, how they’re connected, how they run instructions, how they access memory, etc. (E.g. the ‘H’ in Nvidia H100 stands for the Hopper microarchitecture.)

CPU designers devote a lot of the die’s area to complex control logics, the cache ( auxiliary memory), and features that improve the chip’s performance and ability to make decisions faster. A GPU on the other hand will ‘spend’ more area on many repeating compute blocks and very wide data paths, plus the hardware required to support those blocks, such as memory controllers, register files, display controllers, sensors, on-chip networks, etc.

As a result, GPUs — especially the high-end ones — often have more total transistors than many CPUs, and they aren’t necessarily more densely packed per square mm. In fact, high-end GPUs are often very large. Some GPU packages also place dynamic RAM very close to the GPU die, connected using short wires with high bandwidth. Essentially, the architecture of components needs to ensure the GPU can transfer large volumes of data quickly.

Why do neural networks use GPUs?

Neural networks — mathematical models with multiple layers that learn patterns from data and make predictions — can run on CPUs or GPUs, but engineers prefer GPUs because the networks run many tasks in parallel and move a lot of data.

The maths of neural networks is in the form of matrix and tensor operations. Matrix operations are calculations on two-dimensional grids of numbers, like rows and columns; the numbers in each grid can represent various properties of a single object. The essential problem is to multiply two grids to get a new grid. Tensor operations are the same idea but use higher-dimensional grids, like 3D or 4D arrays. This is useful when the neural network is processing images, for instance, which have more properties of interest than, say, a sentence.

In matrix multiplication, the value of c12 (red-yellow circle) is equal to a11b12 + a12b22. Likewise, the value of c33 (blue-green circle) is equal to a31b13 + a32b23.

In matrix multiplication, the value of c12 (red-yellow circle) is equal to a11b12 + a12b22. Likewise, the value of c33 (blue-green circle) is equal to a31b13 + a32b23.
| Photo Credit:
Lakeworks (CC BY-SA)

A neural network repeatedly adds and multiplies matrices and tensors. Since it’s the same set of mathematical rules, just applied on different numbers, the thousands of cores of a GPU are perfect for the job.

Second, contemporary neural networks can have millions to billions of parameters. (A parameter is a learned weight or bias value inside the network.) So in addition to doing the maths, the network also has to be able to move data fast enough — and GPUs have very high memory bandwidth.

Many GPUs also include tensor cores, which are designed to multiply matrices extremely fast. For example, the NVIDIA H100 Tensor Core GPU can perform around 1.9 quadrillion operations per second of tensor operations called FP16/BF16.

In fact, Google developed chips called Tensor Processing Units (TPUs) to efficiently run the maths that neural networks require.

The green board everything is mounted on is the printed circuit board. The four flat, silver metal blocks arranged in a vertical column near the middle are liquid-cooled packages. The green hoses and the coloured tubes are coolant lines to and from the packages. Each package contains a TPU v4 chip surrounded by four high-bandwidth memory stacks. Four connectors dot the board’s left edge.

The green board everything is mounted on is the printed circuit board. The four flat, silver metal blocks arranged in a vertical column near the middle are liquid-cooled packages. The green hoses and the coloured tubes are coolant lines to and from the packages. Each package contains a TPU v4 chip surrounded by four high-bandwidth memory stacks. Four connectors dot the board’s left edge.
| Photo Credit:
arxiv:2304.01433

How much energy do GPUs consume?

Let’s use a hypothetical example where four GPUs are used to train a neural network to predict the risk of some disease for a person (based on age, BMI, blood markers, some history). Then the same network is put in use.

Each GPU is an Nvidia A100 PCIe, whose board power is around 250 W during training. The GPUs are nearly fully used during training. The training duration is 12 hours.

The energy consumed during training will be 12 kWh and during use, around 2 kWh (assuming only one GPU provides the inferences). The server will also consume power for its CPUs, RAM, storage, fans, and networking, and some power will be lost. It’s typical to add 30-60% of the GPU power for these needs. So the total consumption will be around 6 kWh/day for the network to run continuously.

That’s like running an AC for four to six hours at full compressor power, a water heater for around three hours or 60 small LED bulbs for 10 hours a day.

Does Nvidia have a monopoly on GPUs?

Nvidia technically doesn’t have a monopoly on GPUs; it enjoys a near-complete dominance in some markets and is a very strong market power in artificial intelligence (AI) computing platforms.

In discrete GPUs sold for use in personal computers, industry trackers have reported that Nvidia has roughly 90% market share at least, with AMD and Intel making up most of the rest). As for GPUs used in data centres, Nvidia’s position is strengthened by hardware performance and supply and the CUDA software ecosystem.

CUDA is Nvidia’s software platform to run general-purpose computation (like processing a signal or analysing data) on Nvidia GPUs. As a result, switching away from using Nvidia GPUs also means changing software, which companies don’t like to do. In fact, many buyers consider Nvidia GPUs running CUDA software to be the default platform for training and using neural networks at scale.

The legal definition of monopoly depends on whether a firm can control prices or exclude the competition and whether it maintains that power through unlawful conduct. This is why, for instance, European regulators have been investigating whether Nvidia uses its dominance to lock customers in, mainly by tying or discounting GPU prices when buyers also take Nvidia software or related components.

mukunth.v@thehindu.co.in



Source link

Science

Post navigation

Previous Post: Access Denied
Next Post: Access Denied

Related Posts

  • The mystery of Déjà vu
    The mystery of Déjà vu Science
  • Eye donation and corneal transplant
    Eye donation and corneal transplant Science
  • Understanding the Impact on Elephants and Tigers in India
    Understanding the Impact on Elephants and Tigers in India Science
  • How is global warming affecting sea breeze?
    How is global warming affecting sea breeze? Science
  • Inside India’s ‘Deep Ocean Mission’, a challenge harder than going to space
    Inside India’s ‘Deep Ocean Mission’, a challenge harder than going to space Science
  • Zoopharmacognosy: the study of how animals self-medicate
    Zoopharmacognosy: the study of how animals self-medicate Science

More Related Articles

Science quiz: Tools of writing Science quiz: Tools of writing Science
The Science Quiz: The evolutionary edge to human survival The Science Quiz: The evolutionary edge to human survival Science
How does air pollution affect the brain? How does air pollution affect the brain? Science
Anthropocene epoch declaration unlikely soon, but the idea lives on | Explained Anthropocene epoch declaration unlikely soon, but the idea lives on | Explained Science
‘Ghost particles’ signature can track where spent nuclear fuel goes ‘Ghost particles’ signature can track where spent nuclear fuel goes Science
Can datacentres in orbit solve for AI models’ soaring energy demand? Can datacentres in orbit solve for AI models’ soaring energy demand? Science
SiteLock

Archives

  • September 2026
  • August 2026
  • July 2026
  • June 2026
  • May 2026
  • April 2026
  • March 2026
  • February 2026
  • January 2026
  • December 2025
  • November 2025
  • October 2025
  • September 2025
  • August 2025
  • July 2025
  • June 2025
  • May 2025
  • April 2025
  • March 2025
  • February 2025
  • January 2025
  • December 2024
  • November 2024
  • October 2024
  • September 2024
  • August 2024
  • July 2024
  • June 2024
  • May 2024
  • April 2024
  • March 2024
  • February 2024
  • January 2024
  • December 2023
  • November 2023
  • October 2023
  • September 2023
  • August 2023
  • July 2023
  • June 2023
  • May 2023
  • April 2023
  • March 2023
  • February 2023
  • January 2023
  • December 2022
  • November 2022
  • October 2022
  • September 2022
  • August 2022
  • July 2022
  • June 2022
  • May 2022

Categories

  • Business
  • Nation
  • Science
  • Sports
  • World

Recent Posts

  • Hyena gave birth to two cubs at the Thiruvananthapuram Zoo
  • FPIs turn cautious; withdraw ₹20,974 crore from equities in September amid global uncertainty
  • World leaders return to UN amid wars in West Asia and Ukraine
  • Free mega medical camp promotes Health and Legal awareness in Gajuwaka
  • After October 1, only smaller 5 or 10 kg LPG refills await non-eKYC households

Recent Comments

  1. Glennalbub on UP Teacher Who Asked Students To Slap Muslim Classmate
  2. PhilipLog on UP Teacher Who Asked Students To Slap Muslim Classmate
  3. DennisDum on UP Teacher Who Asked Students To Slap Muslim Classmate
  4. DonaldFrosy on UP Teacher Who Asked Students To Slap Muslim Classmate
  5. DonaldFrosy on UP Teacher Who Asked Students To Slap Muslim Classmate
  • Delhi Painter Gets Hands Back As Organ Donation Meets Surgical Excellence
    Delhi Painter Gets Hands Back As Organ Donation Meets Surgical Excellence Nation
  • A Small Animal Hospital In Mumbai
    A Small Animal Hospital In Mumbai Nation
  • UTT: Goa, Kolkata secure semifinal berths
    UTT: Goa, Kolkata secure semifinal berths Sports
  • Insurance regulator IRDAI retains LIC, GIC Re, New India Assurance as D-Slls  
    Insurance regulator IRDAI retains LIC, GIC Re, New India Assurance as D-Slls   Business
  • Access Denied Business
  • Access Denied
    Access Denied Nation
  • GST Council meeting LIVE: Industry bodies hail two-rate slab and other reforms
    GST Council meeting LIVE: Industry bodies hail two-rate slab and other reforms Business
  • Enforcement Directorate Opposes Arvind Kejriwal’s Request
    Enforcement Directorate Opposes Arvind Kejriwal’s Request Nation

Editor-in-Chief:
Mohammad Ariff,
MSW, MAJMC, BSW, DTL, CTS, CNM, CCR, CAL, RSL, ASOC.
editor@artifex.news

Associate Editors:
1. Zenellis R. Tuba,
zenelis@artifex.news
2. Haris Daniyel
daniyel@artifex.news

Photograher:
Rohan Das
rohan@artifex.news

Artifex.News offers Online Paid Internships to college students from India and Abroad. Interns will get a PRESS CARD and other online offers.
Send your CV (Subjectline: Paid Internship) to internship@artifex.news

Links:
Associate Journalism
About Us
Privacy Policy

News Links:
Breaking News
World
Nation
Sports
Business
Entertainment
Lifestyle

Registered Office:
72/A, Elliot Road, Kolkata - 700016
Tel: 033-22277777, 033-22172217
Email: office@artifex.news

Editorial Office / News Desk:
No. 13, Mezzanine Floor, Esplanade Metro Rail Station,
12 J. L. Nehru Road, Kolkata - 700069.
(Entry from Gate No. 5)
Tel: 033-46011099, 033-46046046
Email: editor@artifex.news

Copyright © 2023 Artifex.News Newsportal designed by Artifex Infotech.