Pradeep Gupta & Stephen Jones on turning CUDA libraries into agent skills
Two NVIDIA veterans explain how CUDA stays vertically integrated yet horizontally open, and why turning its 900+ libraries into agent skills reshapes the developer's job.
Vertically Integrated, Horizontally Open
NVIDIA tunes the whole stack top-to-bottom for performance, then opens every layer so partners can plug in their own hardware, software, data, and models.
We want to open each and every layer to our ecosystem so that everyone can go and be the part of this entire platform, which is what we call an accelerated computing platform.
What CUDA Actually Is
CUDA is the layer that carries any application down to the GPU, and its secret is that it is not one thing but many stacked layers, from frameworks at the top to the driver at the bottom.
CUDA is the magic that connects your application to the GPU hardware. Anything that you are doing that is GPU accelerated, that involves accelerated computing, it goes through CUDA.
The Car and Its Wheels
On top of CUDA sit more than nine hundred CUDA-X libraries, reusable pieces that depend on one another so you extend proven work instead of reinventing it.
I kind of think of it in terms of CUDA as like a whole car it's a whole machine and the libraries are the wheels right nobody can use the car without the wheels. But you wouldn't reinvent the wheel either.
Open Does Not Mean Open Source
NVIDIA's open means CUDA is freely available to take, use, and build a business on without permission, not that it is open source or vendor-neutral, though many libraries and even Nemotron are open source too.
Open means it's freely available. You can go and build your businesses on top of that.
Half His Time on the Other Side
A CUDA software architect spends fully half his time with the hardware team, so each generation from Turing to Vera Rubin is co-designed, and new CUDA advances are pushed back to older GPUs like Ampere and Turing.
I'm in the software organization as one of the architects of CUDA, but fully half my time is spent working with the hardware team.
Every Library Becomes an Agent Skill
NVIDIA is exposing every CUDA-X library as an agent skill, so an agent can pull in the right library for a job, which Pradeep expects to lift adoption by an order of magnitude.
think of CUDA skills to every CUDA library is actually going to drive the adoption of this to, I will say, an order of magnitude higher what it has been done today.
You Manage the Agents
Developers are shifting from writing code to writing prompts and guiding the agents that write it, with Stephen reporting he can now do about three times more.
I must admit that I haven't written code, mostly this year, because I now write prompts, because the agents write the code.