# NVIDIA Spark Chip and the Shift to AI-Native Computing

**Podcast:** a16z Podcast
**Published:** 2026-06-02

## Transcript

Having lived through like a half dozen component shortage things, you just sort of wait them out and you don't let some local Macs or local men determine the future.
This will all correct itself in short order.
This world where you're all gated on dollars per token is a thing that's going to move to your own device, which is exactly what happened with all of computing.
Anytime there's a resource constraint that you have to pay for.
it moves to your device and becomes free.
AI introduces yet another opportunity to change that dynamic for the PC, to have it be forward-looking, not backward-looking.
And I think this is an incredibly important opportunity for Microsoft and for the industry as a whole.
I'm in the Situation Room with Steven Sinofsky, who might have been like the first ever guest on MTS back when we were still doing test streams, I think.
He was the first person I interviewed on a test stream.
He was the president of the Windows division at Microsoft.
He created the Surface program at Microsoft, which we have some very interesting news about today.
We're thrilled to have you on.
Stephen, welcome to MDS.
Welcome back.
Well, thanks so much.
Good to see you.
Hi, everyone.
Yeah, hi.
First question would be, NVIDIA...
And Microsoft and Arm and a few other companies just announced something very interesting at Computex.
What exactly did they announce?
And what does it matter?
Sure.
Well, just so folks know, because it doesn't get into as much, but Computex is this big giant trade show in Taiwan.
And it's the weirdest show because it's like this total inside baseball, you know, Silicon supply chain show.
And normally you never hear about it.
Like, in fact, I never went to it even.
I wanted to, but it actually turns out it was always right around the same time as a big Microsoft sales meeting.
So I never went, but you can think of it as the ecosystem show for everything it takes to build a computing device of any kind.
Totally well, but Jensen in his keynote last night.
did this incredible slide where he walked up and down the whole length of the stage pointing to partners that he was very excited to be there.
And I would bet anyone that anyone watching would have no idea who the companies he was pointing at.
Like these are some of them, like half of the ones he pointed to as kind of being entertaining were just names of companies in traditional Chinese.
And so you didn't even, like you don't even know what they are.
But it's an incredible show.
It's just wild because it's such inside baseball about components and peripherals and chipsets and assembly lines and it's deals and deals and stuff.
And every 10 years or so, it jumps into the mainstream, but never like the past 24 hours.
Just, you never see that.
And actually, it was a lot like, I think it was two years ago, Jensen keynoted CES.
And I've been to...
40 CESs and I'd never seen one with such a broad media reach.
Yeah, it's like Taylor Swift of the tech industry.
It's an incredible level of, it just speaks to the awareness of tech and then the awareness of AI and what NVIDIA has done.
Because I mean, like I, you know, it was bigger than like CES Xbox was like a sideshow compared to this.
Wow.
And he's huge in China too.
Well, the show is always huge in Asia broadly because most of the companies are there, whether they're in mainland or in Taiwan or in Vietnam or Singapore.
And that's sort of the origin of the show.
Like this show is like go there and speak an Asian language forever.
So fascinating.
So the big announce though, I mean, look, there's a zillion things going on, but the one that rose to the top.
was NVIDIA announcing what they called the RTX Spark Super Chip, which is a mouthful.
Before the show, it was broadly called N1X.
And that's sort of what it is.
It's an ARM CPU mounted with NVIDIA parallel processing graphics, basically into one system on a chip.
that has a whole new memory architecture relative to the historic way that PCs have been built.
And the target for it are the PC makers.
So it's very, very exciting.
And if you, the mainstream press, the stock market press, the CNBC going on behind me, like they all looked at this as like NVIDIA entering the PC business, which is, let's call it the mainstream chip business.
Which is so weird because, you know, long before in the stone ages, which is now we're talking about 2011, we actually announced NVIDIA-milking PCs and making the Surface computer, the very first one.
And it was a Computex big announcement, and it was a CES announcement, and it was, you know, I remember that very vividly the partner slide had NVIDIA and Qualcomm and Texas Instruments.
and all the chip makers and PC makers.
So it has a ring of familiarity at, I would imagine, like 1,100 the scale.
Oh, wait, I sent you guys, you could throw it up there.
I sent you the tech meme from that day when we announced the stuff, and it was still a pretty significant role.
So you'll be able to throw it up at some later date.
Yeah.
So in what way is this laptop more like AI native?
So the big thing that's really changed is that the compute burden has shifted from the CPU to the graphics processor and then the associated neural processors, TPUs and those chips.
And that's the thing that really changed.
It's not unlike, you know, 15 years ago what had changed was the bulk of the interesting processing had moved from the CPU to the GPU just for rendering.
And where we are today is just an extension of all of that.
And the difference, it's just insanely important because now that's the compute that we think everybody wants to do.
Now, a lot of people will look at this and go, well, I use ChatGPT.
I do it on my MacBook Air, my Chromebook, my phone.
I don't really need another thing.
The problem is, and this might be a word you guys have used or heard before, but the problem is tokens.
And so the problem is that everybody is gated by the consumption of tokens, which cost money, where you can't get them if you're trying to use them for free.
And so the interesting thing about this device is how much of compute can it move to your local device where you basically have infinitely free tokens?
And that is incredibly interesting and super important.
Now, of course, you've all seen this run up.
to where we are today with the stacks of Mac minis.
Yeah.
And everybody was running their agents on minis.
So why did they do that?
Well, there's a whole bunch of stuff about privacy and sandboxing them.
But the primary thing was if you just want to let something roll for three days while it figures out your best travel itinerary, you really don't want to end up with a $10,000 bill.
So instead, you buy three minis and let it crank away with like...
each minute putting something in isolation or whatever.
And so, but if you just fast forward, you know, six or nine months, it's abundantly clear.
And I say that as a predictor of the future, not like as a, it's obviously intellectually clear, but like, it just seems to me that this world where you're all gated on dollars per token is a thing that's going to move to your own device.
Yeah.
Which is exactly what happened with all of computing.
Anytime there's a resource constraint.
that you have to pay for, it moves to your device and becomes free.
And that, it just, I just don't imagine, I don't know how it can happen.
Sure.
So for someone who wants a more AI-native device in like a year when all of these products have shipped, do you think they're going to want like an NVIDIA Spark laptop or do you think they're going to want to stick with like MacBook Pro or the rumored MacBook Ultra, which is supposed to come out later this year, early next year?
Well, this is just, I mean, this is the huge thing.
And the way that I think this can play out is, well, of course you can play it out in like essentially status quo, which is, you know, the Fortune 500, you know, 80-20, 70-30 rule will be it will just fall to Windows devices running Intel or maybe Spark devices running ARM, but running with a Windows operating system.
And then, you know, the cool people, the...
the bosses, the elites, or whatever you call it, running their MacBook Pros with Chrome or Safari just connecting things, and phones.
But there's another path where it becomes incredibly important to run highly optimized AI stack of software on your device.
And whatever that stack is, is going to get optimized for a particular hardware base.
And that's a thing we've seen.
over and over again.
Now, where we are right now is just so interesting because we don't have enough information to know where things are heading.
At the announcement last night and the press releases and the commentary, you know, Microsoft made it clear, much to my surprise, which we could go into, that the NVIDIA stack of CUDA will be available and supported and part of this Spark.
Now, there are a lot of ways for that to become true.
It could be a download that just runs.
It could be a thing that's installed, pre-installed on a Spark device.
It could be a thing that's part of the OS and updated with Windows Update and administrative permissions and all of this other stuff.
It could be a whole range of things.
And I still don't, nobody knows yet in terms of public announcing how they're going to do that.
The same thing holds for Apple.
And today on a Mac, you can run all the models locally and stuff like that.
You can't really do that on a phone.
And so an interesting question is going to be, what is Apple going to do at WWDC with respect to the CUDA APIs?
Are they going to be native?
Are they going to be a thunking layer?
Lots of stuff could happen there.
Are they distributed?
Is it an App Store app?
Is it an OS component?
Nobody has any idea.
Now, for both companies, the past is very interesting, and most people didn't live through this, but NVIDIA has always been an outsider to the personal computer industry.
It's always been an add-on.
So on the PC, if you ever wanted to use an NVIDIA graphics card, you bought the card and you downloaded drivers from NVIDIA, or before that, they came on a CD or a floppy disk that came with your graphics card.
And so for...
30 years, this whole thing was like, do you have the latest NVIDIA drivers?
Where do you get them?
And we went from getting a new CD to getting a new DVD to FTP to downloading them from the web.
It was a whole cycle.
But it was never a first-class part of Windows until we fixed that in Windows 7 and got them on Windows Update and all this other stuff.
And the APIs on a PC to do graphics, you could always just download the NVIDIA library and call them.
But the official Windows APIs were DirectX.
And they just did the same kind of thing, just completely differently.
That the X is Xbox.
And so Microsoft was all in on the DirectX APIs.
They were a huge part of Windows release called Windows Vista when they first got integrated and then Windows 7 forward.
Then there were the NVIDIA APIs, which at first were just the NVIDIA APIs, then they became CUDA.
Then for graphics, NVIDIA embraced this open thing called OpenGL.
And then Apple went through the same exact thing.
On the Mac, you could download drivers.
You could install an NVIDIA card.
But the APIs, and then they supported OpenGL for a while.
But they always wanted you to use their own stuff.
And the phone did away with all of that.
And Marable did away with all of that.
And it was all in on Apple.
Now, the good news for Apple was that native graphics were just outstanding, and they've always been great.
On the PC side, Intel was so far behind that it just kept pulling both NVIDIA and ATI slash AMD to be, you know, what you used if you used Photoshop or made movies or were just graphic intensive.
And so in the next few weeks, we'll know what Apple is going to do for these APIs, and more importantly, the models themselves, and the runtime.
I mean, NVIDIA has an enormous investment in the open source models and tuning them through their hardware.
And the ecosystem has done a great job, as evidenced by the Mac minis, of tuning those APIs for the Mac.
But that has nothing to do with the phones.
And so that's an operating system difference.
And the number of phone people is large.
And as we know, the hardware is the same.
Now, it's not quite the same and blah, blah, blah.
amount of memory, all that stuff.
But it's very interesting to see the details of Microsoft and then what Apple chooses to do.
Right.
So, like, obviously we are seeing like a memory shortage, right?
And like, so what do you expect the cost to a consumer of this kind of like, you know, very AI native computing device to cost?
Well, certainly for the...
Having lived through like a half dozen component shortage things, you just sort of wait them out and you don't let some local max or local min determine the future.
This will all correct itself in short order.
The history of it, whether it's been DRAM or hard drives or processor shortage, all of these things, we've had them come and go or even smaller components.
So I'm not worried about it at all.
I mean, obviously...
If you're thinking that you need 96 or 128 gig for a standard consumer device versus, say, the 8 on a MacBook Neo, there's a huge difference.
But also that will change in the models too.
Like right now, the models themselves are all tuned to run in hyperscale data centers.
And every month it seems like there's a new paper that says, oh, we cleaved this giant thing off of the inference pipeline.
So now we don't.
need nearly as much memory.
So that all will get fixed.
Not even an inkling of concern I have for that problem.
So another thing that was just announced yesterday was last time we talked, you were very, very excited about the MacBook Neo for Apple as like a category-defining product.
Dell just came out with the new XPS 13 that is, they say, slightly better specs on...
slightly better specs than the MacBook Neo, and it's like $100 more expensive.
So the Neo is like, what is it?
It's $400 for students, $500 for everyone else.
It's $599, $699.
$599, $699.
Yeah.
Yeah, this one.
$499, $599, I think.
Yeah, the SPS 18 is $599, $699.
Yeah, yeah.
MacBook Neo is $499, $599.
Yeah.
So what are your takes on this?
Well, first, You know, kudos to Dell.
Like, Dell is just on an incredible roll.
And Michael Dell is just a legendary CEO.
And read his book, his second book that came out during the pandemic, I think, or right after, right before.
It's fantastic.
But the XPS 13, for a very long time, was sort of my go-to laptop when friends and family and whatever would ask for one.
It is like...
the best laptop.
And then it took a little bit of a dip and went in a funky way on design.
And it is actually back with a vengeance now.
And so XPS 13 is the laptop to get.
Now, this latest one is an attempt to build on that same chassis using Intel.
I, you know, 30 years, I can't keep track of the names or three.
Panther Lake, Python Lake, something, Crystal Lake.
I don't know what it is.
It's some lake.
That's Intel names are always.
places I've never been.
Panther Lake.
Panther Lake.
It's the Intel Rails.
There's a lot of excitement about it.
It does integrate some of the AI compute stuff into it.
But it's not going to be the target machine that the PC ecosystem wants to sell.
And its capabilities are going to be different from the ones that they do want to sell, which is different than the Neo, which has the capabilities that...
need to be targeted.
So the Apple hardware line has a lot of homogeneity in it in terms of capability.
But PCs can become really hit or miss.
And that's always been the difficulty in the PC ecosystem is even when there's a winning machine, it's not the one that when you walk into Best Buy and say, I need to buy a computer, help me, Mr.
or Mrs.
Salesman.
And then they just direct you to something based on SPF and current ads or whatever.
So we'll see.
I'm sure it's a quality machine.
You have the tech people on X talking about taking sides.
It's either the Neo killer because it has an HDMI port or whatever, or it's embarrassing to the PC ecosystem because regardless, it still runs Windows.
Those extremes are stupid.
Both the Neo and this machine are targeted at just people who need a computer.
I think in five years, people who need a computer will also need a computer that runs agents.
But the hardware software world will be unimaginably different in five years.
So this conversation has no relevance to the product lines that will be available in five years.
So, wow, it goes up to 32 gigs of RAM and a terabyte of storage.
It starts at 8 gigs of RAM, which is like not...
Great, I guess.
512 gigs of SSD.
That's not a good number for a PC.
Yeah.
The Mac would do 8.
The PC is, I, you know, like I hate saying it because, and I've truly had a bunch of these over the past month, six months or so, but I spent a lot of energy with our team on getting the memory down to two and four gigs at the time.
But 8 is going to be hard.
Now, Windows is doing a lot of work right now on that.
So we'll see where that goes.
But right now, if you ask me about what PC to buy, I would send you to a 16 gig PC.
It takes work.
It takes like techie work to get it down to be 8 gig reasonable.
Like uninstalling a bunch of stuff, playing around, stuff that you shouldn't tell anyone to do.
Like which PC would you recommend?
Like a specific laptop?
Dell XPS?
Yeah, any.
I would get a...
I would get a Dell XPS 13.
What do you think about the Surface lineup right now?
Well, here, okay, look, I'm obviously not objective about the Surface lineup.
I think that the PC, when we designed Surface, and I wrote like 80 million words on this, which everybody could go see on Hardcore Software.
I'm not going to replay them here about what we did wrong and right.
But originally, Surface was envisioned to be this platform discontinuity in PCs.
It was going to be the move to mobile chips, ARM, and mobile firm factors, like a tablet.
We shipped it as a convertible tablet, but as a tablet.
We actually did an Intel x86-based Surface, and at the time, we called it an objection handler.
And it was to handle the objection of things you didn't like about...
the ARM-based surface, you know, like that.
It didn't run existing software.
It wasn't compatible with old software or whatever.
And so my heart and the strategy for ARM was always to introduce this discontinuity where, look, the world is different now.
The hardware world is different now.
The usage scenarios are different now.
And portability is different.
People want better battery life.
They don't want fans.
They don't want viruses and all that other stuff.
Didn't work.
I moved down to San Francisco area to attract founders and stuff.
And what Microsoft did was sort of basically abandon ARM for the next eight years or so and focus on the objection handler side of things.
So all the services that followed were, in my mind, a niche product because they were just like different Intel PCs.
weren't super important to me.
I mean, I had brought each one of them, but they weren't important to me.
AI introduces yet another opportunity to change that dynamic for the PC, to have it be forward-looking, not backward-looking.
And I think this is an incredibly important opportunity for Microsoft and for the industry as a whole.
But the word, it's different like it was in 2011.
in that it's mobile chips and the scenarios are different?
Even more so.
80% of the typical PC buyers are just running browser-based compute, and they just want the keyboard, the form factor, and they like Macs because they don't wear down over time.
They have all their battery life for real.
They have the viruses and malware, a whole different game.
It's sort of this sealed case that we used to call it.
PCs that did move to ARM also thwarted all of the Windows APIs, which was the thing we chose not to do.
So now the new PCs running ARM are just the old PCs with the same viruses, the same problems with fans, the same lack of quality over time.
The classic Windows thing is, oh, you could just go edit the registry.
Well, if you have an ARM PC, you can still go edit the registry and you can still fork your PC totally and then you're screwed.
And so I just don't think that backward looking is the thing.
Which brings us to last night and all the X comments on the Spark laptops and everybody immediately jumping to two things.
First...
NVIDIA announced that they're all going to run all existing Windows programs, which of course just follows from Microsoft's strategy of learning Win32 to ARM, which wasn't hard.
We'd already done it.
It was just opening up the dev tools and the ability to load the apps and things, which we disabled for ARM because we wanted to move the ecosystem forward to a new OS API.
But then the other part of this is just how you...
You spin the whole thing in terms of backward compatibility.
And then they said, oh, it runs every single app of all time.
It's like, yeah, but you don't want to do that.
And more importantly, the second thing is you don't need it anymore.
But all of the enthusiasts are going nuts because they see it as Intel being replaced by NVIDIA, which is conceptually true.
Except not really.
It's just an alternative.
And what you're going to see in the marketplace is just sort of this price comparison.
And Intel and NVIDIA are just going to drive the prices to each other.
And only one of them can really afford the battle.
But that doesn't change the value proposition for consumers, which is what they really want is to not have that backward compatibility.
They just don't know it.
If they got a PC without a fan that...
You couldn't edit the registry.
You couldn't break it.
You couldn't just go into the system folder and delete stuff.
All of these things that you don't even think about on a Mac anymore and you don't even think about, you can't even think about on a phone.
You don't want them on the PC.
And so it's tough for me to see Microsoft sort of embracing this because I mean, I understand like if you want to sell the enterprise, you have to run that VB app from 2003, but that's not, you don't need to do that.
You could just put it.
on a server and remote into it.
You could put it in a VM on an x86 machine.
There's a million ways to do that.
You just don't need to run it on the machine that you want to run your agents on.
And in the short term, everybody is going to be running terminal anyway.
And these agents in today's agents are all headless anyway.
That will change too.
But right now, that's the big, we make this fork in the road.
And Microsoft has already said the direction that they want to take it, which is they just want NVIDIA chips to do all the things that Windows has always done, which always tests with customers.
And so you say the customer's like, yeah, but those registry editors and admin scripts and stuff really screwed us up.
And like, we know, we know this.
So it's tough for me to see.
Yeah.
Well, we're really excited for the NVIDIA Spark laptops to come out later this year.
We'll see if it can replace my MacBook.
Yeah.
I'm not sure.
We'll see.
We'll do a tech review on stream.
I mean, you could see what it's going to be like if you just have the Spark that you could get from Dell today.
That's the mini device.
I have one of those, and it's incredible.
I mean, it's just...
And there's a tower as well.
All right, well, we'll take a look at that.
Steven, thank you so much for joining us.
Sure.
Thank you, guys.
Thanks again for listening, and I'll see you in the next episode.
This information is for educational purposes only and is not a recommendation to buy, hold, or sell any investment or financial product.
This podcast has been produced by a third party and may include paid promotional advertisements, other company references, and individuals unaffiliated with A16Z.
Such advertisements, companies, and individuals are not endorsed by AH Capital Management LLC, A16Z, or any of its affiliates.
Information is from sources deemed reliable on the date of publication, but A16Z does not guarantee its accuracy.
