# AI Infrastructure Strategy: SLAs, Portability, and CXL

**Podcast:** The CTO Advisor
**Published:** 2026-02-04

## Transcript

All right, this is part two of a data center.
I don't want to use the term modernization, but the reality is AI is changing our data center.
And the last episode, we talked with LANCOM about the impact of the changes that Bracom has made purchasing VMware to our AI infrastructure.
And now we're going to focus today's conversation more on that AI portion of the challenges that we're facing.
Lynn, welcome back to the show.
Thanks, Keith.
I'm excited to have the next conversation here because I, you know, we know that this conversation is evolving really fast in the industry.
And so what was it a month ago that we spoke?
And lots of things have happened in between there.
So lots to cover today.
Yeah, and I want to I think I want to set this one up.
Uh Nature Fresh From Farms shared at a tech field day a few months ago, almost a year ago, about their journey with AI.
They do some pretty amazing things.
Basically, they're looking to improve yields one to three percent every growing season.
And of course, they're using machine learning AI to do this.
They have a cluster of Dell nodes running Intel processors, and they added NVIDIA GPUs to the mix, and they on paper went, you know, you saw the improvements you expect to go from uh using GPUs, I mean using CPUs to GPUs, and AI processing, you know, a 10x performance increase.
But what they discovered was that their SLA for getting the results back from the modeling data was 24 hours, and GPUs reduced it to you know uh the performance down to maybe 30 minutes for processing.
But what they discovered was their CPUs were already meeting the SLA by like 3x the the the requirements, like eight hours.
And at the end of the day, they increased the complexity of their environment trying to manage uh and I'll link it in the in the show notes, but they ended up with increased complexity so much so that they pulled the GPUs out of their cycle.
Are you seeing similar integration pain points with other Intel customers?
Oh yeah, I've seen at least three different situations, can't mention the names, but you know, one of them is is one of the largest, you know, data center owners that you would recognize as name brand value.
And you know, essentially they start out using CPUs because they can't get access to GPUs or they can't get the power.
Or, or you know, they have to wait a year and they can't wait on their deployment model.
And then what they find is once they get it deployed on on the CPUs, that it's it works for the SLAs they have, and I think that that's the main the the nuance here is it's easier to not think about SLAs and to run fast and just think I'll just fix it later.
But the problem is if you don't understand your SLAs, then you're not going to understand how to have uh ROI in your model or how to calculate it.
You'll just, you know, you'll take the safe route, which is specific hardware, but then what you didn't do is actually build in the uh view into what kind of work do we need to get done and what is the definition of efficiency for this kind of situation.
Yeah, I gotta give my good friend Bobby Allen credit for this saying uh many companies ago that he worked for he had the saying that you don't give unlimited capacity and capability to solutions that have a limited business value.
So this idea and we've experienced this with Cloud, right?
That yes, we can architect something to go exceptionally fast.
We can scale up to and this is the context in which he was talking about it, we can scale up to a thousand nodes to a process that doesn't have any business value beyond it running on one VM.
And I'm seeing this pattern being repeated in AI, but I think it's easy to make the argument.
And we talked about this a little bit before the start of the the recording.
I think the the counter-argument to that is that that if these processes are going to eventually end up on AI or end up on GPU, doesn't it make more sense to start it out on GPU, even if you know you're gonna be wasteful in resources and money?
I I don't think many CTOs like that idea, but you know, they don't have to replatform later.
Intel has to be hearing that story in market Lynn.
Talk to me about this this portability as a design goal.
Well, I think that the portability isn't necessarily at the the library level.
That's not where you're really gonna get it.
You're really gonna get it at the partner level, you know.
So some of our partners, Kami Waza, Iterate, Iternal, there's there's a number of them.
And they can run on multiple different configurations.
Um NetApp is another one that is finding that there's some really good use cases.
Um and and not everybody needs the the full-blown um the the highest end agentic.
You can use a lot of small language models.
In fact, NVIDIA just published in September that small language models are in many cases more effective, and we see a number of customers using vertical models.
They're specific to their industry.
And so the question is really is are CPUs and GPUs, you know, CPUs are more general purpose than GPUs.
GPUs are general purpose when they're running these large language models.
And the question is really, what gets you to the industry-specific um implementations.
And then, you know, one of the other things about integration, the operational blockers that IT teams are facing is I need to get breakthrough usage models that can operate my infrastructure without spending any extra money, and I'm not thinking anything about change management.
So I think that the industry is starting to realize we need to slow down to speed up a little bit.
Now, you can tell me I'm wrong on that, but I have find so many instances where there's this.
I start with a CPU because of different assumption set, and then I find I don't really have to move.
And if you do the portability correctly and the architecture correctly, so that you can have a solution that's accelerated or not, and you know, with with Intel Xeon 6, you could be using an accelerator on deck called AMX.
Um, that's really the important design uh criteria for having optionality as you move forward.
Yeah, the uh I don't think you're gonna get a lot of pushback from me today.
I'm out of pushed back a couple of months ago when I built my CTO advisor, uh virtual CTO advisor, and then my CTO advisor is stack builder that kind of implements a lot of our methodology.
It was a really cool solution.
I built it in about a month and I built it on Google Cloud.
And then I got a little worried.
I said, Well, I don't want to be locked into Google models, Google pricing, et cetera.
And I went to move it to an on-prem solution.
I bought an NVIDIA GPU based system.
And when I went to move it, I realized that I was locked in.
I I was indeed locked in.
I I didn't even build to like CUDA anything.
I built to the API and I couldn't move my Google ATX APIs on premises.
So not you're you're not going to get pushback on me on that.
But this does raise the question of governance.
And as architects as CTOs are thinking through governance, what are some of like the smarter things you've seen customers do on the governance side?
Because this is not just a problem when you're talking about do I want to run on X all X86?
Do I want to run on all GPUs?
Or even if I want to abstract it to the on the API layer, the governance question spans all three of those approaches.
Yeah.
You know what's interesting with governance is um since I ended up studying for and clearing the AI governance professional exam, there's a lot of things that have shifted.
You know, you see some of the large cloud service providers saying, uh, I'll give you a a a box.
You're gonna have my you can have a box with my stuff in it, and that will be governed and it will cover the sovereignty and the governance.
You've seen um a lot of of groups in other countries just saying I'm just gonna build data centers.
And then you can ask yourself, you know, is that sustainable?
You know, do they have the full production capacity for that?
Um and so basically there's there's the orchestration and then there's where things are running and then there's where the data is residing.
And I think that um we've gotten so cloud native there's gonna be some unpeeling that has to happen um with where their data's placed and what they're giving access to and how they're doing things and taking advantage of things like um confidential computing and you know TDX and and all of that um platform uh visibility and telemetry that you can get access to.
Um and so it's so multifaceted for multinationals.
I think it's almost simpler if if you're a a a smaller company operating in a more limited region, you're not having to deal with things like the European Union AI Act, and then Singapore has its own AI Act and Canada has its own AI Act, and so does you know Brazil have its own policies.
Um so I think there's this is going to be the next the next wave of AI.
We got the technology there.
There's a lot of great agentic solutions that can provision infrastructure that can check in Jira tickets that can flag when things are happening, maybe even suggest fix fixes.
But at the end of the day, who's signing off on it?
So I'd be curious in your perspective on this and you know one of the other governance questions that I've been asking is if using AI tools to write code means that you don't own the code, then what's the governance over your own innovation there?
So lots to unpack there.
Yeah I've if we we could probably do a whole podcast series on this topic along I've wrote of what I'm calling the reasoning layer but whether you're automating this or doing it manually you have to think through how you're managing governance.
How do I ensure that data that my general model has access to in Europe isn't being leveraged for stuff in the US that European law doesn't allow you to do and you have to really think through this entire challenge of data sovereignty, uh the resulting AI intelligence or AI data becomes as a result to it.
I've talked a lot about the unintended consequences of when you give someone access to something like uh chat GTP or Copilot and they have access to data that they have access to but the insight because you can now uh do data scientist level work with this $20 a month tool to get insights a layer below above your pay grade how do you solve those problems and I think this podcast this single podcast about you know just the general problem of workload management by itself isn't enough time to go over it.
Well it's interesting you mentioned that because like you know I deal with some of these tools internally and what I've found is since the deployment of it we've actually gotten stricter and stricter data silence.
And so then you ask yourself well what's the $20 a month subscription for if things have gotten even tighter in terms of what we have access to.
And so that's become a challenge as well is you you have an overreaction within some organizations where um it becomes very difficult to be your own data analyst.
Yeah I've talked to a lot of C A C AIOs chief AI officers who have simply just said no to a lot of AI tools.
So uh part of RFPs is the question for SaaS solutions.
Can you turn off AI until the business can catch up?
And this is the business.
This isn't an IT thing.
This is the business is saying they want it to slow down until they figure out the repercussions of AI.
Let's shift the conversation back to technology.
I can't help but recognize that AI is a memory-bound workload.
Right.
Intel has worked with CXL for several years now.
Where are you seeing the impact of CXL on these enterprise AI workloads?
Well, you know, first it started out with memory attached, and now, you know, CXL is raising the bar.
So you can have accelerators attached, but that really comes down to higher memory utilization through pooling.
Because, you know, if memory, if networking isn't your bottleneck memory, your storage is going to be your bottleneck or storage over a network.
And so, you know, essentially what CXL is is the envisioning of that was free up stranded memory, allocate it where the workloads need it most, do dynamic memory provisioning.
And so having to reboot, you can basically reallocate flexibly, and then connect or reassign AI accelerators without rigid coupling to the host.
And so, you know, all of those are workloads.
And you know, it's funny is I have the from the Xeon Desk series, and I got some interesting comments on the one about CXL of, you know, it would be really, really um interesting to sit down and talk through some of the challenges, but again, there's so many opportunities with being able to dynamically adjust memory capacity available to a system.
It's gonna be one of those areas that we keep posting and can't mention their name, but I there's a there's a very large, large company that has been able to do um uh a memory mode using CXL attach that has really unlocked some of the data capabilities that are available.
Um and so again, it's it's it's a technology that we envisioned it takes a little while to get through the enabling of the ecosystem, but the the unlock that it can give customers is huge in terms of efficiency, performance, latency, and then not having to have um restarts within their overall uh pipeline of getting things executed.
Yeah, I'm thinking through not just AI, but traditional analytics.
What happens when you know I have my NetApp storage array that's aware of the CXL capabilities, and I'm able to dynamically move or expand compute across several nodes to ingest data as it's streaming in from this NetApp array or uh the ability to uh expand memory and do some really cool stuff like this, stuff that I'm expecting Intel to figure out is this hit rate that you are uh working on when you were putting uh memory DMs, uh store basically storage memory dams inside of inside of servers.
And it's the same I/O problem.
How do I ensure that the compute that uh is needed to process the data and the memory that's needed is closest to the data.
Data has gravity.
And once CXL continues to mature, I can easily see parking back to our example earlier in the podcast of a company like Nature Fresh Farms again extending the need to not have GPUs because they're uh they're they're they're troubling, they're solving the bottleneck until the bottleneck actually becomes real-time latency and they just need to work through a batch process uh you know, within this framework.
When their data grows, if CXL is there, now they don't again, they don't have to buy new equipment and make their environment more complex than needed.
Yep.
Yeah, I mean, I think it's it's it's one of those areas, there's so many degrees of freedom.
The entire data center architecture seems to be completely reopening in terms of you know what do efficiencies look like?
How is it being managed?
And um, you know, back to the integration pain point, I think that there's this huge question around how do you um you know how do you have stability in the middle of all of these new architectures and deployments, and where does that stability point come from?
And I think that's that's one of the main integration pain points or integration challenges for everyone that has deployment.
Do you you mentioned you've got a bunch of C AIOs who basically are was it C A IOs or CTOs that said no, you don't get to put I think C AIOs, these are not even you know, they're not they they're not even IT people.
Right.
And they're saying no.
And you know, a lot of the AI capabilities that seem most natural are coming from vendors that are bolt-ons into the into the properties that are already in in the in the fleet.
And so I think that there's so many considerations.
I I do not envy what some of the CTOs are dealing with these days, but when you start peeling it back to first principles, starting with governance, starting with architectural portability, so you've got optionality, use what you have in the fleet, look at how you can leverage memory and storage technologies more effectively.
You end up with with at least a simplification um in the middle of all of the AI change that that's constant.
All right.
So I'm going to get last question, Lynn.
It's going to be spicy.
Intel is obviously going through a transformation right now.
And you lead an entire team of folks at Intel, and I'm quite sure I've talked to them.
You're asking them to do more with less resources than they've had in the past.
How are you and your team measuring success when it comes to AI projects?
So that not only is Intel as a technology company promoting the use of AI tools, but you're showing the way.
How do you measure success uh with AI projects?
It depends on what type of AI project it is.
You know, an AI project metric of success in a design team is going to be different than in a go-to-market or marketing team like mine.
In the go-to-market and marketing team, a lot of it is coming down to removing hours on non-productive things, like removing hours recrafting content, removing ours sitting in meetings when you can take a transcript, and then you can actually feed it into a pre-built agent, have that agent operate like a coworker, and look at the transcript and then do the write-up, not just of the notes, but you know, potentially technical documentation as a result of sitting sitting in that meeting.
And so right now, what I'm seeing is the ROI, at least for the the collateral production, the messaging, the positioning, and and planning, quite frankly, is do I have to write the document, or can I have something else get the data access that's going to generate the document, and then I can work with that agent so that it's um it is basically going to get 90% of the way there, and I'm having to only correct a little bit more.
I know for a fact that there are um software teams and hardware design teams that are measuring things slightly differently, which is you know, how can I get to a design solution that's 80 or 90% of the way there without actually having to spend the weeks of analysis that we do?
Can we feed all that data and all that history in and then basically have the agent learn from that and then tweak it at the end?
I guess in some sense, that might be um cost to design goes down because you're spending less time.
And in my case, in marketing, it's also time goes down, which is effectively cost.
And so maybe you could argue that those are the same measures.
Um, I'd be curious if you see those as the same as I describe them to you.
Yeah, I I I felt like you're sitting in my own internal meetings.
I stretch a porn across all of that.
I use AI as a collaborator.
I don't, you know, I'm I'm solo again, so I don't have the access to 10 other analyst brains to bounce ideas off of.
So I'll go use a LLM, give it an idea, and say, you know, you are you know this analyst with a opposing idea critique what I've presented to you.
At the same time, I'm developing code.
Man, I haven't developed code in probably 15 years.
Yeah.
And I can now develop code.
I I'm trying to figure out and this is why I asked you the question.
I'm trying to figure out how to codify this.
Like how do you like for a solopreneur, obviously this is 10X my operations.
And when I talk to successful marketers, uh we're working with a great team articulate to talk about how they're quantifying uh the same things.
I'm looking for these patterns and I don't and I and I think it was a little bit unfair question for me to ask because the industry hasn't figured out how you actually quantify the the thing but Lynn you're you're one of the smartest people I know in the industry.
So I'm going to ask you the I'm going to ask you the hard questions.
Yeah.
Well, I mean, I think the one thing that I would say that is the challenge in that, Teeth is um, you know, if if you end up seeing some of the cuts that we're seeing in the industry and everybody's solopreneur, then how much space is there for that in the long run?
Um, so I do think that there's some downside consequences that we tend to whistle past the graveyard about that I'll put back in your camp.
Where do you see things going?
Because it is gonna completely change um the the economy long term or maybe even short term.
Yeah, I'm not I I I'm not bullish on kind of the inference flip happening in the very short term.
I think enterprises have a lot of inertia that solo clears uh get to cut through, but when it is figured out, we're going to see a complete rethinking of organizational charts.
I know you've given this a ton of thought as your team is probably a little bit more advanced and using these AI tools.
How are you going to mentor and how I'm I am I going to mentor the next generation of analysts and advisors?
We're seeing it at PwC, how they're saying, you know what, we're hiring junior analysts out of college, and we're not giving them the same assignments anymore.
We're giving them assignments on managing AI.
How do you take a Keith Towns and four plus one framework on AI infrastructure and give this now to a junior associate, and they can use my stack builder?
They can use my virtual CTO advisor uh instance to now basically bring me into meetings.
That's you know, that's a compelling uh from a productivity is compelling perspective, but what about the middle folks?
What about the folks who did that?
What about the price more?
Yeah, what about the pricing too, right?
For you.
Yeah, what about the price?
This actually puts pricing pressure pressure on me.
If somebody virtual CTO advisors out there, if you go, if you ask it to write a white paper, it will, and it comes very close to my voice.
And I open all of that.
I allow people to do it, and because I can't fight it.
So it's a threat to my own business model.
It is an amazing, amazing and scary time.
Yeah.
So later this is all has always been an incredibly fun conversation.
Folks at Intel is are doing some pretty creative stuff.
You folks are in a really interesting point in the industry.
You're still the de facto x86 leaders.
There's this question as to how much inference capacity do enterprises really need.
There will be an inference flip.
I I I'd be lying to you if I told you I knew if that inference flip is going to mean that people are gonna buy less CPUs, more CPUs.
I think you'd argue more, or if they're going to buy more GPUs or less GPUs.
Well, by the way, not no.
We are the most widely deployed host node.
So, you know, from the standpoint, when I get asked what's Intel's strategy, I mean, our strategy is to be the best solution in all the solution options.
And I think there's a lot to be said.
We're I'm working with your team on other research on the importance of CPUs and the entire AI uh uh data data uh data uh workflow we'll get into that in in future sessions but Lynn I appreciate you stopping by awesome you have you have a lot of work to do I have a lot of work to do uh if this comes out before the end of the year happy new year to everyone if if it at the beginning of the year happy new year to everyone I hope you enjoy a great holiday season.
Uh Lynn I wish you a great holiday season.
Thank you Keith.
It's always good to connect with you
