Skip to main content

Investor Event Transcript

Penguin Solutions, Inc. (PENG)

Investor Event Transcript 2026-02-28 For: 2026-02-28
Added on July 01, 2026

Conference Transcript - PENG 2026-05-12

Matt Kalitri, Analyst — Needham

All right, well, thank you everyone for joining us at our Needham Tech Media and Consumer Conference. I'm Matt Kalitri, I work on the infrastructure software and cybersecurity research team here. It's a pleasure to be joined by Penguin Solutions today. We have CEO Cass Shake and CFO Nate Olmstead. They're gonna run through some slides, do a little background, and then I've got a list of questions on my end, but please feel free to, I wanna make this interactive, So if anyone has anything, we'll have plenty of time for that, too.

Cass Shake, CEO

So, yeah, you guys take it away. Thank you for having us, Matt. We have a few slides, but we will make it conversational after these slides. So starting with some of the trends we are seeing with AI, the fundamental, AI is at a fundamental inflection point. We are always talking to our customers, understanding what's going on and how they are using AI, what's happening in their environment, and why are they investing in AI. Are they seeing ROI, not seeing ROI, and all that stuff. So we'll start with some of the industry trends that we are seeing and how we are helping our customers as AI moves from production to, AI moves from experimentation to production. AI is also moving to the next phase, which is inference. So the first phase primarily focused on model training, hyperscalers using AI for model training, and this next phase where it is much more of the implementation of AI with agentic AI using inference and really automating the workflows, which is one of the reasons the adoption for AI is going up going up because the enterprises are realizing the value of how AI, especially agentic AI, can help them. So for example, the use cases we see our customers using agentic AI range from internal usage. They are using agentic AI to automate the workflows within their customer success, customer support departments where they can have the agents help answer the questions, resolve the cases, deflect the cases. We are also seeing customers use Agent TKI for software development within their IT, even if they are not a tech company. Obviously, IT has developers that is building the solution for whether it is their e-commerce or running their infrastructure is another use case where we see agents automating the workflows, creating this code for the developers. All that is driving the demand for the infrastructure, especially when you have the agents that are working 24 by 7 on the behalf of the users. And these agents consuming more infrastructure, more compute, more memory because they are working on behalf of us 24 by 7. And AI is also moving beyond the hyperscaler to enterprises. Enterprises are deploying it across their operations, as we discussed. They are also using AI to create new business cases, which is, again, driving the demand for the infrastructure. And as AI moves from model training primarily to inference, there is a lot more need for general-purpose compute. And so you may have seen in the agentic AI adoption a lot more demand for processing CPUs beyond the GPUs. And at the same time, these general purpose processors or the CPUs are using a lot more memory beyond just a high bandwidth memory that is used with the accelerators or the GPUs. So a lot more workloads with agentic AI, a lot more usage beyond accelerators and high bandwidth memory to CPUs, as well as the memory that is becoming a key requirement, both in terms of scaling these workloads, because these agentic workloads, some of us have two or three agents. In fact, some of the software developers that are using agentic code development, they have up to 10 agents. So that is a result, a lot more infrastructure requirement beyond the accelerators to compute in the memory. This is a high-level view of some of the things I mentioned. If you look at on the left-hand side, this is the pre-inference or pre-agentic AI era, where primarily, obviously, we were using the chatbots. We are still using the chatbots, but this is before agentic AI. So primarily at that point, you would have the influence, which is primarily the chatbot you're asking a question, chat GPT or whatever you're using, it's going to give you a response. And these workflows were not as intensive because it's like a single-way communication. as AI moved from just a chat bar to agentic AI. These agents are doing the workflows. They are integrated in our system, which you see on the right-hand side. So it drives a lot more demand on the CPUs as well as the rest of the infrastructure beyond just the accelerated compute or the GPUs. So what you see in this right-hand side, the memory requirements become much higher. And this is just the general purpose memory, as we discussed. And at the same time, within the inference environment, a lot more need for the memory for the inference to be making the decisions. So to give you an idea, for example, if the agent is writing a book, There are two ways that, you know, the agent can work. Obviously, let's say the agent has written about 75% of the book. Before writing the next sentence, if the agent is not using the memory, which is direct attached, it will have to do the compute again before writing the next sentence. With having a piece of memory, which is KVCache, CXL-based KVCache is one of the instantiation that we provide allows the LLM to use the book that is written 75% within the memory, access it directly, and write the next sentence. So what is the next net effect? The net effect is LLM responses becomes much faster, and whatever task they are performing as a part of the agent, it becomes much more faster and much more easier. So the net of it is agent TKI with inference driving a need for more memory, which is related to our integrated memory business, as well as we are seeing increased demand with our AI infrastructure business, as these inference workloads require a lot more memory than the model training. This is our AI factory platform. And so we have two main businesses which are AI-driven. We have integrated memory business. We have AI infrastructure business. This is our strategy for these two businesses. We call it AI Factory Platform. And this platform is a combination of products that we provide for our customers to build and manage the AI factories and the services we provide. We provide design, build, deploy, and manage complete end-to-end services. So starting with the first element of this platform, Clusterware. Clusterware is a software. You can think of Clusterware as an operating system. Just like with Windows, for those who are using the laptop with Windows, Windows operating system uses the memory processing capabilities and allows you to run the applications on the laptop. Our Clusterware is an operating system for AI factories. As customers are building this advanced data center AI, native data data centers, AI factories, one of the approaches they can take is to configure each of the GPUs, each of the compute and networking to be able to stand up their AI factories or manage their AI factories. Alternate approaches they can use over clusterware, which allows them to create cluster for all of the infrastructure, including GPUs, including the CPUs, networking equipment, the memory enabled to manage it, build it and manage it in a much more automated way, which is the advantage for us when we are providing solutions to our customer that can simplify their deployment and management. The second piece of AI Factory Platform is a new line of products. We introduced memory AI, and as we discussed, an inference workload, memory becomes a lot more. They are much more memory intensive, and memory can increase their efficiency as well as the response, which is very critical in inference workload. So we introduced a new product in this line of products called KVCache Appliance. And this is an appliance that can accelerate the responses as we discussed. The third element of our AI factory platform is compute. We create custom compute based on the requirements of our customers. And then the fourth element is Origin AI. Origin AI is a blueprint as an end-to-end solution for AI factories that we provide to our customers that allows them to have the requirements that whatever specific workload they want to deploy, we provide them end-to-end architecture with origin AI. The fifth element, as we discussed, is our end-to-end services. We have products. We provide end-to-end services so that customers can have a single point of contact that is responsible for everything, even if some of the products we don't have, such as storage and networking, we will source it for the customers, design it and manage it on their behalf, working with our ecosystem. So that's at the high level of our AI factory platform, which is very helpful, obviously, for enterprises. We are winning customers across enterprise, sovereign AI, and we announced five new logos as a part of our key to earnings announcement, and this is how we are providing them the solution. So at the high level to net it all, AI factory business providing product and services as a full stack AI factories that customers can deploy, large financials, very large enterprises are building their own AI factories so they can have much more performance latency that is needed in inference workloads. Obviously, economics is much better for them. They are using cloud for model training. They are using inference and agentic AI for on-premise factories as they build their factories. Governments around the world, they are building sovereign AI factories. And the idea is they want sovereignty. They don't want to rely on US-based infrastructure. So we built one of the largest AI factories in South Korea with our Hain cluster, which is giving us the visibility to help other countries around the world where we are working with them on building their own sovereign AI factories. And a memory AI line of products, we discussed our memory business, having the memory business and AI infrastructure business gives us the insights, deep insights for our memory architectures, deep insights for what is needed in this next phase of AI within France. And that has allowed us to build these kind of appliances that are, one, very unique. Nobody else in the market has a product like memory AI, which is helping our customers and helping us win the deals. And last but not least, our integrated memory business where we create specialized memory card for large vendors, OEMs such as Cisco, Google, and some of the names you see on the slide. So memory business driven by AI, AI infrastructure business, using some of the differentiation with our understanding of memory to provide full-stack solution for our customers as a part of these two main businesses that we have within our company.

Matt Kalitri, Analyst — Needham

And I will leave you all with this.

Cass Shake, CEO

So, you know, we have three businesses at the high level within the company, as we discussed, integrated memory business, AI factory platform business. We also have an LED business. But the high-level view of our AI-driven business between integrated memory and non-hyperscalary AI HPC business, that makes up for 60% of the business in the first half. And that business in the first half grew about 50%. This is where, you know, this business is AI-driven between memory and AI. So just wanted to show you guys the view of the scale of the business on a 100 basis. You know, there's well over $1 billion growing in excess of about 40 to 50 percent. So we are very excited about all the build-outs that are happening in the market and how customers are using it. And it's an exciting time to be in this business providing memory solutions and AI factory platform. You want to add anything, Mirit, that I may have missed? Thank you.

Nate Olmstead, CFO

You covered it, so thank you.

Matt Kalitri, Analyst — Needham

Great, yeah, well, thank you so much for the breakdown. And it is a truly exciting time to be at that intersection of memory and AI infrastructure like you guys talk about. And I want to get into some of the more product-specific stuff, but maybe to start, Cash, you took over as CEO in February. What drew you to the role, and how the first couple months in the CPEN?

Cass Shake, CEO

Yeah, it's a good question. So I have been in this space, infrastructure across large companies for 30-plus years, between large companies and small companies, actually. So I've spent time at Cisco. When Cisco was building the data center products, entering into the compute space, I was at HPE in the networking business unit. And then, more interestingly, the relevant piece, more relevant piece is my time at Dell. I ran Dell's AI HPC business. This was before AI HPC was as hard as it is right now, between 2016 through 2020. And interesting coincidence was I actually competed against, it was a billion dollar business. Obviously, Dell is a scalable company. It was a billion dollar business when I took over. And we competed against Penguin Computing, which was the company that the parent company acquired. But anyway, having been in this space, seeing the requirements and understanding of the market, when I found out about the opportunity at Penguin, I was excited about that, so knowing the space, as well as the opportunity here, as in, as we discussed the memory business, as well as the AI infrastructure business, that got me excited about the opportunity, And then, you know, obviously talking to the team and talking to former CEO, I realized that there is a lot of opportunity here to grow the business. So that drew me to the business in terms of why Penguin. And to answer the other part of your question about what have I learned, some of the things I shared, I would say, are the high level. I've spent a lot of time, obviously, with the customers around the world. I visited pretty much all of our regions, talking to our team members, started talking to customers, talking to partners. There is an inflection point happening within AI, from model training, experimentation, to production, to agentic AI. And that is creating a unique advantage for us. That's why we have prioritized our AI factory business. We came up with the AI factory platform strategy. We are investing more in innovation, investing more in our memory AI appliances. We are investing more in our clusterware AI. We are planning to make it agentic so that, you know, it can perform all the tasks on its own, only with the human in the loop. But those are some of the things that I have learned in terms of what's happening in the industry. And it's, as we discussed, an exciting time to be in this space.

Matt Kalitri, Analyst — Needham

Yeah, absolutely. That's great. What have been your first couple of key initiatives as you stepped in?

Cass Shake, CEO

So number one is more investment both in product and R&D for our AI factory platform. This is the growth strategy which we believe can help us capture more demand that we are seeing in this segment. Second initiative is just like we discussed, all of these enterprises are using AI within their own operations, company operations. We have a priority where we are investing in AI across the company. We have a software development team that is now using Cursor to develop the code with the agents, and we have introduced Agenting AI in our customer success team. So all of our teams, including myself, we are investing time and resources to what I call drink our own champagne and take advantage of the opportunity and efficiency that AI provides.

Matt Kalitri, Analyst — Needham

Great to hear. You guys reported very solid second quarter results. When you think about what went well in the quarter from an execution standpoint, what stands out and what's your focus as you look to the rest of the year?

Cass Shake, CEO

Yep. So a couple of things, as we discussed in the earnings announcement, memory business is performing really well. And as we discussed, this is the driver for this demand is AI, agentic AI that we are seeing, which is, you know, driving the demand for the business, as well as even if the pricing is one of the advantages, there is a supply constraint in this market, given all the requirements AI is putting on the memory. But the demand, we believe, is much more durable based on the fact that it is AI-driven. The team executed really well in the opportunity for integrated memory business. Our AI HPC business, we acquired five new logos, and some of these are very large customers. Tier 1 Financial, top 10 energy, one of the top 10 energy companies. Deepgram, which was a logo we shared publicly which we acquired as a part of our partnership with Dell and our partnership with NVIDIA is also getting stronger as NVIDIA is entering into enterprise space given some of these hyperscalers which was the driver for their growth are developing their own chips. So that is creating a synergy for us.

Matt Kalitri, Analyst — Needham

Sorry, to stick with memory for a second, You made a point on the call to emphasize the opportunity that the memory AI line unlocks, and particularly as memory becomes more central to AI infrastructure and production. I know you hit on this a bit in the slides, but what exactly is the goal with that product line, and how has the memory business evolved at Penguin?

Cass Shake, CEO

So that product line, Memory AI KVCache Server, is one of the many products that we are going to provide to our customers. So Memory AI product line, the reason we were able to develop is we have deep understanding of the memory. We have deep understanding of memory AI infrastructure architecture. And even if it is built within our integrated memory team, the buyer for that product is essentially the AI infrastructure buyer. So, for example, this customer we acquired, this customer, tier one financial customer we acquired in Q2, they acquired memory KVCache server. So KVCache is a technology that allows you to, as we discussed, keep the content accessible much more faster for the accelerator to use that content and respond to the queries much faster. There are other appliances in the line of memory AI that we are working on. So we were an early investor in Celestial AI, Celestial AI's photonics memory, high bandwidth photonics memory company that got acquired by Marvell for $5.5 billion free revenue. So in addition to being the beneficiary of receiving some of the proceeds from that acquisition as an early investor, we continue to work with Celestial AI to develop photonics memory appliance. That Photonics memory appliance will provide even more access, faster access, to the memory required in the inference workloads. So this is, you know, obviously a growing area, and as enterprises continue to deploy inference, that's going to help them with the memory AI appliances. And the memory AI appliances is really the representation of the uniqueness of the company at the intersection of memory and the AI infrastructure. So having the ability to develop the products and drive the synergy and provide the products to our AI infrastructure, customers are truly a unique advantage for Penguin.

Matt Kalitri, Analyst — Needham

We've seen no shortage of headlines about memory shortages and stuff. And I was even reading recently about, like, worries about a strikeover at Samsung. And what are you guys seeing in that DRAM memory market as far as supply chain, price increases, stuff like that?

Cass Shake, CEO

Yeah, supply is constrained, obviously, as we discussed, because of all of the demand of memory, whether it is high bandwidth memory or the regular memory. One of the advantages we have is we've been in this business for 40 years. So we have a relationship with all large memory manufacturers, which is allowing us to have the access to the memory to serve the demand of our customers. In one of the main suppliers, SK Hynix, we have a deep relationship with them in addition to working with them for decades. SK is an investor within the company. So we work closely with SK Telecom and other SK companies, and having that as a part of our relationship allows us to have the access. But the supply is constrained, so we have a backlog, but based on our relationship, we've been able to access it and serve as much demand as we can. Nate, I don't know if you want to add anything to it.

Nate Olmstead, CFO

You've been reading the same things, right? It's tight. We're fighting every day for supply, building a very strong backlog, going out a few quarters. But Cash is right. Our relationship probably helps us a little bit at the margin, but it's very tight.

Matt Kalitri, Analyst — Needham

Yeah, very important, though. Turning to advanced compute, good quarter there, but actually end up reducing guidance looking forward. What changed during the quarter, and how are you guys looking at that business?

Cass Shake, CEO

Yeah. So that business, very strong pipeline, very strong bookings and bookings growth in Q2, I'd say above market conversion. The challenge in that business is by the time we book to the time we recognize the revenue. It's about three months. However, recently, it is between three to six months. And part of the reason is we are landing new logos, acquiring new logos, as we discussed, five new logos in Q2. When you have a new customer, you have a new process. It takes longer for them to get to the production, get to a state where we can recognize the revenue. The other challenge is also the material availability. In some cases, it's taking longer. So the good news is business is doing really, really well. The challenge is time to book to revenue is increasing. So for example, when we announced our Q2 results about six weeks ago, even at that point, most of the bookings from that point onwards for the second half will be recognized in the first half of next year. So it's really a timing issue, but we are seeing pretty strong demand and it's a matter of execution for us.

Matt Kalitri, Analyst — Needham

And as you work through some of those timing issues and stuff, what's the best way for us to sort of judge the underlying momentum of the business? What are you guys looking at to sort of determine you're on track there?

Cass Shake, CEO

Yeah, so part of that execution is we've been focused on diversifying the business. So acquiring new logos, continue to have new customers is one of the metrics we keep in mind. And diversification is working well as we discussed new logo acquisition. And we have several new logos in the pipe for Q3 that we are working on executing. So bookings, obviously, is a leading indicator. And we are very closely monitoring the bookings. And as we discussed, the bookings are growing pretty strong. And revenue is a lagging indicator. and it's a matter of timing. So making sure we are focused on the right segments, as we discussed, making sure that we are leveraging our differentiation with the AI factory platform, and we feel confident that based on other execution that we will continue to deliver on our commitments.

Matt Kalitri, Analyst — Needham

I want to make sure I'm not chewing up all the time here, too. Do you have any questions from the audience?

Audience Member, Analyst — Audience Member

Yeah.

Cass Shake, CEO

Right, right. So first of all, you know, it depends on the department. So let's say if we are talking about software engineering, which is, as we discussed, we are investing more in software engineering, but making them, first of all, making them the tool available that they can use. In addition to giving them the tools and access to the tools, we are working on providing them the training because they need help in terms of how to effectively use the tool. But at the high level, in this use case, the goal will be how much of the code can be developed by AI. So, for example, in my previous company, which was an agentic AI software-focused company, we were developing 30% by the time I left end of January, we were developing 30% of the code by AI. So what does that mean? Let's say if I had 100 engineers, I was developing the code equivalent to 130 engineers. So that's the value that we saw, and that's kind of the goal over here for the amount of headcount we have. How do we make them more efficient so we get more for what we are investing? Now, in terms of the usage and the tokens, as long as we provide them the right guidelines, as in it's all about providing the guideline how to effectively use the tool, because tool is a tool. If we help them use the tool effectively, then they will be able to create the efficiency we are expecting, and we are not going to spend as much as, you know, they can expend if they are not being trained. So I think training and enablement is equally important as making them the tool available and helping them realize why we are doing it. Because, you know, obviously a lot of people are worried that, you know, if they use the tool, they can lose their job. There's always this discussion. So we've been very upfront. And one of my advantages, having done it in the previous company, and we were on the leading edge and having gone through this process, we've been upfront that we're not focused on reducing the headcount. We want to create a competitive advantage that if we don't use it, we're going to have a disadvantage against those companies that are not using it. But we are not using it to reduce the headcount. We want to move fast. We want to deliver more software capabilities, as an example in this case. That's how we are working with our team members and helping them understand while providing them the training and coaching to use the tools effectively. And that's an example. I mean, different departments have different tools and different outcomes we are expecting. and we are making sure that they understand this is how they should use the tool and this is what we will be measuring for us to see the ROI on our investments across the company.

Matt Kalitri, Analyst — Needham

I'll jump back in here. One, oh, sorry.

Cass Shake, CEO

So it is not, to answer your specific question, that all enterprises are looking at the efficiencies like us. To be honest, the answer is no. I see a range of clarity on why enterprises are using it and if they have specific outcomes in mind. We talked to some of the enterprises where they are still experimenting. So they're making it available to all of the users without having a specific goal in mind, Which, in my view, is still effective because the first thing is getting over the hump of having team members understand this is not going to replace your job, right? But at the same time, while that is the first goal, which is to make sure everybody is using it, having the clarity is also important and may or may not be the case because some of the enterprises are still learning. But then they quickly realize they may not have the goal up front. As soon as they start burning the tokens faster, then it becomes a discussion of, you know, are we seeing the ROI? In order to see the ROI, let's get back to the focus, which is getting more for the investments we are making. So I see a spectrum of companies that may be not exactly using the starting point, where they are coming to the same ending point, which is focusing on the ROI. I think this year, especially in the last three to six months, where we have seen a lot more production deployment, experimentation to production, which is why you see a lot of demand across companies that are providing the infrastructure for that. So I'd say that the adoption will increase significantly in the next one year. because, especially for those who have been using it, the tools are becoming more effective as well. So that's another dynamic because these agents, the agents have made it much more impactful when they are automating the workflows, when they are providing the outcomes in a way that, you know, you're getting more for each of the employees from their outcome perspective. Now, as with any other technology, as we know, some of them, some companies are embracing it sooner than others. So it's going to vary based on where they are in their cycle of using the technology to differentiate and invest in the technology. But for those adopters who have either started or will started, I think the technology will show more returns, which is why the adoption will continue to increase in the next one year.

Matt Kalitri, Analyst — Needham

Well, one thing we've spoken about is before you can get to end user adoption, you have to have everything set up on the back end and just how difficult that is for companies to stand up. And so they'll turn to a solution like Penguin to get that done. What are you guys seeing in terms of the split of organizations that, like, have the internal sophistication that they sort of just need racks as quick as they can get them versus companies that maybe try to do it on their own and then realize they really need this full stack solution from Penguin?

Cass Shake, CEO

Yeah. So let me provide a little more context for the audience in terms of how enterprises are approaching AI infrastructure. So the first thing, whether they are enterprises or mid-market, the easiest thing for them is to access a cloud-based infrastructure, right? Go to AWS or one of the new cloud providers get, you know, access to the infrastructure. That is the facet. They don't have to worry about setting it up. What happens is, you know, acquisition is easier, your team gets started, they start using it. What we are finding is the reflection point for them to move from just using cloud-based to building their own factories is the scale. So when you are at scale, especially very large enterprises, for example, this large tier one financial bank we acquired, they have 40,000 developers within IT. And then when those 40,000 developers started using Agendic AI, the budget they had for one year, they exhausted that budget in three months. And then they started realizing we need to have our own AI factory so that I'm not just using the subscription as I scale. So they started with, you know, they are going to build their own factory. Then they started looking at options. Are we going to build it on our own or we are going to go find another one? They put out an RFP with a very formal process, and we were among one of the many who responded to that RFP because they had decided that they need help, to your point, right? And then at the end, based on our value between having the products and the services helped us win that account, to your point why they are selecting us, is because we provide the full stack end-to-end. So rather than them going to one company for compute, another company for storage, and so on and so forth, we have the ability with our AI factory platform to provide them the solution. So we've provided them the solution, and now they have a setup where they are. So interestingly, I asked the CIO, are they going to move on-premise based on the solution that they have from us with some expansion? The answer was no, they are going to continue with the cloud for model training, and they are going to expand their on-premise factory for inference and agent TKI workloads. So that's kind of the dynamic of why enterprises are building their own factories, and when we go in, why it is more relevant for them to work with a provider like us.

Matt Kalitri, Analyst — Needham

And when you say that they're building out their on-premise for infrastructure, are those conversations becoming more prevalent as we see companies sort of move from training into inference?

Cass Shake, CEO

Yes, we are seeing enterprises. But these are large enterprises that I mentioned. Large enterprises and regulated industries, whether it is oil and gas and others, where there are sovereignty and privacy requirements in addition to the economics, which is at scale, is not feasible for them in the cloud. Awesome. We've probably got time. Yep.

Audience Member, Analyst — Audience Member

Most of what you're doing is.

Cass Shake, CEO

Yeah. So what AI factory platform, we have products that we provide as a part of the system integration. So the cluster where AI is a product, a software product, that is a part of managing, deploying and managing infrastructure. It's a heterogeneous product. So it is vendor agnostic that provides them much more simplicity and it's operating system of AI factories. Memory AI is a product that we provide. That's our own IP that we provide as a part of the solution. And we have compute that we provide. And then we resell the other products that are part of the infrastructure because there's storage and there's networking and so on and so forth. So we have our own IP. And in fact, as a part of the AI factory platform strategy, as we discussed, I'm going to invest more in our IP so that we provide them unique solutions beyond the compute and the storage. We are reselling it, so we have prices, but then we have the resell business as a part of selling the compute.

Nate Olmstead, CFO

Pass through. With some margin on it for the integration work and the supply chain and all that. Yeah, one more quick one.

Cass Shake, CEO

We have not shared the long-term growth rates for those markets publicly, but you can see the growth in those markets, right? So even if, you know, we have not shared the specifics, those are very high-growth markets where we are focused on between the large enterprises, sovereign AI, and neocloud infrastructure. All right.

Matt Kalitri, Analyst — Needham

I think we have to wrap it there, unfortunately. But Cash, Nate, thank you so much for being with us. If anyone is interested in learning more, we'd be more than happy to connect here with the team. And we appreciate everyone's support.