STOCK TITAN

VMware Cloud Foundation Brings Leading AI Models to the Private AI Cloud

(Neutral)
(Very Positive)
Tags
AI

Broadcom (NASDAQ: AVGO) announced that leading AI models from Google, NVIDIA, NEC, Alibaba Cloud, and Z.ai have been tested and validated to run on VMware Cloud Foundation (VCF), enabling enterprises to deploy these models on-premises and offer “model as a service” from their private clouds.

According to Broadcom, VCF is an AI- and Kubernetes-native private cloud platform that can run AI inferencing, agentic applications, containers, and virtual machines together, with MLPerf Inference v5.1 benchmarks showing performance on par with bare metal. Newly validated models include Nemotron 3, Gemma 4, cotomi, Qwen 3.7-Max, and GLM 5.2, expanding VCF’s open ecosystem that already supports mixed compute across AMD, Intel, and NVIDIA and more than 150 open source models via vLLM.

Loading...
Loading translation...

Positive

  • None.

Negative

  • None.

Market Context

The platform record shows an AI-tagged average move of 0.19% across five events, adding historical c...
Analysis

The platform record shows an AI-tagged average move of 0.19% across five events, adding historical context to this model-validation announcement. Recent insider activity was net selling; subsequent disclosures remain the key watchpoint.

Key Figures

Private-cloud AI adoption: 56% Token-efficiency improvement: 40% Open-source models supported: more than 150 models +2 more
5 metrics
Private-cloud AI adoption 56% Enterprises already running or planning production AI inferencing on private cloud
Token-efficiency improvement 40% NEC cotomi model
Open-source models supported more than 150 models VCF with vLLM as the default model runtime
Context length 1 million context Nemotron 3 model
Context window one-million-token Qwen 3.7-Max model

Previous AI Reports

5 past events · Latest: Aug 17 (Positive)
Same Type Pattern 5 events
Date Event Sentiment 24h Move Catalyst
Aug 17 AI conference announcement Positive -0.1% VMware Explore announced private AI cloud sessions, labs, certifications, and keynote programming.
Jun 24 AI processor launch Positive +0.5% Broadcom and OpenAI introduced Jalapeño, an LLM-optimized intelligence processor.
Jun 09 AI financing platform Positive -1.1% Broadcom, Apollo, and Blackstone launched an AI compute financing vehicle.
Jun 09 Private cloud outlook Positive -1.1% Broadcom reported private-cloud adoption trends for production AI inference workloads.
Jun 08 AI security investment Positive +2.8% Broadcom expanded security investments across the Spring and Java ecosystems.

24h Move is the share-price change in the day after each event; other market factors may also have contributed.

Pattern Detected

AI-tagged announcements produced mixed reactions, with three divergences and two alignments between positive news sentiment and the subsequent 24-hour move.

Key Terms

mlperf inference, data sovereignty, multimodal, model context protocol
4 terms
mlperf inference technical
"Independent benchmark testing under MLPerf Inference v5.1 standards"
A standardized benchmark suite that measures how well computer systems run trained machine-learning models to make predictions or classifications in real-world conditions. It reports metrics such as speed, latency and throughput across different hardware and software setups, letting buyers compare performance like a fuel-economy test for cars. Investors use it as a neutral signal of a product’s efficiency, scalability and competitive strength in AI workloads.
data sovereignty regulatory
"a clear path to data sovereignty and cost-effective AI at scale"
Data sovereignty is the principle that digital information is subject to the laws and control of the country or entity where it is stored or processed. For investors, it matters because where data lives affects a company's legal obligations, costs, ability to sell services across borders, and exposure to government access or restrictions — like owning a house that must follow the rules of the town it sits in.
multimodal technical
"open, multimodal models delivers leading accuracy and efficiency"
Multimodal describes an approach, product, or system that uses two or more different types of inputs, methods, or channels — for example combining text, images and audio in a technology product, or blending drugs, devices and therapy in medical care. For investors, multimodal solutions can broaden market reach and competitive differentiation but also add development cost, operational complexity and regulatory hurdles; think of it like a hybrid car that offers more capabilities but requires more parts and oversight.
model context protocol technical
"VCF’s zero-trust security architecture and Model Context Protocol support"
A model context protocol is a set of rules or guidelines that determine how a financial model interprets and applies information within a specific situation. It helps ensure consistent and accurate analysis by clarifying what data or assumptions are relevant in a given scenario. For investors, it provides clarity on how predictions or assessments are made, increasing confidence in decision-making.

AI-generated analysis. How Rhea-AI works. Not financial advice.

See more from StockTitan in Google Search and AI answers. Adds StockTitan as a preferred source · opens Google
Add on Google

Nemotron 3, Gemma 4, cotomi, Qwen3.7-Max, and GLM-5.2 Validated to Run on VMware Cloud Foundation

LAS VEGAS, Aug. 31, 2026 (GLOBE NEWSWIRE) -- VMware Explore 2026 -- Broadcom Inc. (NASDAQ: AVGO), a global technology leader that designs, develops, and supplies semiconductor and infrastructure software solutions, today announced that leading AI models from providers including Google, NVIDIA, NEC, Alibaba Cloud, and Z.ai, are validated to run on VMware Cloud Foundation (VCF), enabling customers to bring these AI models on-premises and deliver “model as a service” to their users. As enterprises accelerate the movement of inference workloads to private cloud driven by privacy, security, governance, and AI tokenomics, VCF provides the AI- and Kubernetes-native platform that gives organizations across industries a production-ready path to deploying a wide variety of AI models.

According to Broadcom's Private Cloud Outlook 2026, 56% of enterprises are already running or planning to run production AI inferencing on private cloud. VCF is a unified private cloud platform capable of running inference workloads, agentic applications, containerized services, and traditional VMs together, eliminating the operational fragmentation of managing separate stacks. Independent benchmark testing under MLPerf Inference v5.1 standards confirms that VCF delivers performance on par with bare metal, making it the ideal platform of choice for enterprises deploying AI at scale on premises.

“Broadcom is committed to giving enterprises the broadest set of AI models for their on-premises infrastructure, all validated on VCF,” said Chris Wolf, global head of AI and advanced services, VMware Cloud Foundation Division, Broadcom. “Working with the world’s leading AI model providers, we’re giving organizations a clear path to data sovereignty and cost-effective AI at scale, with leading models available securely and delivered as a service to their user community through VMware Cloud Foundation’s built-in services.”

Leading Models Validated for VMware Cloud Foundation
VCF empowers enterprises to accelerate AI workload deployment at lower costs through an open and extensible ecosystem. Support for mixed compute across AMD, Intel, and NVIDIA frees enterprises to choose their preferred GPU and CPU hardware for their AI workloads. Leveraging vLLM as the default model runtime gives customers the ability to run more than 150 open source models on VCF. Today, Broadcom is announcing the following models have been tested and validated to run on VCF:

  • Nemotron 3: The NVIDIA Nemotron 3 family of open, multimodal models delivers leading accuracy and efficiency to help agents complete tasks faster. Combining hybrid Mamba-Transformer MoE architecture, 1 million context and multi-environment reinforcement learning, Nemotron 3 enables scalable, long-running agentic workflows across enterprise applications.
  • Gemma 4: Google DeepMind's latest open source, open-weight multimodal model family, purpose-built for developers and the research community for bringing local execution, and enabling enterprises to build and deploy autonomous AI agents.
  • cotomi: NEC's proprietary AI model optimized for Japanese language, trained on curated, highly reliable datasets. It empowers enterprises by seamlessly combining high-speed processing with a 40% improvement in token efficiency.
  • Qwen 3.7-Max: Alibaba's Qwen 3.7-Max is a proprietary multimodal model that offers impressive one-million-token context windows, advanced multimodal reasoning, and agentic-era design, giving global enterprises sovereign, on-premises access to one of the world's most capable AI model families.
  • GLM 5.2: Z.ai (formerly Zhipu AI)'s open source General Language Model enables enterprises to deploy coding and reasoning agents locally for multi-step autonomous workflows with data sovereignty and optimal hardware performance.

Partner and Industry Commentary
“VMware Cloud Foundation gives enterprises the secure, governed foundation they need to put Gemma 4 to work across their most demanding workloads,” said Olivier Lacombe, Director of Product Management, Google DeepMind. “By combining Gemma 4’s multimodal reasoning and agentic capabilities with VCF’s zero-trust security architecture and Model Context Protocol support, enterprises can build and deploy autonomous AI agents entirely within their own infrastructure without compromising on data sovereignty or operational control.”

“NEC’s cotomi was purpose-built for enterprises that need AI to truly understand the nuances of Japanese language and business context, and VMware Cloud Foundation gives those enterprises the private, secure infrastructure to deploy it at scale. With cotomi running on VMware Cloud Foundation, organizations gain an AI model capable of acting as a secure autonomous agent across their workplace systems — without sensitive data ever leaving their own environment,” said Akio Yamada, Chief AI Officer, NEC Corporation. “Furthermore, NEC offers the NEC Private Infrastructure powered by VMware service, providing customers with a private cloud environment hosted in NEC data centers. We will continue to explore opportunities to further integrate cotomi with VMware Cloud Foundation.”

“Enterprise customers need AI inference that respects data sovereignty and regulatory boundaries, particularly in markets where data sovereignty, not just residency, is non-negotiable,” said Craig McLellan, CEO of ThinkOn. “By deploying advanced, enterprise-grade AI models on VMware Cloud Foundation, ThinkOn delivers a sovereign AI framework that gives organizations complete control over their intellectual property. Customers gain production-ready inference capabilities on private cloud infrastructure, with the governance and compliance controls their industries require.”

Additional Resources

About VMware Explore
VMware Explore is the established cloud event for IT professionals to advance their skills and credentials. VMware Explore 2026 will connect attendees with technology experts and solution architects who are solving real-world challenges. Focused technical sessions, Hands-on Labs, and certification opportunities will enable attendees to apply their learnings and deliver greater value back to their organizations. Attendees will leave with the training, tools, and strategies to build and operate a modern private cloud that is AI-native—running any workload while keeping data local and secure. To learn more about VMware Explore, please visit: https://www.vmware.com/explore

About Broadcom
Broadcom Inc. (NASDAQ: AVGO) is a technology leader that designs, develops, and supplies semiconductors and infrastructure software for global organizations' complex, mission-critical needs. Broadcom combines long-term R&D investment with superb execution to deliver the best technology, at scale. Broadcom is a Delaware corporation headquartered in Palo Alto, CA. For more information, visit www.broadcom.com.

Broadcom, the pulse logo, and Connecting everything are among the trademarks of Broadcom. The term "Broadcom" refers to Broadcom Inc., and/or its subsidiaries. Other trademarks are the property of their respective owners.

Media Contact:
Roger T. Fortier
VMware Cloud Foundation Division, Broadcom
+1.408.348.1569
roger.fortier@broadcom.com


FAQ

What did Broadcom (NASDAQ: AVGO) announce about AI models on VMware Cloud Foundation in August 2026?

Broadcom announced that several leading AI models are validated to run on VMware Cloud Foundation, enabling on-premises deployment. According to Broadcom, these models allow enterprises to deliver “model as a service” from private clouds while maintaining privacy, security, governance, and control over data.

Which AI models are now validated on VMware Cloud Foundation for AVGO customers?

The validated models include Nemotron 3, Gemma 4, cotomi, Qwen 3.7-Max, and GLM 5.2. According to Broadcom, these come from providers such as NVIDIA, Google DeepMind, NEC, Alibaba Cloud, and Z.ai, giving enterprises on-premises access to multimodal, language, and agentic-capable AI models.

How does VMware Cloud Foundation support private AI clouds for Broadcom (AVGO) enterprise customers?

VMware Cloud Foundation provides an AI- and Kubernetes-native private cloud platform that runs AI inference, agentic apps, containers, and VMs together. According to Broadcom, this unified stack helps eliminate operational fragmentation and supports mixed compute across AMD, Intel, and NVIDIA for flexible AI infrastructure choices.

What performance claims does Broadcom make about AI workloads on VMware Cloud Foundation?

Broadcom states that independent MLPerf Inference v5.1 benchmarks show VMware Cloud Foundation delivers performance on par with bare metal. According to Broadcom, this makes VCF suitable for large-scale, on-premises AI deployment while still offering virtualization benefits and integrated governance and security controls.

How many AI models can enterprises run on VMware Cloud Foundation using vLLM?

Broadcom indicates that using vLLM as the default model runtime lets customers run more than 150 open source models on VMware Cloud Foundation. According to Broadcom, this broad model coverage expands choice while the newly validated partner models further extend the platform’s AI ecosystem.

How does VMware Cloud Foundation help AVGO customers address data sovereignty for AI workloads?

VMware Cloud Foundation enables enterprises to run AI models entirely within their own infrastructure, supporting data sovereignty. According to Broadcom and partners, customers can keep sensitive data on-premises while deploying autonomous and multimodal AI agents under their own security, governance, and compliance frameworks.