# esx

Published articles for esx.

This is one page of public article previews, not the complete archive. Follow Next page to continue. Summaries are not the original full articles.

## Optimizing AI Deployments with VMware Cloud Foundation (Part 3)

DevFeed: [Optimizing AI Deployments with VMware Cloud Foundation (Part 3)](<https://devfeed.tech/articles/optimizing-ai-deployments-with-vmware-cloud-foundation-part-3-65072.md>)

Original publisher: [Read original article](<https://blogs.vmware.com/cloud-foundation/2026/10/05/optimizing-ai-deployments-with-vmware-cloud-foundation-part-3/>)

Author: hari sivaraman lan vu and uday kurkure

Published: 2026-10-05T21:20:23Z

Content type: article

Language: en

Sources: [VMware](<https://devfeed.tech/sources/vmware-blogs.md>)

Topics: [VMware Cloud Foundation](<https://devfeed.tech/topics/vcf-9-1.md>), [cost-optimization](<https://devfeed.tech/topics/cost-optimization.md>), [AI Infrastructure](<https://devfeed.tech/topics/ai-infrastructure.md>), [quantization](<https://devfeed.tech/topics/quantization.md>)

Tags: [agentic-ai](<https://devfeed.tech/tags/agentic-ai.md>), [ai-inference](<https://devfeed.tech/tags/ai-inference.md>), [ai-infrastructure](<https://devfeed.tech/tags/ai-infrastructure.md>), [ai-infrastructure-optimization](<https://devfeed.tech/tags/ai-infrastructure-optimization.md>), [ai-ml](<https://devfeed.tech/tags/ai-ml.md>), [ai-performance](<https://devfeed.tech/tags/ai-performance.md>), [cloud-infrastructure](<https://devfeed.tech/tags/cloud-infrastructure.md>), [cpu-utilization](<https://devfeed.tech/tags/cpu-utilization.md>), [esx](<https://devfeed.tech/tags/esx.md>), [hammerdb](<https://devfeed.tech/tags/hammerdb.md>), [home-page](<https://devfeed.tech/tags/home-page.md>), [infrastructure](<https://devfeed.tech/tags/infrastructure.md>), [large-language-models](<https://devfeed.tech/tags/large-language-models.md>), [multi-tenancy](<https://devfeed.tech/tags/multi-tenancy.md>), [performance](<https://devfeed.tech/tags/performance.md>), [resource-utilization](<https://devfeed.tech/tags/resource-utilization.md>), [tco-optimization](<https://devfeed.tech/tags/tco-optimization.md>), [tensorrt-llm](<https://devfeed.tech/tags/tensorrt-llm.md>), [tpc-c](<https://devfeed.tech/tags/tpc-c.md>), [vcf-9-1](<https://devfeed.tech/tags/vcf-9-1.md>), [vmware-cloud-foundation](<https://devfeed.tech/tags/vmware-cloud-foundation.md>), [vsphere](<https://devfeed.tech/tags/vsphere.md>), [workload-consolidation](<https://devfeed.tech/tags/workload-consolidation.md>)

### AI overview

This article reports tests of VMware Cloud Foundation 9.1 for running GPU accelerated LLM inference with reduced VM CPU and memory allocations and co-locating inference and CPU intensive workloads. The reported tests found little or no inference performance change across the tested VM configurations and no measurable performance impact when the inference and database workloads shared a server.

### Source excerpt

Maximizing CPU and Memory Utilization for Best TCO Introduction VMware Cloud Foundation (VCF) can help enterprises maximize utilization and optimize AI infrastructure. To demonstrate VCF's efficiency and flexibility for running AI workloads in various configurations, the Broadcom Performance team has completed extensive performance tests. We present these results in our white paper and are publishing ... Continued The post Optimizing AI Deployments with VMware Cloud Foundation (Part 3) appeared first on VMware Blogs.