{"id":58148,"date":"2023-08-25T20:39:56","date_gmt":"2023-08-25T12:39:56","guid":{"rendered":"http:\/\/www.upgrademag.com\/web\/?p=58148"},"modified":"2023-08-25T20:40:00","modified_gmt":"2023-08-25T12:40:00","slug":"vmware-nvidia-unlock-generative-ai-for-enterprises","status":"publish","type":"post","link":"http:\/\/www.upgrademag.com\/web\/2023\/08\/25\/vmware-nvidia-unlock-generative-ai-for-enterprises\/","title":{"rendered":"VMware, NVIDIA unlock generative AI for enterprises"},"content":{"rendered":"\n<p class=\"wp-block-paragraph\"><strong>VMware Inc. and NVIDIA announced the expansion of their strategic partnership to ready the hundreds of thousands of enterprises that run on VMware\u2019s cloud infrastructure for the era of generative AI.<\/strong><\/p>\n\n\n\n<p class=\"wp-block-paragraph\">VMware Private AI Foundation with NVIDIA will enable enterprises to customize models and run generative AI applications, including intelligent chatbots, assistants, search and summarization.\u00a0The platform will be a fully integrated solution featuring generative AI software and accelerated computing from NVIDIA, built on VMware Cloud Foundation and optimized for AI.<\/p>\n\n\n\n<p class=\"wp-block-paragraph\">\u201cGenerative AI and multi-cloud are the perfect match,\u201d said Raghu Raghuram, CEO, VMware. \u201cCustomer data is everywhere \u2014 in their data centers, at the edge, and in their clouds. Together with NVIDIA, we\u2019ll empower enterprises to run their generative AI workloads adjacent to their data with confidence while addressing their corporate data privacy, security and control concerns.\u201d<\/p>\n\n\n\n<p class=\"wp-block-paragraph\">\u201cEnterprises everywhere are racing to integrate generative AI into their businesses,\u201d said Jensen Huang, founder and CEO, NVIDIA. \u201cOur expanded collaboration with VMware will offer hundreds of thousands of customers \u2014 across financial services, healthcare, manufacturing and more \u2014 the full-stack software and computing they need to unlock the potential of generative AI using custom applications built with their own data.\u201d<\/p>\n\n\n\n<p class=\"wp-block-paragraph\"><strong>Full-Stack Computing to Supercharge Generative AI<\/strong><\/p>\n\n\n\n<p class=\"wp-block-paragraph\">To achieve business benefits faster, enterprises are seeking to streamline development, testing and deployment of generative AI applications. McKinsey estimates that generative AI could add up to $4.4 trillion annually to the global economy.<\/p>\n\n\n\n<p class=\"wp-block-paragraph\">VMware Private AI Foundation with NVIDIA will enable enterprises to harness this capability, customizing large language models; producing more secure and private models for their internal usage; and offering generative AI as a service to their users; and, more securely running inference workloads at scale.<\/p>\n\n\n\n<p class=\"wp-block-paragraph\">The platform is expected to include integrated AI tools to empower enterprises to run proven models trained on their private data in a cost-efficient manner. To be built on VMware Cloud Foundation and NVIDIA AI Enterprise software, the platform\u2019s expected benefits will include:<\/p>\n\n\n\n<ul class=\"wp-block-list\">\n<li>Privacy \u2014 Will enable customers to easily run AI services adjacent to wherever they have data with an architecture that preserves data privacy and enable secure access.<\/li>\n\n\n\n<li>Choice \u2014 Enterprises will have a wide choice in where to build and run their models \u2014 from NVIDIA NeMo&#x2122; to Llama 2 and beyond \u2014 including leading OEM hardware configurations and, in the future, on public cloud and service provider offerings.<\/li>\n\n\n\n<li>Performance \u2014 Running on NVIDIA accelerated infrastructure will deliver performance equal to and even exceeding bare metal in some use cases, as proven in recent industry benchmarks.<\/li>\n\n\n\n<li>Data-Center Scale \u2014 GPU scaling optimizations in virtualized environments will enable AI workloads to scale across up to 16 vGPUs\/GPUs in a single virtual machine and across multiple nodes to speed generative AI model fine-tuning and deployment.<\/li>\n\n\n\n<li>Lower Cost \u2014 Will maximize usage of all compute resources across, GPUs, DPUs and CPUs to lower overall costs, and create a pooled resource environment that can be shared efficiently across teams.<\/li>\n\n\n\n<li>Accelerated Storage \u2014 VMware vSAN Express Storage Architecture will provide performance-optimized NVMe storage and supports GPUDirect\u00ae storage over RDMA, allowing for direct I\/O transfer from storage to GPUs without CPU involvement.<\/li>\n\n\n\n<li>Accelerated Networking \u2014 Deep integration between vSphere and NVIDIA NVSwitch&#x2122; technology will further enable multi-GPU models to execute without inter-GPU bottlenecks.<\/li>\n\n\n\n<li>Rapid Deployment and Time to Value \u2014 vSphere Deep Learning VM images and image repository will enable fast prototyping capabilities by offering a stable turnkey solution image that includes frameworks and performance-optimized libraries pre-installed.<\/li>\n<\/ul>\n\n\n\n<p class=\"wp-block-paragraph\">The platform will feature NVIDIA NeMo, an end-to-end, cloud-native framework included in NVIDIA AI Enterprise \u2014 the operating system of the NVIDIA AI platform \u2014 that allows enterprises to build, customize and deploy generative AI models virtually anywhere. NeMo combines customization frameworks, guardrail toolkits, data curation tools and pretrained models to offer enterprises an easy, cost-effective and fast way to adopt generative AI.<\/p>\n\n\n\n<p class=\"wp-block-paragraph\">For deploying generative AI in production, NeMo uses TensorRT for Large Language Models (TRT-LLM), which accelerates and optimizes inference performance on the latest LLMs on NVIDIA GPUs. With NeMo, VMware Private AI Foundation with NVIDIA will enable enterprises to pull in their own data to build and run custom generative AI models on VMware\u2019s hybrid cloud infrastructure.<\/p>\n\n\n\n<p class=\"wp-block-paragraph\">At VMware Explore 2023, NVIDIA and VMware will highlight how developers within enterprises can use the new NVIDIA AI Workbench to pull community models, like Llama 2, available on Hugging Face, customize them remotely and deploy production-grade generative AI in VMware environments.<\/p>\n\n\n\n<p class=\"wp-block-paragraph\"><strong>Broad Ecosystem Support for VMware Private AI Foundation With NVIDIA<\/strong><\/p>\n\n\n\n<p class=\"wp-block-paragraph\">VMware Private AI Foundation with NVIDIA will be supported by Dell Technologies, Hewlett Packard Enterprise and Lenovo \u2014 which will be among the first to offer systems that supercharge enterprise LLM customization and inference workloads with NVIDIA L40S GPU<a href=\"https:\/\/www.nvidia.com\/en-us\/data-center\/l40\/\">s<\/a>, NVIDIA BlueField-3 DPUs and NVIDIA ConnectX-7 SmartNICs.<\/p>\n\n\n\n<p class=\"wp-block-paragraph\">The NVIDIA L40S GPU enables up to 1.2x more generative AI inference performance and up to 1.7x more training performance compared with the NVIDIA A100 Tensor Core GPU.<\/p>\n\n\n\n<p class=\"wp-block-paragraph\">NVIDIA BlueField-3 DPUs accelerate, offload and isolate the tremendous compute load of virtualization, networking, storage, security and other cloud-native AI services from the GPU or CPU.<\/p>\n\n\n\n<p class=\"wp-block-paragraph\">NVIDIA ConnectX-7 SmartNICs deliver smart, accelerated networking for data center infrastructure to boost some of the world\u2019s most demanding AI workloads.<\/p>\n\n\n\n<p class=\"wp-block-paragraph\">VMware Private AI Foundation with NVIDIA builds on the companies\u2019 decade-long partnership. Their co-engineering work optimized VMware\u2019s cloud infrastructure to run NVIDIA AI Enterprise with performance comparable to bare metal. Mutual customers further benefit from the resource and infrastructure management and flexibility enabled by VMware Cloud Foundation.&nbsp;&nbsp;<\/p>\n\n\n\n<p class=\"wp-block-paragraph\"><strong>Availability<\/strong><\/p>\n\n\n\n<p class=\"wp-block-paragraph\">VMware intends to release VMware Private AI Foundation with NVIDIA in early 2024.<\/p>\n","protected":false},"excerpt":{"rendered":"<p>VMware Private AI Foundation with NVIDIA will enable enterprises to customize models and run generative AI applications, including intelligent chatbots, assistants, search and summarization.<\/p>\n","protected":false},"author":6,"featured_media":58154,"comment_status":"closed","ping_status":"closed","sticky":false,"template":"","format":"standard","meta":{"footnotes":""},"categories":[19],"tags":[411,96,2727,1656,670],"class_list":["post-58148","post","type-post","status-publish","format-standard","has-post-thumbnail","hentry","category-headlines","tag-nvidia","tag-technology","tag-technology-adaption","tag-technology-investment","tag-vmware"],"_links":{"self":[{"href":"http:\/\/www.upgrademag.com\/web\/wp-json\/wp\/v2\/posts\/58148","targetHints":{"allow":["GET"]}}],"collection":[{"href":"http:\/\/www.upgrademag.com\/web\/wp-json\/wp\/v2\/posts"}],"about":[{"href":"http:\/\/www.upgrademag.com\/web\/wp-json\/wp\/v2\/types\/post"}],"author":[{"embeddable":true,"href":"http:\/\/www.upgrademag.com\/web\/wp-json\/wp\/v2\/users\/6"}],"replies":[{"embeddable":true,"href":"http:\/\/www.upgrademag.com\/web\/wp-json\/wp\/v2\/comments?post=58148"}],"version-history":[{"count":0,"href":"http:\/\/www.upgrademag.com\/web\/wp-json\/wp\/v2\/posts\/58148\/revisions"}],"wp:featuredmedia":[{"embeddable":true,"href":"http:\/\/www.upgrademag.com\/web\/wp-json\/wp\/v2\/media\/58154"}],"wp:attachment":[{"href":"http:\/\/www.upgrademag.com\/web\/wp-json\/wp\/v2\/media?parent=58148"}],"wp:term":[{"taxonomy":"category","embeddable":true,"href":"http:\/\/www.upgrademag.com\/web\/wp-json\/wp\/v2\/categories?post=58148"},{"taxonomy":"post_tag","embeddable":true,"href":"http:\/\/www.upgrademag.com\/web\/wp-json\/wp\/v2\/tags?post=58148"}],"curies":[{"name":"wp","href":"https:\/\/api.w.org\/{rel}","templated":true}]}}