News: 1701344705

  ARM Give a man a fire and he's warm for a day, but set fire to him and he's warm for the rest of his life (Terry Pratchett, Jingo)

HPE targets enterprises with Nvidia-powered platform for tuning AI

(2023/11/30)


HPE - like many tech companies - is betting big on AI in the hope customers will splash enteprise cash on training or fine tuning models and other areas of interest, rather than risk falling behind their peers.

At the annual Discover event, this year held on Barcelona, HPE lifted the lid on a second generative AI platform co-engineered with GPU maker Nvidia just weeks after the first, and also confirmed that the HPE Machine Learning Development Environment is now available as a managed service on public clouds, starting with AWS and Google Cloud.

HPE's prosaically named "enterprise computing solution for generative AI" is a pre-configured platform comprising a mix of HPE hardware and software plus Nvidia's GPUs, networking and AI software.

[1]

The pair recently announced a [2]supercomputing system for AI that follows a similar theme, but that is built to train generative AI models. This latest platform is instead a more modest one intended for enterprise customers to tune existing models to their requirements and then operate them for inferencing work.

[3]

[4]

HPE describes the enterprise computing solution for generative AI as a rack-scale architecture, meaning that it pulls together multiple components which effectively fill a datacenter rack, or most of one. The compute in this case is provided by 16 [5]ProLiant DL380a Gen11 servers based on Intel Xeon Scalable processors, which can be configured with up to four Nvidia L40 GPUs.

Networking in this platform is provided by [6]Nvidia's Spectrum-X switches and network adapters based on its BlueField-3 DPU chip.

[7]

This configuration is sized to allow it to fine-tune a 70 billion parameter version of Meta's Llama-2 model, according to HPE.

The software to run on this infrastructure is similar to that included with the supercomputing AI rig, namely the [8]HPE Machine Learning Development Environment , and Nvidia's AI Enterprise suite and the NeMo framework for conversational AI.

HPE reiterated that AI calls for a new compute architecture, one that HPE hopes to supply, of course.

[9]

"AI requires a fundamentally different architecture, because the workload is fundamentally different than their classic transaction processing and web services workloads that have become so dominant in computing over the last couple of decades," said Evan Sparks, HPE's chief product officer for AI.

"We think that the next decade is going to require full stack thinking from hardware all the way up to the software layer, as organisations lean into the deployment of AI applications," he added.

However, for many enterprises, finding their way with AI means taking existing models and experimenting to see if they add value into their business processes.

"Many organizations are not going to be building their own foundational models, they're going to be taking a model that has been developed elsewhere, and they're going to deploy it into their business to transform their business processes," said Neil MacDonald, EVP & general manager of HPE's Compute business.

"One of the challenges is building and deploying infrastructure that enables experimentation and fine tuning and then deployment of these models," he claimed. "We feel like enterprises are either going to become AI powered, or they're going to become obsolete."

What's the damage?

Like the supercomputer for AI, HPE has not yet detailed how much this is going to cost customers, but the enterprise computing solution for generative AI is set to be available to order some time in the first quarter of 2024.

HPE also said its Machine Learning Development Environment is now available as a managed service on public clouds. This is a platform to train up generative AI models, and largely based on technology HPE gained from its purchase of Determined AI in 2021.

"This is a fully managed service, fully managed by HPE, meaning that the end users don't have to worry about managing the infrastructure at all in their cloud or cloud accounts," claimed Sparks. It is available first on "popular platforms" like AWS and Google, he added.

The fully managed model is intended to reduce the complexity of AI/ML model training and thus speed the development process, HPE said.

HPE is also beefing up its Greenlake for File Storage offering to keep up with the data demands of AI workloads.

[10]Server sales down 31% at HPE as enterprises hack spending

[11]Intel drops the deets on UK's Dawn AI supercomputer

[12]Tata Consultancy Services ordered to cough up $210M in code theft trial

[13]HPE starts Hybrid Cloud push, will herd users into GreenLake subs service

"Starting in Q2, we're significantly growing that for customers who want to scale up to you know, somewhere in the realm 250 petabytes of data," said Patrick Osborne, SVP and GM of Cloud and Data Infrastructure.

Also coming is support for Nvidia's Quantum-2 networking to allow customers to plug into InfiniBand fabrics for higher throughput, Osborne said.

HPE said the products will be available via the channel and via its subscription-based Greenlake finance model.

"Not everyone wants to spend a ton of CapEx right up front to be able to jump into this opportunity, so we can provide these through very flexible consumption models for our customers, and provide them flexibility both on the technology side as well as financially on the economic side," said Osborne. ®

Get our [14]Tech Resources



[1] https://pubads.g.doubleclick.net/gampad/jump?co=1&iu=/6978/reg_onprem/front&sz=300x50%7C300x100%7C300x250%7C300x251%7C300x252%7C300x600%7C300x601&tile=2&c=2ZWi-tY4sRQlouh2L3tfcdQAAAEQ&t=ct%3Dns%26unitnum%3D2%26raptor%3Dcondor%26pos%3Dtop%26test%3D0

[2] https://www.theregister.com/2023/11/13/hpe_and_nvidia_offer_turnkey/j

[3] https://pubads.g.doubleclick.net/gampad/jump?co=1&iu=/6978/reg_onprem/front&sz=300x50%7C300x100%7C300x250%7C300x251%7C300x252%7C300x600%7C300x601&tile=4&c=44ZWi-tY4sRQlouh2L3tfcdQAAAEQ&t=ct%3Dns%26unitnum%3D4%26raptor%3Dfalcon%26pos%3Dmid%26test%3D0

[4] https://pubads.g.doubleclick.net/gampad/jump?co=1&iu=/6978/reg_onprem/front&sz=300x50%7C300x100%7C300x250%7C300x251%7C300x252%7C300x600%7C300x601&tile=3&c=33ZWi-tY4sRQlouh2L3tfcdQAAAEQ&t=ct%3Dns%26unitnum%3D3%26raptor%3Deagle%26pos%3Dmid%26test%3D0

[5] https://buy.hpe.com/us/en/compute/rack-servers/proliant-dl300-servers/proliant-dl380-server/hpe-proliant-dl380a-gen11/p/1014696168

[6] https://www.theregister.com/2023/11/21/nvidia_supernic_nic/

[7] https://pubads.g.doubleclick.net/gampad/jump?co=1&iu=/6978/reg_onprem/front&sz=300x50%7C300x100%7C300x250%7C300x251%7C300x252%7C300x600%7C300x601&tile=4&c=44ZWi-tY4sRQlouh2L3tfcdQAAAEQ&t=ct%3Dns%26unitnum%3D4%26raptor%3Dfalcon%26pos%3Dmid%26test%3D0

[8] https://www.theregister.com/2022/04/27/hpe_new_ai_products/

[9] https://pubads.g.doubleclick.net/gampad/jump?co=1&iu=/6978/reg_onprem/front&sz=300x50%7C300x100%7C300x250%7C300x251%7C300x252%7C300x600%7C300x601&tile=3&c=33ZWi-tY4sRQlouh2L3tfcdQAAAEQ&t=ct%3Dns%26unitnum%3D3%26raptor%3Deagle%26pos%3Dmid%26test%3D0

[10] https://www.theregister.com/2023/11/29/hpe_fiscal_2023/

[11] https://www.theregister.com/2023/11/13/intel_dawn_ai_uk/

[12] https://www.theregister.com/2023/11/24/tata_210m_code_theft/

[13] https://www.theregister.com/2023/10/25/hpe_hybrid_cloud/

[14] https://whitepapers.theregister.com/



Really?

Mike 137

" enterprises are either going to become AI powered, or they're going to become obsolete "

I remember a Dilbert cartoon from (if I remember right) the late 1980s. Its caption was "there's an engineering solution of every problem". Sadly there isn't, despite what the technocrats think (or at least would have us think). In my domain (information risk) it's widely and correctly recognised that the human equation predominates -- and not just "user behaviour" but management behaviour most importantly. That source of error will remain unless it's controlled by corporate culture before AI is adopted. The big problem currently is that corporate culture is generally antithetical to good practice. That's the high hurdle we need to leap before handing over to AI, unless we want the AI to replicate our mistakes (if for no other reason than that the source of its training will otherwise be poor practice).

Artificial Intelligence vs Natural Stupor

fg_swe

"If you do not buy into the latest hype, bad things will happen to you".

Yours Once-A-Great-Company

Human Brain

fg_swe

100E9 Neurons

Each Neuron 10000 Synapses

That makes 100E13 "Parameters". That is 1E15 == 100 000 Billion Parameters. About 1000x more than the model they mention.

In other words, the "AI" is on the level of a worm or the like.

Free your best employees of the Hyped Nonsense*, use Common Sense and you will win. Also, do not allow newspapers and TVs on campus.

Treat your people well, respect them, nurture their innovative skills. Human brains need 1000x less power than grafic cards for the same number of neurons.

* e.g. "COVID", "UFOs", "MBA leader", "man and woman are socially defined", "eternal guilt from the 18th century", "windmills", "solar cells", "corium melting to middle of earth", ...

QOTD:
"Say, you look pretty athletic. What say we put a pair of tennis
shoes on you and run you into the wall?"