
As AI strikes from mannequin improvement to manufacturing inference, compute demand is accelerating and shifting towards constantly working AI factories that generate tokens at scale. This shift requires entry to massive‑scale, multi‑tenant accelerated computing that may come on-line rapidly, keep extremely utilized and assist the economics of token‑scale AI providers.
Rising AI corporations traditionally have had restricted entry to capital-intensive infrastructure, with even long-term commitments inadequate to unlock financing for compute.
To deal with this, NVIDIA is introducing a brand new enterprise mannequin that opens up compute entry to the quick‑rising AI ecosystem of startups, mannequin builders, enterprises, analysis organizations and regional AI gamers.
This new mannequin allows AI clouds to obtain NVIDIA infrastructure for AI-native, enterprise and ISV prospects via financial alignment with a revenue-sharing and credit-support mannequin. By the partnership, AI clouds will promote NVIDIA-powered cloud providers, with NVIDIA incomes each customary product income and a share of the cloud income on the supported capability. This construction accelerates adoption of NVIDIA platforms among the many high-growth, high-conviction AI native sector, and gives NVIDIA with a recurring, usage-linked earnings stream.
For mannequin builders, inference suppliers, agent platforms and enterprises scaling AI, it will possibly imply sooner entry to full-stack accelerated computing with out ready via website choice, energy procurement, development and {hardware} bring-up.
NVIDIA AI Manufacturing facility Capability Constructed Round Demand
The initiative is already taking form, with AI cloud corporations constructing DSX AI factories designed to serve prospects and workloads throughout areas.
Sharon AI and Firmus are among the many first corporations to work with NVIDIA on this new enterprise mannequin.
Sharon AI is deploying as much as 40,000 NVIDIA Grace Blackwell GB300 GPUs.
“This strategic collaboration with NVIDIA marks a pivotal second in Sharon AI’s mission to ship sovereign, large-scale AI compute infrastructure,” mentioned James Manning, cofounder and CEO of Sharon AI.
Firmus is constructing a DSX AI manufacturing unit campus in Batam, Indonesia. The campus is anticipated to scale to 360 megawatts and as much as 170,000 NVIDIA GPUs.
“AI-native corporations want entry to scalable, energy- and cost-efficient compute infrastructure to compete globally,” mentioned Tim Rosenfield, co-CEO of Firmus Applied sciences. “Firmus AI cloud is constructing a NVIDIA DSX-aligned AI manufacturing unit, which can allow our cloud to assist extra prospects entry the compute they should construct and scale AI.”
AI natives akin to Baseten, Fireworks AI and Collectively AI present the place compute demand is headed: they want instant entry to AI cloud capability to run mannequin coaching, post-training, fine-tuning and high-volume agentic inference for builders, digital natives and enterprises constructing with AI.
Their prospects want dependable entry to large-scale NVIDIA accelerated computing as utilization grows, however in addition they want industrial flexibility as merchandise transfer from pilot to manufacturing.
To safe compute capability and construct and deploy AI fashions, contact Sharon AI and Firmus.
Be taught extra about NVIDIA Cloud Companions and AI factories.
