TLDRocket
Sign in

NVIDIA Unlocks AI Compute at Scale, Inviting Partners to Power the AI Infrastructure Buildout

NVIDIA Colette Kress Covered by 2 sources

NVIDIA introduced a new business model that enables AI cloud providers to access NVIDIA infrastructure through revenue-sharing and credit-support arrangements, allowing startups and enterprises faster access to accelerated computing for AI inference and training. Sharon AI is deploying up to 40,000 NVIDIA Grace Blackwell GB300 GPUs, while Firmus is building an AI factory campus in Indonesia expected to scale to 360 megawatts with up to 170,000 NVIDIA GPUs. The model accelerates adoption of NVIDIA platforms among AI-native companies by removing barriers to large-scale compute access without delays from site selection and infrastructure construction.

Why it matters

As AI moves from model development to production inference, compute demand is accelerating and shifting toward continuously operating AI factories that generate tokens at scale. This shift requires access to large‑scale, multi‑tenant accelerated computing that can come online quickly, stay highly utilized and support the economics of token‑scale AI services. Emerging AI companies historically have […]

Also covered by

Related stories

The daily briefing

Every AI story that matters, in your inbox by 8am.

TLDRocket reads 60+ sources, removes duplicate coverage, and summarises the day in two minutes. Free, no spam, unsubscribe anytime.