Editorial illustration of private cloud inference secured across device, cloud, and GPU layers
Editorial illustration of private cloud inference secured across device, cloud, and GPU layers
+ NVIDIA AI News

NVIDIA says Apple Private Cloud Compute will use Blackwell GPUs on Google Cloud

NVIDIA says Apple Private Cloud Compute is expanding to Google Cloud with Blackwell GPUs and Confidential Computing for server-side Apple Intelligence inference.

NVIDIA says Apple’s Private Cloud Compute is expanding beyond Apple’s data centers to Google Cloud, with NVIDIA Blackwell GPUs and Confidential Computing supporting server-side inference for Apple Intelligence.

That makes Apple’s privacy architecture more complicated, and more interesting. Private Cloud Compute started as an Apple-controlled answer to a hard product problem: some Apple Intelligence requests need server-scale models, but users still expect Apple-style privacy boundaries. NVIDIA’s June 9 post says the next version of that server path will include Google Cloud infrastructure and Blackwell GPUs with hardware-backed confidential computing.

The provider stack is the news

The notable part is not simply that Apple is using more cloud infrastructure. It is the stack: Apple Foundation Models, technologies behind Google’s Gemini family, Google Cloud, and NVIDIA Blackwell hardware all appear in the same server-side path.

NVIDIA describes the work as support for “next-generation Apple Intelligence features.” The company says Apple and Google are custom-building the relevant models, and that NVIDIA is collaborating with both companies to support the inference layer.

That is The AI Feed’s read from NVIDIA’s post, not an Apple capacity disclosure. NVIDIA does not say how much Apple Intelligence traffic will run this way, which features will use it, or when users will see a visible change. The hard fact is narrower: NVIDIA says its confidential-computing GPUs are part of the expanded Private Cloud Compute architecture on Google Cloud.

Confidential inference is the product requirement

NVIDIA’s post explains the relevant security idea plainly: Confidential Computing protects data while it is being processed by isolating workloads in trusted execution environments and letting systems verify that the infrastructure has not been tampered with before sensitive data is sent.

For AI assistants, that matters because the most useful requests are often the most private. A model asked to reason over messages, documents, photos, calendars, or voice conversations cannot be treated like a normal anonymous web query. If that request leaves the device, the user needs a stronger guarantee than “the provider promises to behave.”

Apple’s answer has been Private Cloud Compute. NVIDIA’s pitch is that Blackwell GPUs can bring accelerated inference into that architecture without dropping the privacy bar. The company’s concrete capabilities list includes hardware-rooted trust, encrypted communication paths, remote attestation, and support for accelerated AI inference and training.

Sources

The AI Feed Desk

The AI Feed Desk

Editorial desk

The AI Feed Desk tracks AI provider updates, model releases, agent tooling, and enterprise adoption, turning fast-moving announcements into source-linked context for builders and operators.

Noticed a typo, incorrect information, or translation error?

Tell us so we can fix it.

Help Improve This Article

Related Articles

Editorial illustration of a local text model generating many tokens in parallel on a GPU

Google releases DiffusionGemma for faster local text generation

Google's DiffusionGemma is an experimental open text-diffusion model that generates blocks of text in parallel for lower-latency local workflows.

The AI Feed Desk

By The AI Feed Desk

A large AI factory rack sends green revenue tokens toward a smaller cloud node

NVIDIA turns AI cloud capacity into a revenue-sharing model

NVIDIA's new AI cloud model pairs revenue sharing with credit support, giving emerging cloud providers a way to finance AI factories while tying NVIDIA to downstream token demand.

The AI Feed Desk

By The AI Feed Desk

NVIDIA Blackwell data center hardware used in the AgentPerf benchmark announcement

NVIDIA says Blackwell leads the first AgentPerf benchmark

NVIDIA says GB300 NVL72 runs up to 20x more agents per megawatt than H200 on AgentPerf, a new benchmark for agentic inference.

The AI Feed Desk

By The AI Feed Desk

A secure data center control plane with server racks, agent nodes, lock symbols, and rollback controls shown as abstract infrastructure

NVIDIA and HPE package agent infrastructure as a private-cloud control plane

HPE AI Factory with NVIDIA is adding Vera CPU, NVIDIA Agent Toolkit, confidential computing, local agent registration, and rollback controls.

The AI Feed Desk

By The AI Feed Desk

An AI server rack connects to a warm closed-loop liquid cooling system and dry cooler

NVIDIA says 45 C liquid cooling can reshape AI factory design

NVIDIA says Rubin-generation AI infrastructure can run with 45 C coolant in closed-loop liquid-cooled AI factories, reducing cooling energy and water dependence.

The AI Feed Desk

By The AI Feed Desk