NVIDIA AI News

A next-generation AI compute hall shows accelerator modules scaling into a larger Vera Rubin-style server corridor

NVIDIA gives Safe Superintelligence a 10x compute path

NVIDIA and Safe Superintelligence announced a long-term partnership that includes NVIDIA investment and access to Vera Rubin systems.

The AI Feed Desk

By The AI Feed Desk

A surgical robot simulator streams generated tabletop suturing frames from robot actions inside a controlled research lab

NVIDIA Cosmos-H-Dreams turns surgical world models into a real-time simulator

NVIDIA published Cosmos-H-Dreams, a real-time action-conditioned generative simulator for surgical robotics built from Cosmos-H-Surgical-Simulator.

The AI Feed Desk

By The AI Feed Desk

Open model weights and cybersecurity tools sit between two policy paths labeled access, testing, chips, and safeguards

NVIDIA and Anthropic split over open-weight AI safety

NVIDIA launched the Open Secure AI Alliance while Anthropic argued against blanket open-weight bans and for targeted AI safety controls.

The AI Feed Desk

By The AI Feed Desk

A compact accelerator rack sits inside circular post-training loops on an AI factory floor

NVIDIA frames Vera Rubin around agentic post-training economics

NVIDIA says Vera Rubin is built for continuous post-training loops and can train the largest models with one-fourth the GPUs of Blackwell.

The AI Feed Desk

By The AI Feed Desk

A Japan manufacturing floor connects robotics arms, compact AI PCs, and data-center compute into one NVIDIA stack

NVIDIA uses Japan to package physical AI as a full-stack ecosystem

NVIDIA's July 15 Japan ecosystem update ties RTX Spark, robotics, manufacturing, and local partners into a physical AI stack.

The AI Feed Desk

By The AI Feed Desk

An open dataset map clusters agent workflow samples beside transparent synthetic persona cards

NVIDIA and Hugging Face publish open data for agents

NVIDIA and Hugging Face published a Nemotron data package for agents, including open pretraining data, post-training samples, an interactive Prompt Atlas, and synthetic persona datasets.

The AI Feed Desk

By The AI Feed Desk

A compact AI factory module combines a chip wafer, server racks, cooling pipes, and power equipment

NVIDIA frames U.S. AI buildout as a 43-state supply chain

NVIDIA says its U.S. partner network spans 43 states and plans up to $500 billion of American-built AI infrastructure with semiconductor, system, power, and cloud partners.

The AI Feed Desk

By The AI Feed Desk

A large AI factory rack sends green revenue tokens toward a smaller cloud node

NVIDIA turns AI cloud capacity into a revenue-sharing model

NVIDIA's new AI cloud model pairs revenue sharing with credit support, giving emerging cloud providers a way to finance AI factories while tying NVIDIA to downstream token demand.

The AI Feed Desk

By The AI Feed Desk

A governed cloud workspace connects an AI model core to a high-performance compute rack

Claude reaches Microsoft Foundry with Azure governance and GB300 compute

Anthropic made Claude generally available in Microsoft Foundry, while NVIDIA framed the Azure deployment as a GB300 Blackwell Ultra agent platform.

The AI Feed Desk

By The AI Feed Desk

A green open-model core runs inside a sealed air-gapped compute chamber

NVIDIA and Palantir put Nemotron open models inside air-gapped government AI

NVIDIA says Palantir is using Nemotron open models in isolated environments for U.S. government agencies and critical infrastructure operators.

The AI Feed Desk

By The AI Feed Desk

A mixture-of-experts model is split across GPUs while a single import path feeds the training pipeline

NVIDIA NeMo AutoModel makes MoE fine-tuning a one-import upgrade

NVIDIA's Hugging Face article shows NeMo AutoModel wrapping expert parallelism and custom kernels behind the familiar Transformers loading path for MoE fine-tuning.

The AI Feed Desk

By The AI Feed Desk

An AI server rack connects to a warm closed-loop liquid cooling system and dry cooler

NVIDIA says 45 C liquid cooling can reshape AI factory design

NVIDIA says Rubin-generation AI infrastructure can run with 45 C coolant in closed-loop liquid-cooled AI factories, reducing cooling energy and water dependence.

The AI Feed Desk

By The AI Feed Desk

A secure data center control plane with server racks, agent nodes, lock symbols, and rollback controls shown as abstract infrastructure

NVIDIA and HPE package agent infrastructure as a private-cloud control plane

HPE AI Factory with NVIDIA is adding Vera CPU, NVIDIA Agent Toolkit, confidential computing, local agent registration, and rollback controls.

The AI Feed Desk

By The AI Feed Desk

Rack-scale AI training systems connected by green workload paths in a data center aisle

NVIDIA says Blackwell swept MLPerf Training 6.0

NVIDIA says Blackwell delivered the fastest time to train on all seven MLPerf Training 6.0 benchmarks, including large-scale DeepSeek-V3 and Llama workloads.

The AI Feed Desk

By The AI Feed Desk

NVIDIA Blackwell data center hardware used in the AgentPerf benchmark announcement

NVIDIA says Blackwell leads the first AgentPerf benchmark

NVIDIA says GB300 NVL72 runs up to 20x more agents per megawatt than H200 on AgentPerf, a new benchmark for agentic inference.

The AI Feed Desk

By The AI Feed Desk

Editorial illustration of a local text model generating many tokens in parallel on a GPU

Google releases DiffusionGemma for faster local text generation

Google's DiffusionGemma is an experimental open text-diffusion model that generates blocks of text in parallel for lower-latency local workflows.

The AI Feed Desk

By The AI Feed Desk

Editorial illustration of private cloud inference secured across device, cloud, and GPU layers

NVIDIA says Apple Private Cloud Compute will use Blackwell GPUs on Google Cloud

NVIDIA says Apple Private Cloud Compute is expanding to Google Cloud with Blackwell GPUs and Confidential Computing for server-side Apple Intelligence inference.

The AI Feed Desk

By The AI Feed Desk