#control-plane
← All posts
Modelplane v0.4: NVIDIA Dynamo and AI Cluster Runtime
Modelplane v0.4 can serve models with NVIDIA Dynamo for gang scheduling and fast P2P weight transfer, and can build the software each cluster runs at hardware-validated versions from NVIDIA AI Cluster Runtime. Cloud providers start on demand.

Why Day 0 for Nemotron 3.5 Lightning wasn't a scramble
NVIDIA released Nemotron-3.5-Lightning this morning. It was running on Modelplane by the afternoon, without a line of new Modelplane code, because day-zero model support is built into the design, not a scramble by the team.

Any Engine, Any Topology, Any Infrastructure: How We Designed Modelplane
How we designed Modelplane's fleet-level inference API to fit any engine, in any topology, on any infrastructure — and what's under the hood now that v0.1 has shipped.

Introducing Modelplane: the control plane for AI inference
Today we're open sourcing Modelplane, a control plane that operates AI inference across a fleet of GPU clusters, on cloud, neocloud, and on-premise, as one inference platform.


