#control-plane
← All postsModelplane v0.5: An AI gateway, fleet telemetry, and Civo
Modelplane v0.5 turns the fleet gateway into an AI gateway, collects every engine's metrics under one vocabulary, and adds Civo as a cluster source.

Modelplane v0.4: NVIDIA Dynamo and AI Cluster Runtime
Modelplane v0.4 composes NVIDIA's inference stack across a fleet: a new Dynamo serving stack, and cluster software built from NVIDIA AI Cluster Runtime.

Why Day 0 for Nemotron 3.5 Lightning wasn't a scramble
NVIDIA released Nemotron-3.5-Lightning this morning. It was running on Modelplane by the afternoon, without a line of new Modelplane code, because day-zero model support is built into the design, not a scramble by the team.

Any Engine, Any Topology, Any Infrastructure: How We Designed Modelplane
How we designed Modelplane's fleet-level inference API to fit any engine, in any topology, on any infrastructure — and what's under the hood now that v0.1 has shipped.

Introducing Modelplane: the control plane for AI inference
Today we're open sourcing Modelplane, a control plane that operates AI inference across a fleet of GPU clusters, on cloud, neocloud, and on-premise, as one inference platform.


