#control-plane

← All posts

Modelplane v0.5: an AI gateway and fleet telemetry

Modelplane v0.5 turns the fleet gateway into an AI gateway and collects every engine's metrics under one vocabulary. Civo joins the clouds it can provision on.

Modelplane v0.4: NVIDIA Dynamo and AI Cluster Runtime

Modelplane v0.4: NVIDIA Dynamo and AI Cluster Runtime

Modelplane v0.4 composes NVIDIA's inference stack across a fleet: a new Dynamo serving stack, and cluster software built from NVIDIA AI Cluster Runtime.

Why Day 0 for Nemotron 3.5 Lightning wasn't a scramble

Why Day 0 for Nemotron 3.5 Lightning wasn't a scramble

NVIDIA released Nemotron-3.5-Lightning this morning. It was running on Modelplane by the afternoon, without a line of new Modelplane code, because day-zero model support is built into the design, not a scramble by the team.

Any Engine, Any Topology, Any Infrastructure: How We Designed Modelplane

Any Engine, Any Topology, Any Infrastructure: How We Designed Modelplane

How we designed Modelplane's fleet-level inference API to fit any engine, in any topology, on any infrastructure — and what's under the hood now that v0.1 has shipped.

Introducing Modelplane: the control plane for AI inference

Introducing Modelplane: the control plane for AI inference

Today we're open sourcing Modelplane, a control plane that operates AI inference across a fleet of GPU clusters, on cloud, neocloud, and on-premise, as one inference platform.