One runtime for every model your products call

Astro is a multi-tenant LLM gateway: normalized access to 48 chat models across 17 open inference hosts and the frontier vendors, credential custody that keeps a vendor key out of every calling process, and per-turn metering that reconciles against the vendor's own invoice.

The problem it was built from

Around seventy projects, twenty with real AI needs, and every one of them built its own model-access layer. Counted from source, not asserted:

Every one of those is a place a vendor key can leak and a turn can go unbilled.

What it is

One HTTP service speaking the wires your code already speaks — Anthropic's and OpenAI's — plus a native plane for the things neither expresses. Point a base URL at it and nothing else changes.

What it deliberately refuses to own

A gateway that grows into a platform stops being adoptable. Each of these is refused with a reason rather than deferred.

It also refuses to design around a consumer that should not adopt. One project in the survey keeps the router it already has; that is the correct outcome, not a gap.

Reading further