Package

@flashyos/llm-gateway

Provider-agnostic inference seam: adapters, automatic failover, metering ceilings, and an empty-success guard.

version 0.2.1 · audit of 2026-10-04 · source: flashyos/packages/llm-gateway

workspace package (publication pending)

The seam every estate product routes inference through, so a provider change is a config edit and usage is metered where it is spent.

Learn to build with it: Flashy Academy’s Inference Economy track teaches what inference and compute are, why choosing inference is an economic decision, inference as yield, and getting set up and building with Gatewayz — the unified inference gateway this seam routes to via its gatewayz provider. https://flashy.academy/academy/curriculum

Edge cases — each one paid for once

An empty 200 is a failure wearing a success code

The empty-success guard exists because a provider returning a well-formed nothing passes every status check and fails the user. Guard on content, not on status.

An unpriced model is refused, not passed through

Metering that silently skips a model it cannot price reports a ceiling nobody is enforcing.

Tool use and prompt caching are not yet verified through a gateway

A gateway preserves the common request shape but not every advanced feature: one that silently drops a cache-control directive does not error, it multiplies the bill, and a mangled tool-use request breaks function-calling in ways that read as a model failure. Verify both behave through the gateway before depending on them in production.

← Full catalog · The doctrine behind the tools · Adopt one