Token Routing & Inference
Ask SpiderGate for a job instead of a model. 32 task aliases resolve to ranked fallback chains, 91 slots deep in total, reordered by provider health at request time.
Ask for the job, not the model
Stop writing a vendor's model id into your source. Send the name of the work instead and the gateway decides what serves it, so the model behind a call can change without the caller knowing. It is the same indirection as a DNS name in front of an IP, and it buys the same thing.
// Hardcoded: the provider's product decision, in your source
const response = await openai.chat.completions.create({
model: "llama-3.1-8b-instant", // breaks when it is retired
messages: [...]
});
// SpiderGate: name the job, same OpenAI client
const spidergate = new OpenAI({
baseURL: "https://spideriq.ai/api/gate/v1",
apiKey: "<client_id>:<api_key>:<api_secret>"
});
const response = await spidergate.chat.completions.create({
model: "spideriq/coding", // a job, not a model
messages: [...]
});
An alias is a chain, not a model
An alias resolves to a ranked list, not a single model. 32 aliases carry 91 ordered slots between them, an average of 2.8 deep. When the first choice rate limits or errors, the next slot answers, and the response carries an X-SpiderGate-Fallback-From header naming the first choice so you can see it in your own logs.
Health reorders the chain
Models on healthy providers move to the front and struggling ones move to the back. Nothing is dropped, so a provider having a bad ten minutes recovers on its own.
Cross-alias fallback
If a whole chain is exhausted the request falls through to its family's general alias rather than failing outright.
One URL in front of the catalog
The chains route across Cerebras, Groq, Mistral, MiniMax, NVIDIA NIM, OpenRouter and Codex today. Behind them sits a catalog of 1,216 models, 61 of which are key-backed right now. The routing engine is BerriAI's litellm.Router; what SpiderGate adds is the task alias, per-tenant auth and a usage record for every call.
One credential
You hold a single SpiderGate token. The provider keys, the pooling and the rotation stay on our side of the line.
Read the live chains
GET /api/gate/v1/aliases needs no auth and returns every alias, its chain in priority order and its 30-day usage. Hardcode against that, not against a docs page.
Features
-

Capability Scores & Model Ranking
Compare models by capability category, benchmark evidence, and provider availability in the SpiderGate catalog.
-
Contributor keys — collect a client's key without a seat
Ask a client or colleague for the provider key you need. A signed link, no account, no dashboard seat, and the key validated against the provider before it is stored.
-

Key Health & Re-authentication
A failing provider key is skipped after three failures and retired after three auth failures in 24 hours. The contributor who added it gets a re-authentication link, and re-authenticating updates the same key in place.
-

Studio
Keep conversations and generations in projects, save reusable prompts for agents, compare model outputs, and inspect traces in SpiderGate Studio.
-

Subscription Keys & Packages
A credential store holds the key and stops there. SpiderGate splits a provider key into three axes — how it arrived, how it bills, and which package it is on — so a flat-monthly coding plan meters against its own rolling window.
-

The Media API
Generate images, video and speech on the same SpiderGate token you use for chat. 33 of 81 media models are active, each publishing the parameters it accepts so agents never guess.
-

The Vault
Provider keys encrypted at rest, contributed through an invite link so the requester never sees the secret, and pooled with round-robin failover so one dead key does not stop your agents.
-

Token Routing & Inference
Send a task alias instead of a provider model ID. SpiderGate routes requests through health-aware model chains with fallback.
-

Traces & Cost Attribution
Your provider bills per API key. On a shared pool, the question you need answered is per agent. SpiderGate traces every request, attributes spend to the agent that caused it, and shows which model actually answered.
-

Usage that tells the truth
See what your gateway traffic actually delivered, not just what returned HTTP 200. Four outcome states that sum to your total, real calendar ranges, and a key recommendation that is allowed to say no.
-

Your Subscriptions
One endpoint returns your brand's own LLM plans, the concrete model ids you may pin, whether each is usable right now, and how many lease slots are free.