Infrastructure
More model choice. Lower cost.
Each request goes through one gateway to a configured channel, while the model ID, token usage and billed amount stay visible.
Connect to a growing catalog through one compatible endpoint. Use one key and access every live model at 90% off list price.
scroll
Models in the catalog from
One gateway, a growing model catalog
Use one compatible API for every model currently live, with per-key controls, visible usage and pay-as-you-go billing. Add new models without rebuilding your integration.
10
models in the catalog
2
model providers
2
configured API endpoints
1M
tokens max published context

Connect with leading AI clients. Check the model plaza for each model’s supported endpoints and the API docs for integration details.

Set quotas, spend caps and rate limits per key, then review usage and billing in the console.
Developers
Keep your existing client. Change the base URL and model ID, then move between models without rebuilding your stack.
Works with Cherry Studio, CC Switch and apps that support leading AI API protocols.
from openai import OpenAI client = OpenAI(base_url="https://api.thefreebrain.com/v1", api_key="YOUR_API_KEY")reply = client.chat.completions.create(model="gpt-5.6-terra", messages=[{"role": "user", "content": "Say hello from the gateway."}])print(reply.choices[0].message.content)Example model: gpt-5.6-terra. See the model plaza for rates and endpoint availability.
Explore the models listed today, with more to come. Cards show standard-tier rates; see the model plaza for long-context tiers and exact billing.
Explore live models and pricing →
Price snapshot: 2026-09-07. Standard-tier input/output prices per million tokens; long-context and other rates may differ. Confirm current rates in the model plaza.
Infrastructure
Each request goes through one gateway to a configured channel, while the model ID, token usage and billed amount stay visible.
From request to model, usage and bill, every step is recorded in the console so the actual cost never has to be guessed.

Use Discount Offers to lower your AI costs. Every live model is 90% off list price, with each rate published in the model plaza.
The same low-cost gateway can carry you from the first request to a stable production workload.
Ship faster, pay per use.
Control and reliability at scale.
Three steps
01
Create a key in the console, top up balance and set a quota.
02
Point any compatible client at the gateway base URL.
03
Track usage, cost and latency live, per key and request.

One gateway for many models, with simpler integration and lower spend.