One key per app
You’ll always know which app made which call.
With drop in replacement for OpenAI or Anthropic SDK. Pick a model, a hosting location and set your budget.
No guessing and no fine print. You’ll see the hosting location and the rate for each model before you connect anything, so you can choose the right fit for the job and for your compliance team.

Compare what each one can do, where it runs and what it costs.
Give it the name of the app that will use it.
Copy our example, run it, and watch the usage appear straight away.
You’ll always know which app made which call.
Check what each key has spent, then open the individual calls behind it.
Set a budget per key and see how much remains.
In partnership with SCX
We partnered with SCX so selected models run on infrastructure in Australia. SCX runs isolated inference on purpose-built hardware, with IRAP-aligned controls. SCX states that prompts are not cached or used for training, and are never stored, indexed or replayed.
You can read SCX’s security details at scx.ai. Each model in the catalogue shows exactly where it runs, so you always know what you’re choosing.
Point your existing client at Kubox, swap in your key, and you’re done.
from openai import OpenAI
client = OpenAI(
base_url="https://api.kubox.cloud/v1",
api_key="YOUR_KUBOX_API_KEY",
)
response = client.chat.completions.create(
model="gpt-oss-120b",
messages=[{"role": "user", "content": "Summarise this changelog."}],
)
print(response.choices[0].message.content)Some workloads can’t leave your network at all. We can run open-weight models entirely inside your environment. Tell us what you need and we’ll work it through with you.