Getting started
Point your SDK at the gateway, use a gateway key, and you are done.
1. Get a gateway key
After your organisation is set up, an administrator signs in at Customer login with their email and password and creates a key for your app, choosing its allowed models, budget and rate limit. Request access if you do not have an account yet.
Treat the key like a password. Store it in an environment variable or secrets manager, never in source control.
2. Use the base URL
https://gateway.tokard.com.au/v1
Endpoints
| Method | Path | Description |
|---|---|---|
POST | /v1/chat/completions | Chat completions, including streaming. OpenAI-compatible. |
POST | /v1/responses | Responses API, OpenAI-compatible. |
POST | /v1/embeddings | Text embeddings. |
POST | /v1/messages | Anthropic-compatible Messages API. |
POST | /v1/images/generations | Image generation. |
POST | /v1/audio/speech | Text-to-speech. |
POST | /v1/audio/transcriptions | Speech-to-text transcription. The returned transcript is scanned; the audio itself reaches the provider unredacted. |
Authentication
Send your gateway key as a bearer token: Authorization: Bearer <gateway key>. Provider keys are never required by your application.
Examples
curl
curl https://gateway.tokard.com.au/v1/chat/completions \
-H "Authorization: Bearer $GATEWAY_KEY" \
-H "Content-Type: application/json" \
-d '{
"model": "gpt-4o",
"messages": [{"role": "user", "content": "Hello from the gateway"}]
}' Python
from openai import OpenAI
client = OpenAI(
base_url="https://gateway.tokard.com.au/v1",
api_key=os.environ["GATEWAY_KEY"],
)
resp = client.chat.completions.create(
model="gpt-4o",
messages=[{"role": "user", "content": "Hello from the gateway"}],
)
print(resp.choices[0].message.content) Node
import OpenAI from "openai";
const client = new OpenAI({
baseURL: "https://gateway.tokard.com.au/v1",
apiKey: process.env.GATEWAY_KEY,
});
const resp = await client.chat.completions.create({
model: "gpt-4o",
messages: [{ role: "user", content: "Hello from the gateway" }],
});
console.log(resp.choices[0].message.content); Anthropic SDK (Python)
import anthropic
client = anthropic.Anthropic(
base_url="https://gateway.tokard.com.au",
api_key=os.environ["GATEWAY_KEY"],
)
msg = client.messages.create(
model="claude-sonnet", max_tokens=512,
messages=[{"role": "user", "content": "Hello"}],
) Model names are the ones your administrator has enabled for your key. List them from the console.
When a policy applies
If a request is blocked by policy, the gateway returns an error response explaining which rule applied. Redacted and tokenised requests succeed normally. Tokenised values are restored in the response.
Need help?
Contact us and we will help you connect your first application.
Keep every AI interaction within bounds
Request access and we will help you set up your first policy, key and budget.