Getting Started

  1. Go to /signup
  2. Sign in with Google
  3. Copy your API key (starts with 32 hex characters)
  4. Start making requests

Authentication

Use Bearer token in the Authorization header. All requests must include a valid API key.

Authorization: Bearer YOUR_API_KEY

Base URL

All API endpoints are served under the following base URL:

https://api.atlasbrain.cloud/v1

Available Models

The following models are available through the Atlas Gateway API.

  • Loading models...

Alloy Models

Atlas Alloy is our own frontier model family — an alloy is stronger than any pure metal. You make one standard API call; before the answer reaches you, Alloy has derived it, challenged it, and verified it. How Alloy is built and trained is proprietary — what you get is a single OpenAI-style response that has already survived more scrutiny than ordinary models apply to themselves.

atlas/alloy

  • The flagship. Thinking level low / medium / high via reasoning_effort; code sessions automatically get the code-tuned reasoning
  • Loading live request cost...
  • All tiers

atlas/alloy-pro

  • Our deepest reasoning; at high thinking it reasons in extended multi-pass depth and runs a final verification pass before anything ships
  • Loading live request cost...
  • Pro & Premium only

Thinking level. Both models accept "reasoning_effort": "low" | "medium" | "high" (default medium) on a standard chat completion — low answers fast and cheap, high buys the full extended reasoning. Legacy model ids keep resolving on the wire, so existing integrations never break.

Each alloy call consumes its live catalog cost from your 5h/7d rate windows (bonus requests count too), but only one concurrent slot. Adaptive routes may use less work where supported; the displayed catalog value is the maximum configured charge for the selected preset. Responses are standard OpenAI chat completions with the requested alloy model in the model field.

Measured results. On a 25-task benchmark of objectively verifiable questions (counting, trap comparisons, code-output prediction, date math, probability), atlas/alloy (at low and medium thinking) and atlas/alloy-pro (at high thinking) each scored 25/25 (solo baselines: 24–25/25). In blind, position-swapped head-to-head grading against the strongest solo baseline, atlas/alloy took zero correctness losses across 32 graded pairs. On open-ended prose the honest summary is parity with the best solo model — Alloy's edge is verifiable accuracy and honest, cross-checked confidence. Current numbers, methodology, and the trade-offs we don't hide: the benchmarks page.

Live thinking stream. Streaming alloy requests start showing progress within a second or two: while Alloy works, you may see a reasoning trace in a collapsible “thinking” block (most clients render this natively), then the final answer streams normally. Clients that ignore unknown fields simply see a standard completion. Add "alloy_narrate": false to skip the thinking trace.

Agents & tool use. Alloy models are fully tool-compatible — send tools / tool_choice exactly as you would to any OpenAI-style model, from Claude Code, IDE agents, or your own agent loop. Alloy automatically adapts its effort per turn: simple mechanical turns (like returning a tool result) are handled efficiently at minimal cost, while complex reasoning tasks get the full deep-reasoning treatment. You don't need to manage this — it just works.

Strict client mode. Tool-bearing requests automatically get a byte-for-byte standard OpenAI wire with no extra fields — so strict agentic parsers (Claude Code, IDE agents, SDK tool loops) work seamlessly. Force this on any request with "alloy_strict": true (or the X-Alloy-Strict: 1 header).

cURL

curl https://api.atlasbrain.cloud/v1/chat/completions \
  -H "Authorization: Bearer YOUR_API_KEY" \
  -H "Content-Type: application/json" \
  -d '{
    "model": "atlas/alloy",
    "messages": [{"role": "user", "content": "Explain CRDTs in two paragraphs."}]
  }'

Python (openai library)

from openai import OpenAI

client = OpenAI(
    base_url="https://api.atlasbrain.cloud/v1",
    api_key="YOUR_API_KEY"
)

response = client.chat.completions.create(
    model="atlas/alloy",
    messages=[{"role": "user", "content": "Explain CRDTs in two paragraphs."}]
)
print(response.choices[0].message.content)

Alloy Code — the official CLI

Alloy Code (it also answers to Atlas Code — one product, two names) is our first-party terminal coding agent, built for Atlas Gateway. It plans, writes, and verifies code in your repository using our Alloy 2 frontier models — no API keys to paste into configs, no wire setup.

# 1. Install — one command (SHA-256 verified, signed release):
#    Windows (PowerShell):
irm https://atlasbrain.cloud/install.ps1 | iex
#    macOS / Linux:
curl -fsSL https://atlasbrain.cloud/install | sh

# 2. Sign in (opens your browser; key lands in your OS keychain)
alloy login

# 3. Go
alloy                       # interactive terminal UI
alloy run "fix the bug"     # headless one-shot mission
alloy review                # AI review of your git diff (read-only)
alloy watch "npm test"      # re-check on change, --fix auto-repairs
alloy playbook              # staged mission templates
alloy doctor                # connectivity + auth self-check

What you get. Missions with live plans and receipts, playbooks that survive restarts, a rewindable session timeline, project skills and agents in .alloy/, and a thinking-level slider from quick answers to deep reasoning. Signed releases — every archive is verifiable against the published release key.

Manual download →  ·  Free tier works out of the box; Pro & Premium unlock alloy-pro deep reasoning.

Preview channel (beta)

Prereleases carry the newest features first — less battle-tested, same signed and hash-verified install path. Install the current preview build with:

# Windows (PowerShell):
irm https://atlasbrain.cloud/install-beta.ps1 | iex
# macOS / Linux:
curl -fsSL https://atlasbrain.cloud/install-beta | sh

When a preview proves stable it graduates to the official release above. Running a preview? Tell us what breaks — that feedback is what gets features promoted.

Use with Coding Agents

Atlas Gateway speaks three wire protocols — OpenAI Chat Completions, the Anthropic Messages API, and the OpenAI Responses API — so every major coding agent connects with a copy-paste config. Tool-bearing requests automatically get strict byte-standard wire output.

Claude Code

# macOS / Linux
export ANTHROPIC_BASE_URL="https://api.atlasbrain.cloud"
export ANTHROPIC_AUTH_TOKEN="YOUR_API_KEY"
export ANTHROPIC_MODEL="atlas/alloy"
export API_TIMEOUT_MS=300000
claude

# Windows (cmd)
set ANTHROPIC_BASE_URL=https://api.atlasbrain.cloud
set ANTHROPIC_AUTH_TOKEN=YOUR_API_KEY
set ANTHROPIC_MODEL=atlas/alloy
set API_TIMEOUT_MS=300000
claude

ANTHROPIC_API_KEY works too; API_TIMEOUT_MS=300000 is recommended — deep-reasoning turns on atlas/alloy-pro can take a couple of minutes. Code sessions on atlas/alloy automatically use the code-tuned reasoning.

Codex CLI (≥0.142)

Add this to ~/.codex/config.toml, then run codex with ATLAS_API_KEY set in your environment:

model = "atlas/alloy"
model_provider = "atlas"

[model_providers.atlas]
name = "Atlas Gateway"
base_url = "https://api.atlasbrain.cloud/v1"
wire_api = "responses"
env_key = "ATLAS_API_KEY"
preferred_auth_method = "apikey"

OpenCode

Create opencode.json in your project directory:

{
  "$schema": "https://opencode.ai/config.json",
  "provider": {
    "atlas": {
      "npm": "@ai-sdk/openai-compatible",
      "name": "Atlas Gateway",
      "options": {
        "baseURL": "https://api.atlasbrain.cloud/v1",
        "apiKey": "YOUR_API_KEY"
      },
      "models": { "atlas/alloy": { "name": "Alloy 2" } }
    }
  },
  "model": "atlas/atlas/alloy"
}

Continue.dev / Cursor / any OpenAI-compatible tool

Point the OpenAI-compatible provider at https://api.atlasbrain.cloud/v1 with your Atlas key and any model id from /v1/models. Streaming, tool calls, and JSON mode work as standard.

Endpoints

POST /v1/chat/completions
OpenAI-compatible chat completions. Same request/response format as the OpenAI API. Supports all models listed above.
POST /v1/messages
Anthropic-compatible Messages API. Same request/response format as the Anthropic API, including streaming events, tool use, and count_tokens. Send your Atlas key in the x-api-key header. Works natively with Claude Code: set ANTHROPIC_BASE_URL=https://api.atlasbrain.cloud, ANTHROPIC_API_KEY to your Atlas key, and ANTHROPIC_MODEL to any Atlas model id.
POST /v1/responses
OpenAI Responses API wire. Required by Codex CLI ≥0.142 (wire_api="responses"). Same Atlas key and models as chat completions; streams end with response.completed.
POST /v1/completions
Legacy completions endpoint. OpenAI-compatible format.
GET /v1/models
List all available models and their current status.
GET /v1/usage
Check your current usage, rate window stats, and bonus request balance.

Rate Limits

Rate limits use sliding time windows per plan tier. Bonus requests (earned via ads) are consumed first and bypass window limits.

Free

  • 50 req / 5h
  • 200 req / 7d
  • 20 RPM
  • 2 concurrent

Pro

  • 500 req / 5h
  • 3,000 req / 7d
  • 100 RPM
  • 8 concurrent

Premium

  • 2,500 req / 5h
  • 10,000 req / 7d
  • 500 RPM
  • 24 concurrent

RPM = Requests Per Minute. When you hit a limit, the API returns 429 — free tier: "Limit reached. Watch an ad to continue now, or wait for reset."; paid tiers: "Rate limit reached (concurrent requests or usage window). Please retry in a few seconds."

Bonus & Top-Up Features

Free users can earn bonus requests that bypass rate windows. Bonus priority mode (configurable by admin) controls whether bonus is used before, after, or randomly against the rate window.

  • Earn rewards — Loading current bonus values...
  • Promo Codes — Redeem codes via POST /api/topup/promo with {code}
  • Multi-Key — Create up to 3/5/10 API keys (Free/Pro/Premium) sharing one rate pool
  • Referral — Share your link; both users get bonus when the referral watches 3 ads
  • Request Log — View your last 20 requests via GET /api/requests/log
  • Dashboard Chat — Built-in streaming chat using your own API key
  • System Status — Public health at /system-status and /api/public/status

Code Examples

cURL

curl https://api.atlasbrain.cloud/v1/chat/completions \
  -H "Authorization: Bearer YOUR_API_KEY" \
  -H "Content-Type: application/json" \
  -d '{
    "model": "claude-opus-4.8",
    "messages": [{"role": "user", "content": "Hello!"}]
  }'

Python (openai library)

from openai import OpenAI

client = OpenAI(
    base_url="https://api.atlasbrain.cloud/v1",
    api_key="YOUR_API_KEY"
)

response = client.chat.completions.create(
    model="claude-opus-4.8",
    messages=[{"role": "user", "content": "Hello!"}]
)

print(response.choices[0].message.content)

Python (requests)

import requests

resp = requests.post(
    "https://api.atlasbrain.cloud/v1/chat/completions",
    headers={"Authorization": "Bearer YOUR_API_KEY"},
    json={"model": "claude-opus-4.8", "messages": [{"role": "user", "content": "Hello!"}]}
)

print(resp.json()["choices"][0]["message"]["content"])

JavaScript

const response = await fetch(
    "https://api.atlasbrain.cloud/v1/chat/completions",
    {
        method: "POST",
        headers: {
            "Authorization": "Bearer YOUR_API_KEY",
            "Content-Type": "application/json"
        },
        body: JSON.stringify({
            model: "claude-opus-4.8",
            messages: [{role: "user", content: "Hello!"}]
        })
    }
);

const data = await response.json();
console.log(data.choices[0].message.content);

Claude Code CLI

# Configure Claude Code to use Atlas Gateway
# (ANTHROPIC_BASE_URL is the origin only - no /v1 suffix)
claude config set apiKey YOUR_API_KEY

Pass via environment variables:

ANTHROPIC_API_KEY=YOUR_API_KEY \
ANTHROPIC_BASE_URL=https://api.atlasbrain.cloud \
claude

Codex CLI (OpenAI-compatible)

OPENAI_API_KEY=YOUR_API_KEY \
OPENAI_BASE_URL=https://api.atlasbrain.cloud/v1 \
codex

Cursor IDE

In Cursor settings, set the following under the OpenAI-compatible provider:

  • OpenAI API Key: YOUR_API_KEY
  • OpenAI Base URL: https://api.atlasbrain.cloud/v1

Or use the custom endpoint option with any OpenAI-compatible provider. Select claude-opus-4.8 as your model.

Continue.dev (VS Code extension)

{
    "models": [{
        "title": "Atlas Gateway",
        "provider": "openai",
        "model": "claude-opus-4.8",
        "apiKey": "YOUR_API_KEY",
        "apiBase": "https://api.atlasbrain.cloud/v1"
    }]
}

Aider

export OPENAI_API_KEY=YOUR_API_KEY
export OPENAI_API_BASE_URL=https://api.atlasbrain.cloud/v1
aider --model claude-opus-4.8

Checking Usage

Query your current usage and rate window stats with a simple GET request:

curl -H "Authorization: Bearer YOUR_API_KEY" \
    https://api.atlasbrain.cloud/v1/usage

Response includes rate window usage, bonus request balance, and token counts.

Response format

{
    "tier": "free",
    "rate_5h_used": 3,
    "rate_5h_limit": 50,
    "rate_7d_used": 3,
    "rate_7d_limit": 200,
    "bonus_requests": 15,
    "total_tokens": 1520
}

Models Endpoint

GET /v1/models returns all available models and their current operational status. Useful for programmatically discovering which models are available and healthy.

curl -H "Authorization: Bearer YOUR_API_KEY" \
    https://api.atlasbrain.cloud/v1/models

Each model entry includes the model ID, owned_by field, and availability status.