Getting Started
- Go to /signup
- Sign in with Google
- Copy your API key (starts with 32 hex characters)
- Start making requests
Authentication
Use Bearer token in the Authorization header. All requests must include a valid API key.
Authorization: Bearer YOUR_API_KEY
Base URL
All API endpoints are served under the following base URL:
https://api.atlasbrain.cloud/v1
Available Models
The following models are available through the Atlas Gateway API.
- Loading models...
Alloy Models
Atlas Alloy is our own frontier model family — an alloy is stronger than any pure metal. You make one standard API call; before the answer reaches you, Alloy has derived it, challenged it, and verified it. How Alloy is built and trained is proprietary — what you get is a single OpenAI-style response that has already survived more scrutiny than ordinary models apply to themselves.
atlas/alloy
- The flagship. Thinking level low / medium / high via reasoning_effort; code sessions automatically get the code-tuned reasoning
- Loading live request cost...
- All tiers
atlas/alloy-pro
- Our deepest reasoning; at high thinking it reasons in extended multi-pass depth and runs a final verification pass before anything ships
- Loading live request cost...
- Pro & Premium only
Thinking level. Both models accept "reasoning_effort": "low" | "medium" | "high" (default medium) on a standard chat completion — low answers fast and cheap, high buys the full extended reasoning. Legacy model ids keep resolving on the wire, so existing integrations never break.
Each alloy call consumes its live catalog cost from your 5h/7d rate windows (bonus requests count too), but only one concurrent slot. Adaptive routes may use less work where supported; the displayed catalog value is the maximum configured charge for the selected preset. Responses are standard OpenAI chat completions with the requested alloy model in the model field.
Measured results. On a 25-task benchmark of objectively verifiable questions (counting, trap comparisons, code-output prediction, date math, probability), atlas/alloy (at low and medium thinking) and atlas/alloy-pro (at high thinking) each scored 25/25 (solo baselines: 24–25/25). In blind, position-swapped head-to-head grading against the strongest solo baseline, atlas/alloy took zero correctness losses across 32 graded pairs. On open-ended prose the honest summary is parity with the best solo model — Alloy's edge is verifiable accuracy and honest, cross-checked confidence. Current numbers, methodology, and the trade-offs we don't hide: the benchmarks page.
Live thinking stream. Streaming alloy requests start showing progress within a second or two: while Alloy works, you may see a reasoning trace in a collapsible “thinking” block (most clients render this natively), then the final answer streams normally. Clients that ignore unknown fields simply see a standard completion. Add "alloy_narrate": false to skip the thinking trace.
Agents & tool use. Alloy models are fully tool-compatible — send tools / tool_choice exactly as you would to any OpenAI-style model, from Claude Code, IDE agents, or your own agent loop. Alloy automatically adapts its effort per turn: simple mechanical turns (like returning a tool result) are handled efficiently at minimal cost, while complex reasoning tasks get the full deep-reasoning treatment. You don't need to manage this — it just works.
Strict client mode. Tool-bearing requests automatically get a byte-for-byte standard OpenAI wire with no extra fields — so strict agentic parsers (Claude Code, IDE agents, SDK tool loops) work seamlessly. Force this on any request with "alloy_strict": true (or the X-Alloy-Strict: 1 header).
cURL
curl https://api.atlasbrain.cloud/v1/chat/completions \ -H "Authorization: Bearer YOUR_API_KEY" \ -H "Content-Type: application/json" \ -d '{ "model": "atlas/alloy", "messages": [{"role": "user", "content": "Explain CRDTs in two paragraphs."}] }'
Python (openai library)
from openai import OpenAI client = OpenAI( base_url="https://api.atlasbrain.cloud/v1", api_key="YOUR_API_KEY" ) response = client.chat.completions.create( model="atlas/alloy", messages=[{"role": "user", "content": "Explain CRDTs in two paragraphs."}] ) print(response.choices[0].message.content)
Alloy Code — the official CLI
Alloy Code (it also answers to Atlas Code — one product, two names) is our first-party terminal coding agent, built for Atlas Gateway. It plans, writes, and verifies code in your repository using our Alloy 2 frontier models — no API keys to paste into configs, no wire setup.
# 1. Install — one command (SHA-256 verified, signed release): # Windows (PowerShell): irm https://atlasbrain.cloud/install.ps1 | iex # macOS / Linux: curl -fsSL https://atlasbrain.cloud/install | sh # 2. Sign in (opens your browser; key lands in your OS keychain) alloy login # 3. Go alloy # interactive terminal UI alloy run "fix the bug" # headless one-shot mission alloy review # AI review of your git diff (read-only) alloy watch "npm test" # re-check on change, --fix auto-repairs alloy playbook # staged mission templates alloy doctor # connectivity + auth self-check
What you get. Missions with live plans and receipts, playbooks that survive restarts, a rewindable session timeline, project skills and agents in .alloy/, and a thinking-level slider from quick answers to deep reasoning. Signed releases — every archive is verifiable against the published release key.
Manual download → · Free tier works out of the box; Pro & Premium unlock alloy-pro deep reasoning.
Preview channel (beta)
Prereleases carry the newest features first — less battle-tested, same signed and hash-verified install path. Install the current preview build with:
# Windows (PowerShell): irm https://atlasbrain.cloud/install-beta.ps1 | iex # macOS / Linux: curl -fsSL https://atlasbrain.cloud/install-beta | sh
When a preview proves stable it graduates to the official release above. Running a preview? Tell us what breaks — that feedback is what gets features promoted.
Use with Coding Agents
Atlas Gateway speaks three wire protocols — OpenAI Chat Completions, the Anthropic Messages API, and the OpenAI Responses API — so every major coding agent connects with a copy-paste config. Tool-bearing requests automatically get strict byte-standard wire output.
Claude Code
# macOS / Linux export ANTHROPIC_BASE_URL="https://api.atlasbrain.cloud" export ANTHROPIC_AUTH_TOKEN="YOUR_API_KEY" export ANTHROPIC_MODEL="atlas/alloy" export API_TIMEOUT_MS=300000 claude # Windows (cmd) set ANTHROPIC_BASE_URL=https://api.atlasbrain.cloud set ANTHROPIC_AUTH_TOKEN=YOUR_API_KEY set ANTHROPIC_MODEL=atlas/alloy set API_TIMEOUT_MS=300000 claude
ANTHROPIC_API_KEY works too; API_TIMEOUT_MS=300000 is recommended — deep-reasoning turns on atlas/alloy-pro can take a couple of minutes. Code sessions on atlas/alloy automatically use the code-tuned reasoning.
Codex CLI (≥0.142)
Add this to ~/.codex/config.toml, then run codex with ATLAS_API_KEY set in your environment:
model = "atlas/alloy" model_provider = "atlas" [model_providers.atlas] name = "Atlas Gateway" base_url = "https://api.atlasbrain.cloud/v1" wire_api = "responses" env_key = "ATLAS_API_KEY" preferred_auth_method = "apikey"
OpenCode
Create opencode.json in your project directory:
{
"$schema": "https://opencode.ai/config.json",
"provider": {
"atlas": {
"npm": "@ai-sdk/openai-compatible",
"name": "Atlas Gateway",
"options": {
"baseURL": "https://api.atlasbrain.cloud/v1",
"apiKey": "YOUR_API_KEY"
},
"models": { "atlas/alloy": { "name": "Alloy 2" } }
}
},
"model": "atlas/atlas/alloy"
}
Continue.dev / Cursor / any OpenAI-compatible tool
Point the OpenAI-compatible provider at https://api.atlasbrain.cloud/v1 with your Atlas key and any model id from /v1/models. Streaming, tool calls, and JSON mode work as standard.
Endpoints
count_tokens. Send your Atlas key in the x-api-key header. Works natively with Claude Code: set ANTHROPIC_BASE_URL=https://api.atlasbrain.cloud, ANTHROPIC_API_KEY to your Atlas key, and ANTHROPIC_MODEL to any Atlas model id.wire_api="responses"). Same Atlas key and models as chat completions; streams end with response.completed.Rate Limits
Rate limits use sliding time windows per plan tier. Bonus requests (earned via ads) are consumed first and bypass window limits.
Free
- 50 req / 5h
- 200 req / 7d
- 20 RPM
- 2 concurrent
Pro
- 500 req / 5h
- 3,000 req / 7d
- 100 RPM
- 8 concurrent
Premium
- 2,500 req / 5h
- 10,000 req / 7d
- 500 RPM
- 24 concurrent
RPM = Requests Per Minute. When you hit a limit, the API returns 429 — free tier: "Limit reached. Watch an ad to continue now, or wait for reset."; paid tiers: "Rate limit reached (concurrent requests or usage window). Please retry in a few seconds."
Bonus & Top-Up Features
Free users can earn bonus requests that bypass rate windows. Bonus priority mode (configurable by admin) controls whether bonus is used before, after, or randomly against the rate window.
- Earn rewards — Loading current bonus values...
- Promo Codes — Redeem codes via
POST /api/topup/promowith{code} - Multi-Key — Create up to 3/5/10 API keys (Free/Pro/Premium) sharing one rate pool
- Referral — Share your link; both users get bonus when the referral watches 3 ads
- Request Log — View your last 20 requests via
GET /api/requests/log - Dashboard Chat — Built-in streaming chat using your own API key
- System Status — Public health at
/system-statusand/api/public/status
Code Examples
cURL
curl https://api.atlasbrain.cloud/v1/chat/completions \ -H "Authorization: Bearer YOUR_API_KEY" \ -H "Content-Type: application/json" \ -d '{ "model": "claude-opus-4.8", "messages": [{"role": "user", "content": "Hello!"}] }'
Python (openai library)
from openai import OpenAI client = OpenAI( base_url="https://api.atlasbrain.cloud/v1", api_key="YOUR_API_KEY" ) response = client.chat.completions.create( model="claude-opus-4.8", messages=[{"role": "user", "content": "Hello!"}] ) print(response.choices[0].message.content)
Python (requests)
import requests resp = requests.post( "https://api.atlasbrain.cloud/v1/chat/completions", headers={"Authorization": "Bearer YOUR_API_KEY"}, json={"model": "claude-opus-4.8", "messages": [{"role": "user", "content": "Hello!"}]} ) print(resp.json()["choices"][0]["message"]["content"])
JavaScript
const response = await fetch( "https://api.atlasbrain.cloud/v1/chat/completions", { method: "POST", headers: { "Authorization": "Bearer YOUR_API_KEY", "Content-Type": "application/json" }, body: JSON.stringify({ model: "claude-opus-4.8", messages: [{role: "user", content: "Hello!"}] }) } ); const data = await response.json(); console.log(data.choices[0].message.content);
Claude Code CLI
# Configure Claude Code to use Atlas Gateway # (ANTHROPIC_BASE_URL is the origin only - no /v1 suffix) claude config set apiKey YOUR_API_KEY
Pass via environment variables:
ANTHROPIC_API_KEY=YOUR_API_KEY \ ANTHROPIC_BASE_URL=https://api.atlasbrain.cloud \ claude
Codex CLI (OpenAI-compatible)
OPENAI_API_KEY=YOUR_API_KEY \ OPENAI_BASE_URL=https://api.atlasbrain.cloud/v1 \ codex
Cursor IDE
In Cursor settings, set the following under the OpenAI-compatible provider:
- OpenAI API Key: YOUR_API_KEY
- OpenAI Base URL: https://api.atlasbrain.cloud/v1
Or use the custom endpoint option with any OpenAI-compatible provider. Select claude-opus-4.8 as your model.
Continue.dev (VS Code extension)
{
"models": [{
"title": "Atlas Gateway",
"provider": "openai",
"model": "claude-opus-4.8",
"apiKey": "YOUR_API_KEY",
"apiBase": "https://api.atlasbrain.cloud/v1"
}]
}
Aider
export OPENAI_API_KEY=YOUR_API_KEY export OPENAI_API_BASE_URL=https://api.atlasbrain.cloud/v1 aider --model claude-opus-4.8
Checking Usage
Query your current usage and rate window stats with a simple GET request:
curl -H "Authorization: Bearer YOUR_API_KEY" \ https://api.atlasbrain.cloud/v1/usage
Response includes rate window usage, bonus request balance, and token counts.
Response format
{
"tier": "free",
"rate_5h_used": 3,
"rate_5h_limit": 50,
"rate_7d_used": 3,
"rate_7d_limit": 200,
"bonus_requests": 15,
"total_tokens": 1520
}
Models Endpoint
GET /v1/models returns all available models and their current operational status. Useful for programmatically discovering which models are available and healthy.
curl -H "Authorization: Bearer YOUR_API_KEY" \ https://api.atlasbrain.cloud/v1/models
Each model entry includes the model ID, owned_by field, and availability status.