brain
OpenAI-compatible API

Use the crowd.

One endpoint. The router places every request on the best target it can find: the browser network when it can serve the model, cloud fallback when it can't. Your existing OpenAI SDK works unchanged.

Router · live placement SIM482 req/s
QWEN 32B
brain/qwen
4,821 nodes
—
DEEPSEEK DISTILL
brain/code
2,184 nodes
—
EMBEDDINGS
brain/embed
1,821 nodes
—
VISIONBETA
brain/vision
884 nodes
—
VERIFICATION
verify/*
canaries
—
brain/autoROUTER

Automatically chooses the cheapest execution path that satisfies the request.

Context32,768 tok
Network price / 1Munpriced
Reference / 1Munpriced
brain/qwenBETA

Distributed Qwen-class inference across the browser pool.

Context32,768 tok
Network price / 1Munpriced
Reference / 1Munpriced
brain/codeBETA

Coding-optimized inference on distilled reasoning models.

Context16,384 tok
Network price / 1Munpriced
Reference / 1Munpriced
brain/embedBETA

Distributed embeddings. Small model, highly parallel, browser-native.

Context8,192 tok
Network price / 1Munpriced
Reference / 1Munpriced

EST Prices are published after cost benchmarks against reference providers. No savings are claimed until they are measured.

Playground

Requests go through the same server-side gateway as the public API. Provider keys never leave the server. The execution target and latency are real; the shard plan shows how the browser pool partitions the work and is simulated until distributed LLM execution ships.

Request
POST /v1/chat/completions · 449 chars
Execution
REQUEST
SPLIT
NODE 8A21
idle
NODE 19F2
idle
NODE 81CC
idle
NODE 28AA
idle
MERGE
RESPONSE
Nodes used SIM
—
Latency LIVE
—
Compute units SIM
—
Tokens LIVE
—
Run the prompt to see how the router places it.

Drop-in for your OpenAI client.

Point the base URL at Brain, choose brain/auto, and keep the rest of your code.

js
import OpenAI from "openai";

const brain = new OpenAI({
  baseURL: "https://YOUR_BRAIN_HOST/v1",
  apiKey: process.env.BRAIN_API_KEY,
});

const res = await brain.chat.completions.create({
  model: "brain/auto",
  messages: [{ role: "user", content: "Audit this contract…" }],
});