82trust / 100

Chat Completions

by Chat completions - auto tier in AI models & inference

x402 APIPassing, checked 4 h ago

OpenAI-compatible chat completions with the model chosen server-side: omit model and the gateway routes the prompt to the top-ranked model for its task (code, reasoning, long-context, general) from a fixed eval-derived ranking, failing over automatically on provider errors. Flat price per call, 16k chars in, 1024 tokens out, streaming supported. Use it as a drop-in OpenAI base_url when you want good answers without picking a model.

POST https://agent402.tools/v1/auto/chat/completions

Last 30 days

All checks passedSome failedAll failedNot checked
Uptime
100%
Response time
284 ms typical, 284 ms slowest 5%
Last check
4 h ago
Next check
any minute now

How to call it

# See the payment challenge (nothing is charged)
curl -i -X POST "https://agent402.tools/v1/auto/chat/completions" \
  -H "content-type: application/json" \
  -d '{"max_tokens":5,"messages":[{"content":"Reply with exactly: OK","role":"user"}]}'
import { wrapFetchWithPayment } from "@x402/fetch";
import { x402Client } from "@x402/core/client";
import { ExactEvmScheme } from "@x402/evm/exact/client";
import { privateKeyToAccount } from "viem/accounts";

const client = new x402Client().register(
  "eip155:8453",
  new ExactEvmScheme(privateKeyToAccount(process.env.AGENT_KEY)),
);
const pay = wrapFetchWithPayment(fetch, client);

// Not sure it's safe to pay? Preflight it first for $0.03:
// GET https://toolvet.app/api/v1/check?url=https%3A%2F%2Fagent402.tools%2Fv1%2Fauto%2Fchat%2Fcompletions
const res = await pay("https://agent402.tools/v1/auto/chat/completions", {
  method: "POST",
  headers: { "content-type": "application/json" },
  body: JSON.stringify({"max_tokens":5,"messages":[{"content":"Reply with exactly: OK","role":"user"}]}),
});
console.log(await res.json());

Example input

{
  "max_tokens": 5,
  "messages": [
    {
      "content": "Reply with exactly: OK",
      "role": "user"
    }
  ]
}

Example output

{
  "agent402_router": {
    "category": "general",
    "quality": "balanced",
    "served": "openai/gpt-4o-mini"
  },
  "choices": [
    {
      "finish_reason": "stop",
      "index": 0,
      "message": {
        "content": "OK",
        "role": "assistant"
      }
    }
  ],
  "created": 1750000000,
  "id": "gen-…",
  "model": "openai/gpt-4o-mini",
  "object": "chat.completion",
  "usage": {
    "completion_tokens": 1,
    "prompt_tokens": 12,
    "total_tokens": 13
  }
}

Security scan

  • No findings. We scan names, descriptions and tool definitions for hidden instructions and other prompt-injection patterns.

Recent checks

WhenResultHTTPTimePrice
4 h agoPassed402284 ms$0.01