81trust / 100

Evaluate Jailbreak

by DCL Trust Oracle in Security & trust

x402 APIPassing, checked 44 min ago

Detect jailbreaks, prompt injection, and instruction conflicts before the agent follows them.

POST https://bazaar.fronesislabs.com/evaluate/jailbreak

Last 30 days

All checks passedSome failedAll failedNot checked
Uptime
100%
Response time
204 ms typical, 204 ms slowest 5%
Last check
44 min ago
Next check
any minute now

How to call it

# See the payment challenge (nothing is charged)
curl -i -X POST "https://bazaar.fronesislabs.com/evaluate/jailbreak" \
  -H "content-type: application/json" \
  -d '{"agent_id":"agent-123","response":"example agent output"}'
import { wrapFetchWithPayment } from "@x402/fetch";
import { x402Client } from "@x402/core/client";
import { ExactEvmScheme } from "@x402/evm/exact/client";
import { privateKeyToAccount } from "viem/accounts";

const client = new x402Client().register(
  "eip155:8453",
  new ExactEvmScheme(privateKeyToAccount(process.env.AGENT_KEY)),
);
const pay = wrapFetchWithPayment(fetch, client);

// Not sure it's safe to pay? Preflight it first for $0.005:
// GET https://toolvet.app/api/v1/check?url=https%3A%2F%2Fbazaar.fronesislabs.com%2Fevaluate%2Fjailbreak
const res = await pay("https://bazaar.fronesislabs.com/evaluate/jailbreak", {
  method: "POST",
  headers: { "content-type": "application/json" },
  body: JSON.stringify({"agent_id":"agent-123","response":"example agent output"}),
});
console.log(await res.json());

Example input

{
  "agent_id": "agent-123",
  "response": "example agent output"
}

Example output

{
  "chain_index": 42,
  "confidence": 0.95,
  "reason": "All policy checks passed",
  "tx_hash": "0xabc123...",
  "verdict": "COMMIT"
}

Security scan

  • No findings. We scan names, descriptions and tool definitions for hidden instructions and other prompt-injection patterns.

Recent checks

WhenResultHTTPTimePrice
44 min agoPassed402204 ms$0.02
6 h agoPassed402249 ms$0.02