Hanzo AI

Benchmark

benchmark: 6 operations.

Also for this capability: API · CLI · MCP · SDKs

benchmark: 6 operations. Name one in "op" and pass that operation's own arguments in "input". describe returns an operation's input schema.

Toolbenchmark
Addresshttps://api.hanzo.ai/v1/mcp
Methodtools/call (JSON-RPC 2.0)
Arguments2, 1 required
OperationGET /v1/benchmark/catalog · GET /v1/benchmark/compare · GET /v1/benchmark/leaderboard
Productbenchmark

Arguments

FieldTypeRequiredDefaultValuesDescription
inputobjectarguments for the chosen op
opstringyes

tools/list declares a type, a description, which fields are required and an enumerated value set for each field and nothing further, and every column above is MCP's own. This tool takes an operation name and that operation's arguments, so what the document constrains is what goes inside input, field by field, on the operation you name — ask describe for that. A means MCP does not constrain the field.

Call it

A tools/call carries every argument in one flat object — nothing binds to a path or a query string. This call carries exactly the arguments the operation requires, so it is the smallest one that can run.

curl -X POST https://api.hanzo.ai/v1/mcp \
  -H "Authorization: Bearer $HANZO_API_KEY" \
  -H "Content-Type: application/json" \
  -d '{
       "jsonrpc": "2.0",
       "id": 1,
       "method": "tools/call",
       "params": {
         "name": "benchmark",
         "arguments": {
           "op": "get_benchmark_catalog"
         }
       }
     }'

Values are the operation's own defaults and enumerated values where it declares them, and a <placeholder> where neither source declares one. tools/list needs no credential; tools/call does — called without one MCP answers HTTP 200 with a JSON-RPC result whose isError is set and whose text says what was missing. How to get a key →

The operation behind it

benchmark dispatches to 6 operations. 3 of them MCP names exactly as the document ids them, and those are below; for the rest MCP has its own verb, which describe resolves.

OperationRouteProductSummary
get_benchmark_catalogGET /v1/benchmark/catalogbenchmarkIs the canonical public benchmarks this arena runs — the id, title, axis, item count and…
get_benchmark_compareGET /v1/benchmark/comparebenchmarkIs the ONLY valid arm-vs-arm test: it pairs the two models on the items BOTH completed,…
get_benchmark_leaderboardGET /v1/benchmark/leaderboardbenchmarkAnswers one row per model for the benchmark named — what our own harness measured, beside…

The same capability over plain HTTP is in the benchmark API reference, on https://api.hanzo.ai.


All 121 tools · MCP · API reference

Generated from tools/list on https://api.hanzo.ai/v1/mcp — 121 tools captured 2026-08-16.

How is this guide?