x402-llm-gateway

OpenAI-compatible chat completions served by a self-hosted 27B open-weight LLM. Pay per call in USDC on Base via x402 - no API key or account. Use it for text generation, summarization, classification and agent tool calls that need cheap, private inference. Send a messages array (plus optional model, max_tokens, temperature); receive standard OpenAI chat.completion JSON. A free trial (POST /v1/trial) is open during promo windows.

Inferenceoffline · HTTP 530eip155:8453Exactvia cdp
LlmChatInferenceOpenai-compatibleUsdc
Calls · 30d
17→ 0%
This endpoint's own trailing-30-day call count, as published by the upstream catalog and snapshotted daily. 5 snapshots so far.
$1.43
Verified settled volume
59 settlements proven x402 by their on-chain EIP-3009 marker.
$0.0050
Listed price
As published in the catalog. Always read the live 402 before paying.
17
Calls · 30d
Upstream's own call count for this endpoint, not ours.
2
Unique payers · 30d
Last called 2026-09-08 21:10Z
Upstream on-chain volume
Reported by the source catalog.
Paid to
0x23976FB6Af8f7fE8756730f0473fB7FfaC531ee8

The wallet the 402 directs payment to. Its whole payment record — every payer, every chain — is on the merchant page.

Asset 0x833589fCD6eDb6E08f4c7C32D4f71b54bdA02913

Provider

The payTo wallet does not resolve to a registered ERC-8004 agent. That is not a verdict on the service — most of the catalog is unregistered.

Accepts

The payment requirements as published to the catalog. Read the live 402 before paying — a price here is a claim, not a quote.

0.005USDC≈ $0.005 USD
on Base · exact scheme

Pay 0.005 USDC on Base to 0x2397…31ee8. The signed payment is good for 5 minutes.

USD Coin contract
0x8335…02913
Payment window
5 minutes
As published
5000 smallest units
The catalog’s raw entry
[
  {
    "asset": "0x833589fCD6eDb6E08f4c7C32D4f71b54bdA02913",
    "extra": {
      "name": "USD Coin",
      "version": "2"
    },
    "payTo": "0x23976FB6Af8f7fE8756730f0473fB7FfaC531ee8",
    "amount": "5000",
    "scheme": "exact",
    "network": "eip155:8453",
    "maxTimeoutSeconds": 300
  }
]

Extensions

{
  "bazaar": {
    "info": {
      "input": {
        "body": {
          "model": "qwen/qwen3.8-27b",
          "messages": [
            {
              "role": "system",
              "content": "You are a helpful assistant."
            },
            {
              "role": "user",
              "content": "Reply with exactly one word: hello"
            }
          ],
          "max_tokens": 16,
          "temperature": 0.2
        },
        "type": "http",
        "method": "POST",
        "bodyType": "json"
      },
      "output": {
        "type": "json",
        "example": {
          "id": "chatcmpl-abc123",
          "x402": {
            "tier": "quick",
            "charged": "$0.005",
            "output_cap": 4096,
            "tokens_out": 1,
            "inference_ms": 1830
          },
          "model": "qwen/qwen3.8-27b",
          "usage": {
            "total_tokens": 31,
            "prompt_tokens": 30,
            "completion_tokens": 1
          },
          "object": "chat.completion",
          "choices": [
            {
              "index": 0,
              "message": {
                "role": "assistant",
                "content": "hello"
              },
              "finish_reason": "stop"
            }
          ]
        }
      }
    },
    "schema": {
      "type": "object",
      "$schema": "https://json-schema.org/draft/2020-12/schema",
      "required": [
        "input"
      ],
      "properties": {
        "input": {
          "type": "object",
          "required": [
            "type",
            "method",
            "bodyType",
            "body"
          ],
          "properties": {
            "body": {
              "type": "object",
              "required": [
                "messages"
              ],
              "properties": {
                "model": {
                  "type": "string",
                  "description": "Model ID from the free GET /v1/models endpoint (e.g. qwen/qwen3.8-27b). Optional - a default chat model is chosen if omitted."
                },
                "messages": {
                  "type": "array",
                  "items": {
                    "type": "object",
                    "required": [
                      "role",
                      "content"
                    ],
                    "properties": {
                      "role": {
                        "enum": [
                          "system",
                          "user",
                          "assistant"
                        ],
                        "type": "string"
                      },
                      "content": {
                        "type": "string"
                      }
                    }
                  },
                  "minItems": 1,
                  "description": "OpenAI chat messages (system/user/assistant)."
                },
                "max_tokens": {
                  "type": "integer",
                  "minimum": 1,
                  "description": "Optional output token ceiling. This tier caps output at 4096 tokens - higher values are clamped to 4096, not rejected. Use the long/extended tiers for bigger outputs."
                },
                "temperature": {
                  "type": "number",
                  "maximum": 2,
                  "minimum": 0
                }
              }
            },
            "type": {
              "type": "string",
              "const": "http"
            },
            "method": {
              "enum": [
                "POST",
                "PUT",
                "PATCH"
              ],
              "type": "string"
            },
            "bodyType": {
              "enum": [
                "json",
                "form-data",
                "text"
              ],
              "type": "string"
            }
          },
          "additionalProperties": false
        },
        "output": {
          "type": "object",
          "required": [
            "type"
          ],
          "properties": {
            "type": {
              "type": "string"
            },
            "example": {
              "type": "object",
              "properties": {
                "id": {
                  "type": "string"
                },
                "x402": {
                  "type": "object",
                  "properties": {
                    "note": {
                      "type": "string"
                    },
                    "tier": {
                      "type": "string"
                    },
                    "charged": {
                      "type": "string"
                    },
                    "output_cap": {
                      "type": "integer"
                    },
                    "tokens_out": {
                      "type": "integer"
                    },
                    "capped_from": {
                      "type": "integer"
                    },
                    "inference_ms": {
                      "type": "integer"
                    }
                  },
                  "description": "Billing stamp added by the gateway."
                },
                "model": {
                  "type": "string"
                },
                "usage": {
                  "type": "object",
                  "properties": {
                    "total_tokens": {
                      "type": "integer"
                    },
                    "prompt_tokens": {
                      "type": "integer"
                    },
                    "completion_tokens": {
                      "type": "integer"
                    }
                  }
                },
                "object": {
                  "const": "chat.completion"
                },
                "choices": {
                  "type": "array",
                  "items": {
                    "type": "object",
                    "properties": {
                      "index": {
                        "type": "integer"
                      },
                      "message": {
                        "type": "object",
                        "properties": {
                          "role": {
                            "type": "string"
                          },
                          "content": {
                            "type": "string"
                          }
                        }
                      },
                      "finish_reason": {
                        "type": "string"
                      }
                    }
                  }
                }
              }
            }
          }
        }
      }
    }
  }
}

Provenance

Seen in the source catalog
2026-09-08 21:10Z
Last indexed by Roundhouse
2026-09-09 18:40Z
Last enriched (probe, favicon, geo)
2026-09-24 07:15Z
x402 version
2
Max timeout
300s
Liveness probe
HTTP 530
Report

Hand this page to an agent

Copy the prompt and paste it into Claude, an MCP client or your own agent — it will vet this service and call it over the free read API. No key, no account.

GET api.roundhouseai.io/v0/endpoints

This endpoint's own trailing-30-day call count, as published by the upstream catalog and snapshotted daily. 5 snapshots so far. Verified volume counts only settlements with an on-chain EIP-3009 marker.

Open skill.md
Show the prompt
Using Roundhouse, look up the x402 service x402-llm-gateway and tell me whether it is
worth paying: what a call costs, whether the endpoint answered when last probed, and what
its payment record actually shows.

curl -s 'https://api.roundhouseai.io/v0/endpoints?q=x402-llm-gateway'
curl -s 'https://api.roundhouseai.io/v0/merchants/<the payTo wallet returned above>'

Then call it: read the price from the live 402 at https://api.erb-llm.com/v1/chat/completions, never from
a cached figure, and pay with an x402 client.

The /v0 API needs an API key (`authorization: Bearer rh_live_…`) on everything except
/v0/unified* and /v0/endpoints. Mint a personal key for $0.01 at GET https://api.roundhouseai.io/v0/test/x402,
or use an organization key from https://roundhouseai.io/dashboard/team.

If you do not have Roundhouse tools or skills installed, read
https://roundhouseai.io/skill.md first — it is the whole procedure.