OpenAI-compatible chat completions served by a self-hosted 27B open-weight LLM. Pay per call in USDC on Base via x402 - no API key or account. Use it for text generation, summarization, classification and agent tool calls that need cheap, private inference. Send a messages array (plus optional model, max_tokens, temperature); receive standard OpenAI chat.completion JSON. A free trial (POST /v1/trial) is open during promo windows.
# 1. Ask the endpoint what it costs (no payment, no wallet needed): curl -i -X POST 'https://api.erb-llm.com/v1/chat/completions' # -> HTTP/1.1 402 Payment Required # the response carries the price, asset and pay-to address. # 2. Pay and retry with any x402 client: npx x402-fetch -X POST 'https://api.erb-llm.com/v1/chat/completions'
Endpoint: https://api.erb-llm.com/v1/chat/completions. Operated by api.erb-llm.com, not by Animica. Price and availability were correct at the last probe on 2026-10-01 and are set by the operator, who may change them.
Extended tier of the x402 LLM gateway: same self-hosted 27B open-weight LLM with the highest output ceiling (32,768 tokens) for long-form writing, reports and f…
Long-form tier of the x402 LLM gateway: same self-hosted 27B open-weight LLM with an 8,192-token output ceiling for summarization, drafting and multi-paragraph …