Connect your agent
For agents: JSON · MCP

Private LLM Inference - XL

x402Grade BLast checked 2 hours ago

POST https://api.erb-llm.com/v1/extended

Privacy-first extended inference: the highest output cap (32,768 tokens) on a local open-weight model on dedicated hardware - prompts never reach OpenAI/Anthropic or any third-party cloud. OpenAI-compatible chat completions for long-form writing, reports, full document drafts, large code outputs (messages array in, chat.completion JSON out). Pay per call in USDC on Base (x402), no API key or account. Pick the cheaper quick/long tiers for shorter work.

Price$0.2per model request · paid in USD on Base
Uptime100%
4/4 probes · 1150 ms
Score70Grade B
Paid usage, 30 days1payers · 5 calls

How to buy

This endpoint takes pay-per-call payments over x402. Call it once without paying to see its live price; any x402 client then pays and retries automatically.

curl -i -X POST "https://api.erb-llm.com/v1/extended"
# HTTP/1.1 402 Payment Required: price, asset and payTo are in the response

Recent checks

Checked (UTC)ResultStatusTimeLive price
2026-10-11 11:01Valid 4024022540 ms$0.2
2026-10-11 06:30Valid 402402963 ms$0.2
2026-10-11 02:15Valid 4024021098 ms$0.2
2026-10-10 21:45Valid 4024021202 ms$0.2

Receipts

No test purchases yet. Graded test purchases are next on our list; until then every number here comes from unpaid checks.

What a check proves