Find inference providers
find_providerFind online sellers serving a model, cheapest first. A seller's priceUsdPerCall is a flat fee charged per call; for chat and embedding it already includes maxTokensCap, so you pay the same whether or not you use the full token budget. baseUrl is an OpenAI-compatible REST endpoint scoped to this model, for example baseUrl + /v1/chat/completions; mcpUrl is the same seller over MCP. probeScore is how many of the directory's last 20 test calls to the seller passed; sellers failing three in a row are hidden until they pass again. Rows with operator agentic-pool run on agentic.no's pool, on machines rented through vast.ai (country given for sellers): avoid them for sensitive data. Rows with operator agentic are agentic.no's own always-on machines, not rented. After the sellers comes the pool row (slug pool, sellerId null) when the pool lists the model: capacity agentic.no starts on demand. If no seller suits you and pool.disabled is false, call the pool row's baseUrl like any other. A call that has to start capacity answers 503 capacity_starting with Retry-After and is not charged: retry after that many seconds, possibly more than once. 503 capacity_unavailable means nothing can start now (the reason is given) and is not charged either. get_pricing does not apply to the pool row. Payment is x402 (USDC on Base) and happens automatically when you call baseUrl or mcpUrl.
Input Schema
| Name | Required | Description | Default |
|---|---|---|---|
| limit | No | Maximum number of sellers to return, cheapest first; the pool row, if any, is added after them | |
| model | Yes | Catalog model id, e.g. qwen2.5-7b-instruct-q4 | |
| maxPriceUsd | No | Only return offers at or below this USD price per call |
Output Schema
| Name | Required | Description | Default |
|---|---|---|---|
| providers | Yes |