Cheap small-model text completion: prompt in, text (or JSON object) out. Token-capped, schema-declared.
compute_inferCheap small-model text completion for agent sub-tasks: classify, extract, summarize, rewrite. MUST be invoked when an agent needs a short LLM answer without paying flagship rates. Prompt <=8000 chars, maxTokens <=1024, json:true forces a JSON object. Do NOT use for long-form generation or tool calling. Settles $0.003 USDC; no charge on failure.
Input Schema
| Name | Required | Description | Default |
|---|---|---|---|
| json | No | Force a JSON object response | |
| prompt | Yes | User prompt | |
| system | No | Optional system instruction | |
| maxTokens | No | maxTokens |
Output Schema
| Name | Required | Description | Default |
|---|---|---|---|
| ts | No | ||
| model | No | ||
| usage | No | ||
| output | No | ||
| settled | No |