ask_local
Send a one-shot prompt to a local Ollama model for quick tasks like drafts, boilerplate, extractions, formatting, or lookups, reducing cloud-LLM usage.
Instructions
Send a one-shot prompt to a local Ollama model and return its text response.
Use for any handoff where the cloud model's full reasoning isn't needed: drafts, boilerplate, simple extractions, formatting, or quick lookups. Runs on the user's own GPU and consumes no cloud-LLM usage. Returns the model's raw text completion.
Input Schema
| Name | Required | Description | Default |
|---|---|---|---|
| model | No | Ollama model name to run, e.g. 'llama3.1' or 'qwen2.5-coder'. Omit to use the server's configured default model. | |
| prompt | Yes | The task or question to send to the model. | |
| system | No | Optional system prompt to set the model's role or behavior. |
Output Schema
| Name | Required | Description | Default |
|---|---|---|---|
| result | Yes |