chat_local
Conduct multi-turn chats with a local Ollama model, preserving full conversation context for handoffs that need more than one turn. Runs locally at no cloud cost.
Instructions
Hold a multi-turn chat against a local Ollama model.
Use instead of ask_local when the handoff needs more than one turn of
context — a running conversation or a system + user + assistant history.
Runs locally at no cloud cost. Returns the model's next assistant message
as text.
Input Schema
| Name | Required | Description | Default |
|---|---|---|---|
| model | No | Ollama model name to run, e.g. 'llama3.1' or 'qwen2.5-coder'. Omit to use the server's configured default model. | |
| messages | Yes | Conversation as a list of {"role": "user"|"assistant"|"system", "content": str} messages, in order. |
Output Schema
| Name | Required | Description | Default |
|---|---|---|---|
| result | Yes |