summarize_local
Summarize long files, logs, or transcripts locally on your GPU to reduce cloud token usage and keep frontier models focused. Pass optional focus to bias the summary toward key areas.
Instructions
Summarize a block of text using the local model.
Use to offload long files, logs, transcripts, or docs the cloud model does
not need to fully ingest — call this instead of reading a large blob into the
cloud context. Runs on the user's GPU at no cloud cost. Returns a concise
prose summary; pass focus to bias it toward what matters.
Input Schema
| Name | Required | Description | Default |
|---|---|---|---|
| text | Yes | The content to summarize; may be long (the local context window is configurable). | |
| focus | No | Optional hint to steer the summary, e.g. 'errors and stack traces' or 'API surface only'. | |
| model | No | Ollama model name to run, e.g. 'llama3.1' or 'qwen2.5-coder'. Omit to use the server's configured default model. |
Output Schema
| Name | Required | Description | Default |
|---|---|---|---|
| result | Yes |