analyze_inference_config
Optimize LLM inference by scanning GPU details, benchmarking memory bandwidth, and checking ML runtimes and environment variables.
Instructions
Performs a deep scan for LLM inference optimization: GPU details, real memory bandwidth benchmark, ML runtimes (Ollama, Docker, WSL), and environment variables.
Input Schema
TableJSON Schema
| Name | Required | Description | Default |
|---|---|---|---|
No arguments | |||