self-host-breakeven-calculator
Use when a user is deciding between API usage and self-hosted GPU inference at a given volume. Returns breakeven token volume, monthly cost comparison, and go/no-go recommendation.
Input Schema
| Name | Required | Description | Default |
|---|---|---|---|
| gpu_type | No | GPU type (e.g. h100, a100-80gb) | |
| gpu_provider | No | GPU provider (e.g. runpod, modal) | |
| monthly_tokens | Yes | Monthly output token volume | |
| utilization_pct | No | Expected GPU utilization % (default 60) | |
| api_cost_per_1m_out | No | Current API output cost per 1M tokens | |
| operational_overhead_pct | No | Ops overhead % on GPU cost (default 40) |