evaluate_prompt
Get pre-run feedback on a prompt's clarity and specificity before running a model comparison, using your credentials or server-side keys.
Instructions
Get pre-run feedback on a prompt's clarity/specificity before running a comparison.
An explicit, separately-triggered LLM call (uses your credentials) — not run
automatically as part of run_comparison. Rate-limited independently from
run_comparison's 3-per-8h budget. Pass `creds` as {"openrouter"?: str,
"bedrock"?: {...}, "vertex"?: {...}, "foundry"?: {...}} to use Amazon Bedrock,
Google Vertex AI, or Microsoft Foundry; a bare `api_key` is treated as an
OpenRouter key. `judge_backend` picks which backend runs the evaluation.
If the operator has set server-side keys for a backend (see https://github.com/thejaredchapman/evalforge-lite/blob/main/docs/hosting-and-server-keys.md), those are used for it automatically and creds for it are not needed.
Input Schema
| Name | Required | Description | Default |
|---|---|---|---|
| creds | No | ||
| prompt | Yes | ||
| api_key | No | ||
| judge_backend | No | openrouter |