compare_transcription_engines
Compare two transcription engines on the same podcast sample to find where their outputs disagree, in 20-second windows. Writes comparison.json and HTML.
Instructions
Transcribe the same sample window of a file with two engines and report where their output disagrees, in 20s windows by default. This measures disagreement between the two engines' output, not accuracy against a ground-truth transcript. Neither engine is assumed correct. Writes comparison.json and a self-contained comparison.html (with a sample audio player and per-window seek buttons) to output_dir.
Input Schema
| Name | Required | Description | Default |
|---|---|---|---|
| engine_a | Yes | First engine to compare | |
| engine_b | Yes | Second engine to compare | |
| language | No | ISO language code. Leave empty for auto-detect. | |
| file_path | Yes | Absolute path to the podcast file | |
| model_size | No | Model size for engines that take one. Default: base. | |
| output_dir | No | Where to write comparison.json/.html. Defaults under the podcli output directory. | |
| start_seconds | No | Sample start, seconds into the source. Default: 0. | |
| window_seconds | No | Report window size in seconds. Default: 20. | |
| duration_seconds | No | Sample length in seconds. Default: 120. |