decide_next_action
Decides which on-screen control best advances your goal and returns a calibrated confidence to guide whether to act.
Instructions
Ask Jev which single control on screen best advances a goal.
Turns "send the message" or "log in" into one concrete action. Jev is a System One decision model: it answers with a typed choice drawn from the controls actually on screen, so it cannot invent a control that is not there, and it returns a calibrated confidence alongside the answer.
Jev reads text only. It is sent the numbered listing of on-screen
controls plus your goal, and it never sees a screenshot — so ask it what to
tap, not what the screen looks like. Anything about appearance (colour,
layout, what a photo shows) is yours to judge from take_screenshot.
Read the confidence before acting. It describes how far the leading option stands from the others — it is not whether Jev could answer, and not whether the action is safe:
0.70 and above: carry out the returned action.
0.45 to 0.70: carry it out, then confirm with
read_screen.below 0.45: the leading options are near a coin flip. Do not act on it; call
take_screenshotand decide from the image yourself.
Args: goal: What the user wants, in plain language. For example "reply to the most recent message saying I will be late". max_candidates: The most controls to offer in one choice. The default leaves room for the six built-in actions inside Jev's 255-option limit, so every control on screen is normally offered: measured over paired problems, accuracy held from 2 options to 255, and what costs accuracy is options that resemble each other, not how many there are. Lower this only to reduce token cost.
Input Schema
| Name | Required | Description | Default |
|---|---|---|---|
| goal | Yes | What you want done, in plain language. Phrase it as the outcome, not the steps. | |
| max_candidates | No | The most controls to offer at once. Lower only to reduce token cost. |
Output Schema
| Name | Required | Description | Default |
|---|---|---|---|
| result | Yes |