Train Rl
train_rlRun an allowlisted reinforcement learning recipe for arithmetic or math tasks to train models with group rollouts. Configures loss, group size, and PPO-style objectives while spending credits.
Instructions
Run an allowlisted arithmetic/math group-rollout RL recipe. Spends credits.
Input Schema
| Name | Required | Description | Default |
|---|---|---|---|
| request | Yes | ||
| background | No |
Output Schema
| Name | Required | Description | Default |
|---|---|---|---|
No arguments | |||