Models

Preview

Inspect trained versions, candidate lineage, evaluations, and deployment decisions.

Open Reinforcement Learning → Models at Models to inspect model versions produced by training. This collection is distinct from the hosted model catalog in LLM Gateway.

Review a candidate

Open the version and follow its source training run, starting model, dataset, recipe, grader releases, and evaluation evidence. Inspect artifacts and any readiness or compatibility issues before using it in another run or serving it.

Compare the candidate with the current model on held-out tasks under the same scoring rules. Inspect critical failures and cost or latency changes as well as the aggregate score. An absent or failed check is not evidence of acceptance.

Decide how to use it

Keep the current model, investigate the candidate further, or explicitly accept the candidate where the available controls allow it. Training completion and candidate acceptance are separate from creating or activating a serving deployment. Open Serving for endpoint management and its usage and API guide.

Previous runs retain their captured versions and scores. A later model or grader change does not rewrite that history. See Experiments for fresh comparisons and Training for the creation workflow.