Skip to main content
For Chat Completions, use reasoning_effort with a model whose catalog entry supports reasoning. low, medium, and high are the common values exposed by the SERV Playground; the upstream provider determines the exact behavior for each model.
Run your evaluation set with low first. Increase the effort only when the results improve enough to justify the added cost and latency.