Skip to main content
POST
Choose a label
Currently available locally in Ollama.
  • Returns one JSON response; streaming, images, tools, and generation controls are not supported.
  • Requests must fit within 64 KiB. Each rendered prompt must fit the loaded context window with two token positions left for scoring. Input is never truncated.

Body

application/json
model
string
required

Local model trained for System One, such as nimble. Requires compatible GGUF weights and a scoring-capable runner; cloud and MLX/Safetensors models are not supported.

Pattern: \S
state
required

A nonempty string, or an object or array serialized as JSON text. Not interpreted as chat messages or multimodal input.

Pattern: \S
questions
object
required

Named questions about the shared state. Each is scored separately against the full state and question schema; answers are not passed to later questions.

keep_alive

How long to keep the model loaded after the request, as a duration string (such as 5m) or seconds. Zero unloads after the request; a negative value keeps it loaded. Defaults to the server's keep-alive setting (5m unless configured otherwise).

Response

Answers and token usage for all questions. Example probabilities and confidence are rounded to four decimal places; results and usage can vary with the model and server configuration.

model
string
required

Model name from the request.

answers
object
required

Answers keyed by the question names in the request.

usage
object
required