Run an inference with a base model
Start a generation with the given base model. The request body is the model's inference form (see GET /base-models/{slug}/inference-schema for its shape). Returns immediately with an inference id; poll GET /workspaces/{workspace_id}/inferences/{inference_id} for results.
Path parameters
Id of the workspace that owns the resource.
Id of the workspace that owns the resource.
The base model, named by the id the list endpoint returns, or by its slug, display name, or alias.
The base model, named by the id the list endpoint returns, or by its slug, display name, or alias.
Query parameters
Organize this run under a named session. A new session is created if none matches.
Organize this run under a named session. A new session is created if none matches.
Headers
Opaque key making this request safe to retry. The first request with a given key executes; a later request with the same key returns that first response with Idempotent-Replayed: true instead of executing again. The key covers the method, path, query, and body it was first used with; reusing it for a different request is rejected with 422. Replayable for 24 hours.
Request body
Response
Successful Response
Changes
No recorded changes to this endpoint across all 1 revision of this API.