Before you start
You needcurl, jq, and an API credential authorized for inference. The examples attribute the request to support-assistant; replace that reference with your project.
1. Discover model IDs
CallGET /ai/v1/models:
list-models.sh
2. Generate a response
SetMODEL_ID, then call POST /ai/v1/chat/completions:
generate-response.sh
3. Read the answer
Inspect the response before extracting the assistant message:read-response.sh
finish_reason if generation ends unexpectedly.
This is a direct inference request. To execute a configured agent and track its run history, follow Run an agent.
API references
- Create a chat completion and model discovery.
- List models through MCP. The available operations differ across interfaces.