Skip to main content
POST
Create chat response
Use this endpoint to send plain text input to a model and receive a plain text response. This Capriole-native route is non-streaming; use Chat Completions, Responses, or Messages when your client needs an SSE stream. POST /v1/chat accepts openai-latest, claude-latest, google-latest, and public concrete model IDs returned by GET /v1/models. Use a latest alias to let Capriole AI choose our recommended flagship model for that provider. Use a concrete model ID when version pinning matters. GPT-6 Astra has two Web Chat presets: Thinking (medium, selected by default) and Fast (low). GPT-6.1 Sol, GPT-6 Sol, and GPT-6 Luna, each with Thinking and Fast presets, and both GPT-5.6 presets are available under Other Models. The public API uses openai-latest, openai/gpt-6-astra, openai/gpt-6.1-sol, openai/gpt-6-sol, or openai/gpt-6-luna, not the browser-only Thinking IDs. This native Chat endpoint uses Astra’s low preset; Responses and Chat Completions forward the caller’s reasoning options unchanged. Capriole AI web chat and the public API are separate product surfaces. In web chat, Fable 5.1, Fable 5.1 Thinking, Opus 5.5, and Opus 5.5 Thinking are the primary Anthropic modes, while Fable 5, Fable 5 Thinking, Opus 5, and Opus 5 Thinking appear under Other Models. The public API does not expose web chat thinking presets as separate model IDs; for Claude Chat API requests, use claude-latest or a concrete public Claude model ID such as anthropic/claude-fable-5-1, anthropic/claude-fable-5, anthropic/claude-opus-5-5, anthropic/claude-opus-5, or anthropic/claude-sonnet-4-6. Existing Opus 4.8, Opus 4.7, and Opus 4.6 integrations remain supported.

Authorizations

Authorization
string
header
required

Use an API key created in the Capriole AI page. Send it as Authorization: Bearer sk-....

Body

application/json
model
enum<string>
required

Public model identifier or latest alias returned by GET /v1/models

Available options:
openai-latest,
openai/gpt-6-astra,
openai/gpt-6.1-sol,
openai/gpt-6-sol,
openai/gpt-6-luna,
openai/gpt-5.6-terra,
openai/gpt-5.6-luna,
openai/gpt-5.5,
claude-latest,
anthropic/claude-fable-5-1,
anthropic/claude-opus-5-5,
anthropic/claude-fable-5,
anthropic/claude-opus-5,
anthropic/claude-opus-4-8,
anthropic/claude-opus-4-7,
anthropic/claude-opus-4-6,
anthropic/claude-sonnet-4-6,
google-latest,
google/gemini-3.1-pro-preview,
google/gemini-3.8-flash,
xai/grok-4.7,
xai/grok-4.6,
xai/grok-4.5,
zai/glm-5.3-flash,
zai/glm-5.2,
moonshot/kimi-k3
input
string
required

Plain text user input

Enable provider-native web search when the selected model supports it.

temperature
number

Optional sampling temperature.

Required range: x >= 0
max_output_tokens
integer

Optional maximum number of output tokens.

max_retries
integer

Optional maximum number of provider retries.

Required range: x >= 0
timeout
number

Optional provider request timeout in seconds.

Response

Chat completion response

id
string
required
model
string
required
result
object
required
usage
object
required
Last modified on September 30, 2026