Skip to main content

Create a chat completion

POST 

/v1/chat/completions

Runs an OpenAI chat completion request through the full proxium loop: authentication, memory, the response cache, the budget, routing, failover and cost. A cache hit costs no call slot. The answer is the provider's body without change. With stream: true, the answer is server-sent events. A cache hit is sent as events too.

Request​

Responses​

The provider's answer without change, or the cached answer.