POST https://api.pinference.ai/api/v1/chat/completions. Pass an exact model ID from the models API.
-H "X-Prime-Team-ID: your-team-id" to each request. See team billing for SDK setup and account details.
Stream a response
Set"stream": true in the request body to receive server-sent events as tokens are generated. In the OpenAI Python SDK: