OpenAIAnthropicGooglexAIPerplexity
← Back to all updates
Published ·xAI• Minor

xAI added per-request priority scheduling for text inference via `service_tier: "priority"`.

The read

Priority is now explicit at request time, with billing tied to actual priority use.

The frame

It sits on Chat Completions and Responses, giving builders a response field to verify the applied service tier.

Primary source
xAI
Keep reading
See what Nova3 builds
Related updates

This is the world we build in. It moves this fast. We move with it.

Bring us your project