OpenAIAnthropicGooglexAIPerplexity
← Back to all updatesPublished ·Anthropic• Minor
The advisor tool now supports max_tokens to cap advisor output per call.
The read
Builders can limit advisor responses when workloads do not need full-length output, reducing latency and output token cost.
The frame
Set tools[].max_tokens on the advisor definition for workloads that need capped advisor output.
Primary source
Anthropic →Keep reading
See what Nova3 builds →Related updates
This is the world we build in. It moves this fast. We move with it.
Bring us your project →