Skip to main content
You can ask the model to spend more effort on a reply. The response you get back is still the reply itself. It does not include a reasoning trace, a thinking block, or hidden reasoning text. Leaving the field out is the same as asking for no extra effort. The examples use lume-3.5.

Chat Completions

POST /v1/chat/completions needs the scope chat:completions. Set reasoning_effort to none or high.
The success body is an ordinary chat completion. Read the reply from choices[0].message.content. tool_calls or function_call appear only when the model decided to call a tool. There is no reasoning field on the message. Any other value for reasoning_effort is HTTP 400:
The response includes X-Request-Id.

Messages

POST /v1/messages needs messages:create and anthropic-version: 2023-06-01. You can set effort in either of two places. thinking.type maps like this: output_config.effort maps low to no extra effort, and medium, high, and max to high effort. If you send both, output_config.effort wins. effort set to low means no extra effort even when thinking.type is enabled or adaptive. The two fields do not combine. If you send only one, that field is the one that counts.
The success body is an ordinary Messages reply. content holds text or tool_use. It does not hold a thinking block, a redacted_thinking block, or reasoning text. If you send those blocks on an earlier assistant turn, they are checked and then dropped.

Next

Generate text

The same reply with no extra effort.

Structured outputs

Ask the reply to be JSON.