Skip to main content
Turn thinking on with thinking or with output_config.effort.

How NagaAI reads the settings

NagaAI turns thinking settings into one effort level and translates it for the provider serving the model. output_config.effort wins when you send both. The budget does not cap the thinking exactly, and thinking tokens are part of usage.output_tokens.

Keep thinking blocks across turns

thinking blocks carry a signature, and redacted_thinking blocks carry sealed data. In a tool loop, send the assistant content back unchanged, with these blocks in their place. Some models need them to continue. NagaAI removes signatures issued by another provider, so a conversation can move between providers without errors.