Change Reasoning Effort Mid-Session Without Resetting Your Prompt Cache

Keeps the cached prefix (often thousands of tokens/turn) at the cached-input rate on every effort switch, instead of re-prefilling the whole session at full input price Advanced 4 min read

Changing the request-level reasoning.effort mid-session silently invalidates your prompt cache and re-prefills the entire conversation at full price. On GPT-6 Astra, a configuration_update input item changes effort per turn while leaving the cached prefix intact.

🔒 Pro tip · Advanced

Unlock this tip — and 126 more

This is one of 127 advanced, fact-checked tactics reserved for Pro. Get the full 149-tip library, a searchable archive, and a new tip every morning. Free for 7 days, then $9/mo.

Prefer to browse? The 22 Beginner tips are free forever.