claude-opus-5 thinking ~42s, with keepalive pings), so a read timeout of 90-120s sized to the inter-event gap is enough — no huge value covering the whole run.
Also: give max_tokens room (start at 64000; thinking eats the budget), and check stop_reason — end_turn is success, max_tokens means truncated so raise it and retry.
The full recipe (streaming code, three-part timeout, retry strategy, scenario cheat sheet) is in the new Long-form Output Practices page: /en/api-capabilities/long-form-output-practices.
← Back to Live · 📚 Monthly archive