Skip to content
Discussion options

You must be logged in to vote
Admin verified this answer by houtanb Jun 25, 2026

For a few reasons:

  • scratchpad / chain of thought prompting worked better in older models than it does in newer models
  • our zero short prompts were attaining SOTA performance (backs up the first point) so the scratchpad ones were redundant
  • our desire to redirect that budget towards models with different reasoning levels and tool use (forthcoming in the next 1 or 2 rounds)

Replies: 1 comment

Comment options

You must be logged in to vote
0 replies
Answer verified by Admin Jun 25, 2026
Sign up for free to join this conversation on GitHub. Already have an account? Sign in to comment
Category
Q&A
Labels
None yet
2 participants