Source Intelligence

DisclaimerUnofficial, and not affiliated with Anthropic. Nearly all of this is read straight out of what ships: npm bundles, captured prompts, published docs. Anthropic's own notes go in verbatim, marked as theirs. The rest is my reading, and every entry carries the strings behind it. If one looks wrong, vote it down and say why.

All of v2.1.242 Home All releases olderv2.1.241 v2.1.243newer
Claude Code v2.1.242

Eval output options and judge prompt loosened

Under the hood
Useful1 Signal2
Elsewhere

Eval runs can output mock call records and grade results on mock calls alone.

What

Eval runs can select mock call records as an output alongside trace, last message and files, and results carry the mocks block and a flag for runs graded on mock calls alone.

Details
  • The suite adoption record now tracks a separate consent decision alongside the adoption decision, both derived rather than fixed to true.
  • The judge's closing instruction, previously the hardcoded "Respond with exactly one word: PASS or FAIL.", is now supplied by the caller.
Evidence

"mock_calls"

Strings lifted out of the shipped bundle, so the claim above can be checked against them.

See this entry in the whole of v2.1.242 →