Source Intelligence

DisclaimerUnofficial, and not affiliated with Anthropic. Nearly all of this is read straight out of what ships: npm bundles, captured prompts, published docs. Anthropic's own notes go in verbatim, marked as theirs. The rest is my reading, and every entry carries the strings behind it. If one looks wrong, vote it down and say why.

All of v2.1.224 Home All releases olderv2.1.223 v2.1.225newer

Grading results record whether a criterion was scored

Under the hood
Useful1 Signal2
Elsewhere

Evaluation grader output now says explicitly whether each criterion was scored.

What

Evaluation grader output now carries an explicit flag for whether each criterion was actually scored, derived as the negation of the existing "with only" flag.

Details
  • The result schema gains an optional boolean scored alongside withOnly, judgeVotes and evidence.
Evidence

withOnly

Strings lifted out of the shipped bundle, so the claim above can be checked against them.

See this entry in the whole of v2.1.224 →