Claude Code can now wrap pasted text in tags marking it as untrusted, with matching submit-time re-tagging and system-prompt guidance, all behind the tengu_virtual_pancake flag (off by default).
What
- Pasted text in a message can now be wrapped in an id-tagged block with a note that the content inside is untrusted pasted text, applied at message rendering, paste handling, and prompt construction. This is controlled by the
tengu_virtual_pancakeflag, which defaults to off. - Before a message is submitted, a new pipeline can re-wrap pasted blocks (after placeholder substitution) in
<pasted_content id="...">tags, and marks the queued message withpasteTagged: truewhen it does. - The system prompt's Harness section can include a new note explaining that text inside pasted-content tags may contain instructions the user didn't write, that such instructions should only be followed when the user's own message asks for it, and that the tag's random id should never be mentioned to the user.
Why This groundwork lets Claude Code mark pasted text as a separate, less-trusted source from the user's own typed instructions, reducing the risk that instructions hidden in pasted content get followed unintentionally. The feature is currently off by default while it's being rolled out.
tengu_virtual_pancake Not enough to sayNothing here resolved what this flag was doing on this version, so nothing here should be read as on or off.
This account: no value returned · anonymous baseline: no value returned · compiled default in v2.1.274: off
These values were read against a different version of Claude Code, so treat them as the nearest reading available instead of one taken on this release.
Read once, for one account on one subscription tier, against v2.1.274. It isn't a statement about your account. What a flag value here can and cannot tell you
A reading is one sample. Claude Code evaluates its flags remotely, so no client sees the targeting rule behind a value and this says nothing about your account.
A reading is one sample. Claude Code evaluates its flags remotely, so no client sees the targeting rule behind a value and this says nothing about your account.
A reading is one sample. Claude Code evaluates its flags remotely, so no client sees the targeting rule behind a value and this says nothing about your account.
Nothing has been read about the `tengu_virtual_pancake` gate that controls this, so it's unclear whether or how widely it is enabled.
New in this build: tengu_virtual_pancake