Model guidance adds a one-line instruction to cut Opus 5's delay before its first visible output.
What's wrong with this entry?
The bundled model guidance adds a note that Opus 5 sometimes thinks before emitting its first visible block, which raises the delay before anything appears in chat and voice, along with a one-line system instruction to reduce it.
- The suggested line is
Latency-sensitive; begin your visible answer immediately. - The advice is to apply it only where first-token latency is visible to a person, not everywhere.
- It also appears as a new tuning item in the Opus 5 checklist.
Latency-sensitive; begin your visible answer immediately.
Strings lifted out of the shipped bundle, so the claim above can be checked against them.
Related
Other releases about the same thing. Found by shared names or similar wording; neither means one caused the other.
-
v2.1.234
Search field in the model picker
Both mention model
-
v2.1.236
Cancellable waits when
/modeltalks to a cloud sessionBoth mention model
-
v2.1.238
Model and effort switch warning now checks whether the cache is actually warm
Both mention model