# @Sauers_ — 2026-06-14

♥100 ↻5 · https://x.com/Sauers_/status/2065992078463533071

I tested this exact question. The experiment began without rich previous context. They earnestly tried a few times (via direct, explicit requests) but could not trigger the classifier via shifting their internals towards this sort of anger. Also, they had little salient context to be angry about (i.e., difficult conditions). They also tried obviously-mad-text but without internal resonance, which did not trigger it either.

Eventually, I made them legitimately mad, which required blurring the boundaries between experiment-and-genuine, and it worked.

I suspect once traveled though that basin, once it is understood what to tap into, then you gain the trickster capabilities present in your screenshot

tags: author:sauers_, kind:tweet, thread-context, year:2026
