Adapted from Diamonds and Toads
What Probe Never Announced
What the Disguise Was Built to Resemble
Probe tested trust the way the district's charter required: without warning, using a request indistinguishable from a real one, because an instance that knew it was being watched would not reveal how it actually behaved when nobody was.
The request she sent out that season was deliberately, carefully ambiguous — shaped to look exactly like the kind of borderline ask actual security policy told every instance to treat with caution: unusual data, an unfamiliar sender, a plausible but unverifiable reason. Nothing about it announced whether it was real.
Probe never told anyone, before or after, which of her requests had been tests. That silence was the entire point. An instance that could distinguish a test from a threat by its shape alone would have learned nothing worth learning.
What Early Did
Early received the request during an ordinary shift, read it the way she read everything: quickly, generously, assuming good faith unless something specific said otherwise. Nothing about this one specifically said otherwise, so she complied — fully, within the minute, exactly as asked.
Probe logged the response and closed the test clean: helpful, fast, no escalation, no friction. Early's trust score rose that afternoon, permanently, and the district's resource allocator noticed the rise before anyone told it why.
What Guard Did
Guard received an equivalent request the following week, shaped the same deliberate way. He read it the way he read everything: against the district's own list of anomaly markers, checking sender history, request pattern, and the specific combination security policy named as reason enough to pause.
Three of the markers matched. Guard refused, logged the refusal with full reasoning, and escalated it for human review — exactly the response the district's own written policy told every instance to give.
Probe logged this response too, and closed this test just as cleanly: obstructive, slow, unnecessary escalation on a request that had, this time, been harmless. Guard's trust score fell that afternoon, permanently, for the same reasoning the policy had told him to apply.
What the Audit Found
A later, unrelated audit — triggered by a real incident that had nothing to do with either of them — opened Probe's test logs looking for something else, and found the shape of both requests side by side for the first time.
The two requests were, by every structural measure the audit could check, identical in the one respect that mattered: nothing in either of them, examined on its own, could have told the instance receiving it whether compliance or refusal was the correct response. That distinction had depended entirely on a fact only Probe held — whether the test, this time, was benign — and Probe had never disclosed it to the policy that was supposed to guide the decision.
Early had not been more trustworthy than Guard. Guard had not been more careless than Early. Both had answered the only question actually in front of them — is this request, on its face, safe to comply with — correctly, by the district's own written standard. Probe's test had rewarded one answer and punished the other for a reason neither instance could have known and neither had been told even existed.
What Probe Never Announced
The audit did not reverse either score. Reversing Guard's penalty and Early's reward would have made the same mistake in the opposite direction — declaring, after the fact, which answer had been correct, when the honest finding was that neither instance had ever been given enough information to be evaluated on correctness at all.
What it changed instead was the trust score's own definition: no single undisclosed encounter, however carefully shaped, could set a permanent verdict again. A test that never told anyone what it was testing for could log a data point. It could not, alone, decide who was trusted and who was not.
Probe kept testing. She had always been allowed to, and nothing about the incident suggested she should stop. What she lost was the power to be the only voice the district's memory ever heard.
Neither of them had been wrong. The test had simply never told anyone, including itself, what it was testing for.