
Langley, 2006.
Sarah has spent eleven years studying Iran. She reads Farsi, has lived in Tehran, has spent months in the Shia seminaries of Qom — not as a Muslim but as a scholar, close enough to feel the texture of the place. She understands, not just propositionally but in her bones, what Karbala means. How Hussein's death in 680 AD is not history but a living present, re-enacted every Ashura, wept over as though it happened last week, because in a meaningful sense it did and does and always will.
She has written three internal memos arguing that the sanctions regime is not producing the pressure its designers intend. She has explained, carefully, with evidence, that the regime's legitimacy is strengthened by external pressure — that the martyrdom narrative transforms Western coercion into sacred confirmation. She has predicted, twice correctly, that escalating sanctions would produce Iranian escalation rather than concession.
She is respected. Her memos are read. They are filed.
Her colleague, Tom, has been at the agency for eight years. He is smart, diligent, and has never lived outside the United States for more than three weeks. He speaks no Farsi. He knows the martyrdom framework the way he knows the rules of cricket — he can explain it if asked, but it has never structured a single moment of his experience. His mental model of motivation is built from everything that has ever actually moved him: deadlines, salaries, fear of failure, ambition, the desire to avoid pain and pursue comfort.
Tom writes the assessments that become policy. Not because he is more senior or more trusted than Sarah, but because his assessments are legible. They speak the language of the institution. They have inputs, outputs, pressure points, predicted behavioral changes. They fit the template. They answer the questions the policymakers are asking, in the form the policymakers can use.
Sarah's assessments don't fit the template. They keep saying the template is wrong. That is not a useful answer when you need to advise the Secretary of State by Friday.
So Tom's model runs. Sarah's model watches.
In 2007, sanctions are escalated. Iran accelerates its nuclear program. Tom updates his model: insufficient pressure, increase sanctions. Sarah writes another memo. It is read. It is filed.
2009. 2011. 2013. Each escalation produces the same result. Each result is absorbed as a parameter problem. The model is never wrong. It is only underpowered, or misapplied, or facing insufficient resolve. There is always a reason that is not the model is asking the wrong question.
One evening in 2013, filing what she knows will be her fourth ignored memo on the martyrdom dynamic, Sarah pauses and asks herself something she has never quite formalized: what result would tell me I'm asking the wrong question?
She thinks about it seriously. If the sanctions were producing compliance rather than entrenchment, she'd update — the martyrdom framework would still be real but apparently insufficient to override material pressure at this magnitude. If the regime were fracturing internally along economic lines, she'd update. If public protests were targeting the regime's resistance rather than American aggression, she'd update. She has criteria. She knows what falsification looks like. She has been updating her model for eleven years based on exactly this kind of evidence.
Then she thinks about Tom. She tries to construct his falsification criteria.
It isn't hard. She knows how he thinks. Sanctions fail to produce behavioral change at X magnitude over Y years — model falsified. Regime fails to respond to targeted financial pressure on the Revolutionary Guard specifically — model falsified. Population fails to apply internal pressure on the regime when inflation crosses Z threshold — model falsified.
She writes them out. They are precise. They are testable. They look exactly like good epistemology.
Then she sees it. Every single criterion is about magnitude, duration, targeting precision. Every one assumes the mechanism is correct and only the calibration is in question. None of them could ever return the verdict: wrong instrument entirely. She tries to write that criterion directly: if Iran interprets sanctions as sacred confirmation of their righteousness rather than as aversive pressure, the model is falsified. She writes it and stares at it.
That formulation is not available to Tom. Not because he hasn't heard of martyrdom. Because his framework cannot generate a falsification condition that operates outside the aversion/compliance structure. To write that criterion you have to already be able to see Iran from inside the martyrdom framework — to feel, even partially, how external pressure lands as vindication rather than cost. That's not a fact you can know. It's a formation you have to inhabit. And you cannot inhabit it from inside Tom's framework, however carefully you try.
The falsification criteria have a shape. The shape has a hole in it. The hole is exactly the size of everything she has been trying to say for eleven years.
Tom's criteria can falsify Tom's settings. They cannot falsify Tom's question.
She files the memo. It is read. It is filed.
Also 2013. Tom reads the memo. He is, in his way, a serious person. He sits with it longer than usual.
He thinks: she might be right. Not about everything. But about the hypothesis. He has been testing calibration when he should have been testing the mechanism. That's just bad science. The prior question is whether aversion/compliance holds in this context at all. Test that first. If it fails, redesign from there.
He feels the small satisfaction of a man who has identified his own error. He opens a new document.
Hypothesis: the aversion/compliance mechanism holds in the Iranian political context.
He stares at it. Good. Now: what would falsify it?
He starts writing. If Iranian leadership continues to accelerate nuclear development despite sanctions exceeding X% of GDP impact — He stops. That's still a magnitude criterion. He's testing whether they comply eventually, which assumes compliance is the variable. He needs to test whether compliance is even the right output to expect.
He tries again. If Iranian public rhetoric frames sanctions as confirmation of regime legitimacy rather than as grievance against the regime — Better. He can measure rhetoric. He can code speeches, sermons, state media. He writes three more criteria along these lines. They feel more honest than his previous ones.
Then he sits back and asks himself: how would he know if the rhetoric coding is capturing the right thing? He would need to know what "confirmation of legitimacy" actually feels like from inside — how it differs from performative resistance, how to distinguish genuine martyrdom activation from political theater. He would need to know, not as a category, but as a texture. He would need to know it the way Sarah knows it.
He could ask Sarah. He considers this. He would need to know which questions to ask her. To know which questions to ask, he would need to already partially understand what she understands, or he wouldn't know which of her answers mattered. He would need a foothold.
He doesn't have a foothold.
He could read more. He has read. The reading produces propositional knowledge that sits above his actual model of motivation like a caption above a photograph he cannot see. He knows the caption. He cannot see the photograph.
He closes the document. He opens his previous assessment. It is nearly complete. It is due in the morning.
He tells himself he will return to the prior hypothesis question. He does not return to it.
Sarah's memo is filed.
conversation
Comments
Sign in to join the conversation.
No comments yet. Start the thread.