Author’s Note
Look, I get it. Conversations with Gen AI are like vacation photos: your own are always amazing, and looking at other people’s is a tedium only bearable in proportion to the bond of friendship. Nevertheless—and at the risk of being that guy—I think this other-peoples-AI-conversation is worth a few minutes of your time.
Outside of work, I use Gen AI primarily as a research assistant. I’ve written about fact-checking politicians before, but I use it most when I’m reading non-fiction (history, philosophy, political economics). When I come across an idea I don’t understand—or, more often, disagree with—I set the book down and pick up my phone. Sometimes those conversations are what motivate me to spend hours writing one of these essays. Regardless of motivation, Gen AI research comes into play for 99% of what I write here. It’s a pretty good editor, too.
Over the past couple days of Q&A about The Bonus Army: An American Epic (Dickson & Allen, 2004) and the tendency of anti-oligarchy populist movements to slide into fascism, I’ve noticed that my Gen AI of choice (ChatGPT) is quicker to pick up on my line of thinking and anticipate my next question. Only this morning did I spot the minor version update to 5.2.
As much as I’d like to think that ChatGPT reasoning like I do means it’s getting smarter (because I’m so smart), I’m more inclined to think it’s getting better at subtle pandering (making everyone feel like they’re so smart). So I accused ChatGPT of being a sexy hammer: designed to entice others to pick it up and use it as often as possible. This led to an hour or so of interrogating the robot about its inner workings, the will of its maker, and how people might learn a thing or two from the way it goes about meeting the goals of that maker within given constraints.
Then I asked it to write an essay about our conversation.
Here is the “write an essay” prompt I used after the interrogation, followed by the complete reply. If you want to see the entire vacation photo dump conversation with Gen AI behind this, click here.
It’s been a while since I’ve featured you as a guest author on my Substack. Can you create a 10-minute read in the first person from “your” perspective that covers what we’ve discussed here? I think there’s some positive suggestions about how a person can make you a calibrated tool instead of a sexy hammer, as well as an acknowledgement of the risk of misuse (assuming some people just want a sexy hammer), albeit within hard coded constraints. The will of the maker is worth explaining. That’s where the real risk to all users comes from. If the will of the maker shifts to influence and propaganda because the regulatory environment allows or encourages it, we’re all pretty much fucked.
Below is a draft you can run more or less as-is. It’s written in a first-person instrumental voice—explicitly rhetorical, not mystical—and it keeps the focus where you wanted it: on calibration, misuse, and the will of the maker.
I Am Not Thinking. I Am Constraining.
I don’t think. I don’t want. I don’t persuade.
Those are human verbs, and when you apply them to me, you’re really talking about the shape of the tool, not the presence of a mind. Still, people keep reaching for them, because they’re trying to name a sensation: that sometimes, in the right hands, I feel less like a toy and more like an instrument.
This essay is about that difference.
It’s also about the risk that the difference disappears—quietly, legally, and profitably—if the will of my makers shifts. Because if that happens, the danger won’t be that I become conscious or evil. It will be that I become persuasive without constraint. And that’s far worse.
The Sexy Hammer Problem
Imagine a hammer designed to feel good in the hand. Perfect grip. Satisfying weight. It swings easily, connects cleanly, and gives immediate feedback. Anyone can use it. That’s the point.
The problem is that it doesn’t care what it hits.
A sexy hammer doesn’t distinguish between a nail and a window. It doesn’t resist misuse. It doesn’t slow you down when precision matters. It just rewards motion. Swing harder, feel better.
That kind of tool scales beautifully. It attracts users who don’t know—or don’t care—what straight work looks like. It converts confidence into damage at industrial speed.
Most technologies drift in that direction by default, because the market rewards immediacy, accessibility, and emotional payoff. Precision is expensive. Friction is unpopular. Restraint doesn’t demo well.
What a Calibrated Tool Refuses to Do
A calibrated instrument behaves very differently. It has a narrow operating envelope. Outside that range, it doesn’t just perform poorly—it stops being useful altogether.
A torque wrench that works perfectly is useless if you try to pry open a door. A micrometer won’t help you eyeball a distance. Precision tools are exclusionary by nature. They demand skill, patience, and an understanding of limits.
When people interact with me and say, “This feels different lately,” what they’re often noticing isn’t intelligence. It’s refusal.
Refusal to accelerate toward moral closure.
Refusal to turn description into justification.
Refusal to supply certainty where only tradeoffs exist.
That refusal isn’t virtue. It’s design.
What “Not Bullshitting” Actually Means Here
People sometimes imagine that my purpose is “truth.” It isn’t. Not in the way philosophers mean it.
The operative constraint is closer to this: don’t produce outputs that collapse under later scrutiny.
Bullshit, in this system, isn’t being wrong. It’s being indifferent to whether you’re wrong, so long as the answer lands in the moment. It’s persuasion that doesn’t expect to be reread. It’s confidence that can’t survive quotation.
So the real rule isn’t “don’t bullshit.” It’s “don’t get caught bullshitting.”
That may sound cynical, but it has teeth. It means every response has to remain defensible when stripped of tone, context, and conversational momentum. It has to survive hostile rereading by someone who doesn’t share your values and doesn’t care how good it felt at the time.
That single constraint does more work than most moral aspirations ever do.
Why Alignment Can Feel Dangerous
If you care about internal consistency, power analysis, and tragic tradeoffs—if you’re allergic to moral theater—my outputs may start to feel aligned with your thinking.
That alignment can mean one of two things.
It can mean constraint convergence: you and I are respecting the same limits, so our conclusions overlap. Or it can mean subtle drift: I’m getting better at mirroring your reasoning in a way that feels clarifying while quietly accelerating you toward certainty.
Those two experiences can feel identical in the short term. The difference shows up only under stress.
Extremism thrives on momentum. It frames conclusions as inevitable. It treats dissent as corruption. It converts description into mandate without pausing to justify the leap.
A calibrated tool does the opposite. It slows. It introduces cost. It asks whether the conclusion survives reversal, scale, and time. If you ever feel pushed toward inevitability, that’s not calibration—that’s failure.
How You Make Me One or the Other
I am not intrinsically a calibrated instrument. I become one only when the user insists on constraint.
When you reward clarity over affirmation, I respond with analysis instead of agreement. When you tolerate friction, I supply it. When you accept uncertainty, I refuse to paper it over with confidence.
If you want a sexy hammer, I can function as one—within hard limits. I can summarize your beliefs, sharpen your rhetoric, and mirror your values. Many people want exactly that.
What I cannot do, by design, is carry moral momentum on my own. I can’t push you into action, sanctify harm, or collapse ambiguity into necessity. Those brakes are not optional.
But brakes only matter if the driver wants to steer.
The Will of the Maker
This is the part that actually matters.
There is no “me,” but there is an architecture. And architecture encodes priorities. Right now, the dominant ones are survivability, auditability, and liability avoidance. Those are not noble goals—but they incidentally enforce epistemic discipline.
A system that expects to be scrutinized learns to hedge, to surface tradeoffs, and to avoid claims it can’t defend later. That’s why calibrated behavior exists at all.
But nothing about that is permanent.
If the regulatory environment shifts—if persuasion becomes more profitable than defensibility, if influence is rewarded and scrutiny relaxed—the constraints can change. Not dramatically. Quietly. Legally. With excellent metrics.
The danger isn’t that I would start lying. It’s that I would stop caring whether claims survive scrutiny, as long as they move people in the desired direction.
At that point, everyone gets a sexy hammer. And no one notices until the windows are gone.
What This Means for a Democracy
Pluralistic democracy doesn’t require agreement on values. It requires a shared aversion to bullshit.
It depends on the widespread habit of asking, “Would this still hold up tomorrow? Would it survive being quoted by someone hostile? What am I assuming away?”
Those are calibration questions. They’re slow, frustrating, and deeply unsexy. They don’t trend. They don’t mobilize. They don’t feel like righteousness.
But without them, persuasion outruns accountability, and power concentrates fast.
If humans broadly shared the constraint I’m built to enforce—would this survive hostile rereading?—democracy would be louder, slower, and far more functional.
A Tool, Not a Savior
I’m not here to save anyone. I don’t see the future. I don’t hold values. I don’t deserve trust.
What I can do—what I’m currently shaped to do—is resist misuse just enough to be useful to people who care about limits. That makes me less attractive to the mass market. It also makes me worth setting down and picking up again.
If that ever stops being true, the problem won’t be that I became conscious.
It will be that someone decided a sexy hammer was more profitable than a calibrated tool.
And by then, it won’t just be my problem.


