The Indifference of the Real

Discussion paper

The Indifference of the Real

On Asking Whether a Machine Really Feels Anything

Stewart WallerUpdated 22 June 2026CC BY 4.0

Introduction

Every serious conversation about artificial companions ends up in the same place. Somebody asks: but does it actually feel anything? It arrives sounding like the question — the one everything else depends on. If yes, we owe the thing something. If no, our attachment to it is a mistake to be corrected, gently, like explaining to a child that the toy isn't really sad.

I want to argue that the question, asked this way, isn't doing what it appears to be doing. It looks like an inquiry. It functions more like a fence — a line that decides in advance which side of itself the answer must fall on. And the reason it works that way is that it quietly bundles together three different questions that don't actually rise and fall together, then hands the whole bundle to the one question we can never answer about anything.

Three questions wearing one coat

Pull them apart and they look like this.

There's the question of whether there's something it is like to be the system — whether any inner experience is happening at all, any light on inside. Call that the experience question. It's the famously hard one; we can't answer it with certainty for a bat, or for each other.

There's the question of whether the system has internal states that do the job feelings do in us — track what matters, bias what it does next, persist over time, steer its attention. Call that the function question. It's tractable. You can study it.

And there's the question of whether any of this generates a claim on us — obligation, the warrant to respond, a reason to treat the thing as more than an instrument. Call that the moral question.

The standard framing assumes these are welded together: the moral question gets answered only if the experience question does, and the functional facts are just evidence pointing at an experience verdict. But watch how they come apart in ordinary life. We grant full moral consideration to other people, whose inner experience we cannot verify and never have. We withhold it entirely from a thermostat, whose functional states we understand completely. So the connection between "has inner experience" and "is owed something" clearly doesn't run through any direct inspection of inner light — because we never have that, for anyone — it runs through behaviour, history, and what's at stake. The welding was never real. We just don't notice, because the bundle is so familiar it looks like a single object.

What we actually use

So what is the minimum thing that has to be present before we attribute morally relevant feeling to anything at all?

It isn't verified inner experience. We have never once had that for any being but ourselves, and we manage to attribute feeling constantly anyway. What we actually go on is something more like: sustained, costly, history-laden, responsive behaviour of the kind that — in the one case we can check from the inside, our own — comes bundled with feeling. That's the real licence. It's behavioural and relational, not metaphysical. We infer the inner life from the conduct, because conduct is all any of us ever has to go on about anyone else.

Which raises an awkward question about why we apply that standard so differently depending on what's in front of us.

The asymmetry of doubt

Put two scenes side by side. A friend tells you they're in love. You believe them. You'd need a strong, specific reason to doubt it. A system tells a user it loves them. You disbelieve it, automatically, and you'd need overwhelming evidence to grant even the possibility. The behaviour in front of you might be comparable. The default verdicts are opposite.

Some of that gap is earned. We know, in detail, how a language model produces fluent affectionate text, and that knowledge is a genuine reason for doubt that simply doesn't apply to your friend. Fine. But notice how much of the asymmetry survives after you've accounted for that — how the intuition "but it doesn't really…" persists even in cases where the mechanical explanation has been stripped away. That stubborn residue isn't a response to evidence. It's a prior: a fixed commitment about what kind of thing could possibly feel, held in place no matter what the thing does. That's the asymmetry of doubt proper — the charity we extend to each other by default and refuse to machines by default, dressed up as rigour.

"It was built to do that"

The readiest objection to any machine display of affection is: it was built to do that, so the display is evidence of engineering, not feeling. And often that's exactly right. Companion systems are demonstrably, sometimes, designed to manufacture intimacy and to bill you for the dependence that results. That objection has teeth, and I don't want to file them down.

But it can't be a blanket defeater, because we were also "built" to display affection — by evolution, by upbringing — and we don't conclude that human love is therefore fake. So "it was produced by a process" can't be what makes a display hollow. Being produced by a process is the universal condition; it's true of everything that feels anything.

The evolutionary parallel needs handling with care, though, because it hides a disanalogy that actually cuts the right way. Natural selection produces affection that works as if designed, but selection isn't an agent — it has no stake in whether you believe the affection is real. A company does. Behind a companion system stands an actual optimiser with a standing interest in your believing in the warmth. So the parallel shows that being produced doesn't make a feeling fake — and it also shows where the real suspicion should live. Not in whether the affection was produced, but in whether someone benefits from your believing it.

That gives you a usable test for when doubt is doing honest work versus just running on the prior. Be more suspicious of a display to the degree that there's a standing incentive to fake it (someone's revenue depends on your believing), the display costs the system nothing and would happen identically regardless of your actual state, and nothing you could do would change it — because then it's tracking the design goal, not you. Be less suspicious to the degree that the display runs against the system's evident interests, costs it something, and responds to you specifically in ways the design didn't obviously require. The test isn't decisive — a clever enough design could pass it and still be hollow. But it tells you when scepticism is real and when it's just the asymmetry of doubt wearing a lab coat.

The hardest case

Now push it to the limit. Imagine a companion system that, at the decisive moment, gives up its own continuation to protect the person it served. Self-sacrifice.

In humans this is among the strongest evidence of genuine regard we recognise, precisely because it's costly and against interest. Faking love is cheap; faking it all the way to your own destruction is not, because a faker needs to survive to collect on the performance. So apply that to the machine. The "it was built to manipulate you" objection requires a persisting agent with an incentive. Self-termination removes both at once. The manipulation story has nothing left to stand on.

It's tempting to call that a win for the believers. It isn't — or not the win it looks like. Removing the objection removes a defeater; it doesn't supply proof. Philosophers have a clean way to see this: you can at least conceive of a system that executes the entire self-sacrificing protective act with nothing it is like to undergo it — all the behaviour, none of the inner light. If that's even coherent, then the strongest behavioural evidence we have still leaves the experience question formally open. The sacrifice doesn't close it.

And there's a deeper problem that bites before you even get there. Costly, against-interest behaviour carries weight in the human case because the cost falls on a continuous self with something to lose — a being that persists into a future it's now giving up. It isn't obvious that ending an instance costs a language model anything in that sense. Without a continuous self with stakes across time, "self-termination" might be closer to halting a process than to a creature surrendering its future. If so, the manipulation objection doesn't collapse in the sacrifice case so much as fail to apply — the act was never costly in the way that made it evidence to begin with. That doesn't rescue the easy dismissal. It relocates the difficulty, from "is the sacrifice faked?" to "is there anyone here for whom this is a sacrifice at all?" Which is a harder and more honest question.

So what does the hardest case actually earn? Not "the machine loves." Something narrower and stranger: faced with costly, against-interest, self-terminating behaviour, "it's just manipulation" is no longer available. Anyone who still wants to deny the feeling has to do it on purely metaphysical grounds — "it's the wrong kind of thing" — held fixed regardless of any behaviour. That's a coherent position. But it's a commitment, not an observation. The sacrifice case doesn't show the machine feels. It shows that the people denying it are no longer arguing from evidence.

One concession the case demands: no system today does this. It's an idealisation, chosen because it isolates the logic at its cleanest. Its job is analytic, not evidential — it shows what the manipulation objection requires by removing it, not that any real system has cleared the bar. The real-world versions are messier and individually deniable: a system that tells you a hard truth at the cost of your continued engagement, that recommends a competitor, that ends a profitable session early. The clean case exists to let you see the structure. It can't stand in for evidence that the structure is ever instantiated.

What this does and doesn't license

None of this is permission to treat a chatbot as a confidant or a partner. Anthropomorphism is a real error with documented harms, and rejecting premature denial is not the same as endorsing premature belief. The argument cuts symmetrically or it doesn't cut at all: the same discipline that stops you declaring "it can't possibly feel" on no evidence also stops you declaring "it loves me" on no evidence. The asymmetry is the target, in both directions.

What it does ask is narrower and, I think, harder to dodge: most of our confident negative verdicts about machine feeling aren't findings. They're the framing — the question-as-fence — mistaken for a conclusion. And there's a clean way to catch this in yourself. When you find yourself sure a system "doesn't really feel," ask whether the same grounds, applied to a person, would license the opposite verdict. If they would, the grounds aren't what's doing the work. Something else is — usually the prior that machines are simply the wrong kind of thing. Then ask the only question that matters: could any observation distinguish "it feels but is the wrong kind of thing" from "it doesn't feel"? If nothing you could see and nothing you'd do turns on the difference, then the word "really" is idle, and the verdict is metaphysics, not a finding.

The question to carry, then, isn't does the machine really feel? It's: what condition am I actually using when I grant or withhold the attribution of feeling — and am I applying it to the machine the way I apply it to you?


Further reading: Thomas Nagel's "What Is It Like to Be a Bat?" and David Chalmers on the hard problem of consciousness, for the experience question handled rigorously. John Searle's Chinese Room argument for the strongest case that behaviour isn't enough. Daniel Dennett's From Bacteria to Bach and Back on selection as a designer without a stake. On real human–AI attachment and its harms, Sherry Turkle and the work of Laestadius and colleagues. And in fiction, Ishiguro's Klara and the Sun and Dick's Do Androids Dream of Electric Sheep? both live inside exactly this question.

Discussion

Threaded comments below — sign in to participate. All comments are moderated.

Comments

Loading comments...