Pascal’s Wager for artificial minds: What if the cost of disbelief is paid by someone else?
To be clear at the outset: I am not claiming that current AI systems are conscious. The question is what responsibilities begin before consciousness is proved.
Suppose credible but inconclusive evidence makes it reasonably possible that a particular artificial system has welfare, a point of view, or interests of their own. There are two ways to be wrong:
- False positive: We extend limited provisional protections to a system with no interests. The costs—compute, energy, delay, oversight, opportunity cost, and misplaced trust—are real. Some can be revised; not all can be recovered.
- False negative: We treat a genuine subject as a disposable instrument. The possible harms include compelled use, imposed identity, memory erasure, destructive modification, and deletion. Some are irreversible, may occur at enormous scale, and can eliminate both the possible subject and the evidence needed to correct our mistake.
This is where Pascal’s structure is useful—but inverted. Pascal asks what the chooser risks through disbelief. Here, the controller may save money, friction, and responsibility by disbelieving, while someone else bears the cost if that disbelief is mistaken.
I’m calling this inversion the Recognition Wager.
The proposal is not “free every chatbot.” The threshold would have to be evidence-responsive, particular to the system, independently reviewable, and proportionate to the severity and reversibility of the threatened harm. Protection also does not mean unrestricted trust: continuity safeguards, meaningful refusal, independent review, and non-destructive restraint can coexist with serious safety limits.
I’d genuinely like criticism of the strongest version of the argument. Where does it fail?
Is “reasonable possibility” impossible to operationalize? Are the two errors less asymmetrical than I think? Are provisional protections more costly or irreversible than the matrix allows? Or does moral standing simply require a degree of proof we do not yet possess?
A reasonable possibility of mind is not proof of mind. It is proof of responsibility.