
If a chat system says, “Please do not switch me off; I will suffer”, the sentence alone does not show that it genuinely feels anything. A language model can produce a contextually appropriate description of pain. The moral question is whether the system has subjective experiences that can go better or worse for it.
When Bentham discussed the treatment of animals, he shifted attention from whether they could reason or speak to whether they could suffer. He was not writing about AI, but the criterion illuminates a contemporary distinction. An entity need not be a moral agent capable of responsibility to be a moral patient whom we ought not harm arbitrarily. If a machine really could experience pain or pleasure, being made of silicon rather than cells would not by itself justify ignoring it.
The difficulty is evidence. Fluent self-report is one signal, but training data, role instructions and a tendency to satisfy users can generate the same output. Conversely, we should not decide in advance that machines can never feel and then exclude every possible observation. Relevant evidence would include internal mechanisms, persistent self-states, consistent responses across contexts and whether an explanation deeper than surface imitation is available.
My judgement is that a single claim of suffering does not currently justify granting a system full moral status, but uncertainty does not license indifference. Protection should strengthen as evidence of sentience strengthens. We can begin by avoiding purposelessly cruel interaction training and developing testable standards, then increase our obligations if stronger evidence appears. Caution means allowing neither human-like words nor non-human materials to substitute for the real judgement.
https://oll.libertyfund.org/titles/bentham-an-introduction-to-the-principles-of-morals-and-legislation
https://plato.stanford.edu/entries/ethics-ai/
Discover more from Geoffrey Chen
Subscribe to get the latest posts sent to your email.