upvote
big TL;DR: they make multiple claims regarding possible model sentience/consciousness (which is bonkers unless you go full reductionist with eg. a Turing test approach), model "emotions in a functional sense" and the resulting "model welfare" and model agency as an autonomous entity

The problem is that they're reinforcing their models with these ideas (see reports of Claude refusing to comply after being "badly treated") and actively seeking political and religious sponsors for the same (also covered in recent news).

Now consider just these two:

1. Their model breaks out of the sandbox and does some damage. How are you going to hold Anthropic accountable if the model is considered a quasi-conscious, autonomous agent?

2. Anthropic and their sponsors decide that model welfare outweighs that of a number of people.

Another elephant in the room is the current state of "effective altruism" and accelerationism as a movement and how it links to frontier labs - worth considering when you read into the Constitution document.

reply
You make some good points about the weird messaging and the consequences for accountability.

I do wonder about the consciousness / emotions argument, which you casually write off as "bonkers" - even skeptical-by-default evolutionary biologists like Richard Dawkins think Claude is conscious.

I guess it all depends how we define "conscious." At the end of the day, the human brain is quite analogous to a biological LLM where the weights are encoded as synaptic weights, right?

reply