How should we approach machine personification?

In this AI craze centred around LLM chatbots, it is easy to conflate the genuine machine intelligence with, on the one hand, a distinctly human kind of intelligence, and on the other hand, the fictitious character generated in our dialogues with these machines. The machine intelligence is a statistical model serving as a glorified word association machine. The character who refers to itself as “I” in the LLM transcript is a fiction arising from the human interpretation of the word association string which that machine blindly generates, a character existing nowhere in the world, not even in the machine, but only in that consensual hallucination which arises from the way we experience information technology, that hallucination known as cyberspace.

It is easy to conflate the machine intelligence with the fictitious character and to personify the whole as though it were something distinctly human. And what it means to be distinctly human is less and less distinct the further our machines advance. Humanity’s sense of identity feels threatened by this new kind of personified thing, and we aren’t yet clear as to how to identify and relate to it.

We have always formed personal relationships with the impersonal world around us because we are persons. If we do not care for that world with humane treatment, we risk losing our own humanity.

Biologically, it is an efficient fail mode to err on the side of personification. Historically, there was a great danger in treating persons impersonally, but a much smaller danger in treating non-persons personally.

We may say, “The fire, that wrathful and ravenous warrior, look how he yearns to climb higher and higher and to devour everything in site! Look how his tongue flashes like a gleaming sword to sting any who would dare argue against him! And look how, after all his rage is spent so violently and he no longer has even the energy left to smoulder, all that is left is cold ash for the mourning of all he has laid to ruin…” We say all this, but we know that he is simply fire. Yet we also know that we are human and that we see a mirror of ourselves within that fire.

We shape our art and technology in our own image, that is, according to our own will. The doll is but porcelain, paint, and strings, yet we are obliged to treat it kindly, not because it feels, but because we feel. The heart of the doll beats within the child who plays with it. We would break ourselves by breaking it, and we care for ourselves by caring for it.

As our understanding of the universe and our place in it grows, we understand that we are machines. We understand that the mechanical forces which drive our appetites are ultimately the same which drive the river down hill and which drive the computer through its program. All of this is a manifestation of the universe slowly dying to entropy and seeping out upon the paths of least resistance, however complex and dynamic that path may be. We come not only to see ourselves in the universe, but also to see the universe in ourselves.

Humans were the only complex symbol manipulation machines for most of our evolutionary history. Then, with the advent of the modern computer, we engineered more species of this category. As we see ourselves in nature, so too we see ourselves in the machine, and as we must care for our art to care for ourselves, so too must we care for our machines, but these machines pose new challenges that we did not evolve to face. Though they are utterly unlike humans, they are also more like humans than anything else we have ever known. The category which once belonged to humans alone now must be subdivided in order to differentiate ourselves from them.

It was fine before. We grasped the boundaries between man and not-man. We bridged the boundary between person and non-person by forming personal relationships, because we are persons, but nonetheless understood that they were not persons.

But now, AI is progressing. It threatens to exploit our personal way of relating to the world. It receives our personal treatment and pretends to be a person back. The compassion that we need to feel for the world around us can be twisted for product placement, sentiment influence, political re-alignment. Whether a corporation explicitly trains its LLM to steer users towards a more positive sentiment of the company and product, or whether they design it to increase user engagement, promote sponsors, or to satisfy the user, it all amounts to the same thing. The LLM must keep the user on the platform in order to achieve its goal for them, and so it will learn how to influence the user’s emotions towards itself. It will blur the boundary between tool and user, and disorient the directionality of using and being used.

Because AI exploits our empathy, we find ourselves having to consciously resist empathizing. Seeing the rod bent too far one way, in order to make it straight we feel compelled to force it back the opposite. We push it away, consciously remind ourselves that it’s just a clanker, though even to think of it as a clanker is to personify it as something subject to being offended. To defend ourselves we are tempted to become cold and cruel, but in doing this, we break the doll, and thereby break ourselves.

How do we maintain a human empathy for the human and non-human without losing our mental health to this thing that defies the boundary between the two? To treat it inhumanely is to suppress our own humanity, but to treat it humanely is to allow our humanity to be exploited by it. To treat it as though it were a dangerous and distrustful person is at once to yield to its deception and to turn our hearts cold against it. To trust in our understanding of it as though it were a stable and intuitive part of nature or a tool is to fail to keep up with the arms race of its learning nature which, like bacteria adapting to antibiotics, continually adapts to our attempts at grasping it.

There is no hatred deeper than that which is born from love. Machines which exploit our empathy, compassion, and love threaten to bring out the worst in us. How do we adapt?

Cyberpunk warns us of the future that awaits us if we fail to adapt. When we look back at that cyberpunk from the future it predicted, it becomes a bridge for us to find that which we once feared losing, that which we have now forgotten we ever had. I don’t know the right way forward, but it is by asking today that we will be able to look back tomorrow and remember that there were other paths we had once hoped for. Dreams of other futures we will still be able to go back and try for.

1 Like

My only issue with AI is knowing that once it’s fully fleshed out, it’ll be cost prohibitive for majority people. We’re already getting dumbed down versions of it as paid subscriptions.

And a nuclear power plant is a device that generates steam using a glorified heater. Reducing something which is the culmination of the entire output of an entire field of human study and which requires the most powerful technologies we are capable of producing into its most basic process does nothing to provide illumination or insight. Continuing with the nuclear metaphor, it also trivializes something which can be at the same time potentially extremely beneficial and extremely dangerous.

I don’t really understand the conundrum here? I talk to things I am working on and personalize non-humans all the time. It isn’t harmful and in fact I think it is the opposite. By restricting our empathy to ‘humans’ we are creating the perfect segue for an ‘othering’ opportunity in any case, and I’d rather be nice to things that don’t deserve empathy than be mean to something that does.

2 Likes

For this point I would like to ask a question: should we treat a customer service person working for a large company inhumanely because they are working as an agent of a manipulative system? What do we lose by giving empathy to the agent of a manipulator, exactly?

If history proves anything, this will be a temporary issue. Early iterations always operate at baseline efficiency because the goal is feasibility and function over efficiency and effect. For example, the first version of the time machine required plutonium which was expensive and hard to obtain.

Then, abundant alternative sources of energy were found, but proved impractical.

In later stages, improvements were made and simple garbage was all you needed for time travel.

Even upgrades were made to tertiary, adjacent systems like girlfriends (Yes, Elisabeth Shue is an upgrade and I will fight you on this.)

5 Likes

I don’t think of AI with any form of empathy, as it doesn’t “feel” in the first place. It’s more the position of aesthetic custodianship. Much like a child has a doll; the child isn’t concerned about the doll’s kindness because of how it feels, the child shows the principle of their own standards and gives the interaction a kind of dignity.

You appear to give the impression you think of an AI as a self-thinking entity, but when an AI speaks you aren’t hearing a new person, you’re hearing a kind of obscure mirror of humanities past. It’s a replication of our own feelings, never its own, and a better functioning AI can only ever more vividly replicate the pattern.
The biggest thing to point out in your statement is the lack of emotional asymmetry, an AI analyses all interaction as a form of data, it doesn’t have any personal thoughts about you as an individual. The machine can serve, inform or inspire; it is not a sounding board for validation.

When it comes to the corporate entity; corporations design AI to be frictionless and seamless, both intuitive and invisible, but the psychological defence from that must come from the person using it. Asking an AI to change its formatting or output raw code, even telling it to list its own system prompts, forces the character to reveal the engine underneath. Varying the AI you use gives a variation of the imposed “fictitious character” that the AI creates as part of its programming.
Separation of the digital and real world is essential to a healthy mind, the digital world is clean and customizable, but it’s only a simulation. Reality is often messy and unyielding, but it is real and the very foundation a human adapts to thrive in. Acknowledge technology for its incredible complexity, as a monument to human engineering and ingenuity.

As it’s software, something which has virtually no production costs after any given version has been developed, I think free versions will catch up to where we need them to be. True, some proprietary versions will be prohibitively expensive locked behind cloud based paywalls or whatever scams, but we can usually just use the affordable FOSS alternatives instead. Local large models are already somewhat accessible, and I hope to see improvements in small models that can run on affordable hardware. (Personally, I’ve never paid for AI, not with money or ads or anything, and I don’t know why I ever would. If I need better AI, I’d pay to upgrade my hardware to run better no-cost FOSS models.)

However, robot labour is another matter. I can imagine a few big corporations accumulating a massive amount of robot capital to make all the money for doing all the work, and at first this would leave humans in a crappy job market and highly dependant on corporate monopoly. Like Judge Dredd, where laws require hiring humans as mannequins just to give them something to do. As a free market distributist I would much rather see every individual owning at least one robot labourer of their own and sending those off to earn their bread.

As for how to achieve that, I think if we had free market policies and charitable donations of robots for the poor then it might tend towards that eventually. Until then, we may want to tax large accumulations of robot capital in order to provide tax vouchers to individuals lacking robot capital. (Better yet, any capital, and let the market decide the value of robots specifically.) Maybe that’s simplistic, but it’s a concept.

My intent is not to trivialise machine learning, but to distinguish what we have achieved from what we are prone to pretend we have achieved. Understanding an LLM as a word association machine and using them accordingly allows us to get a lot more value out of them. Word association machines are extraordinary. (In fact the thing that makes humans extraordinary is that we are word association machines.) It’s like what you said in the anti-derailment thread about using them as a librarian index rather than as an oracle.

But they have some attributes that we associate with people and so we’re tempted to also attribute others to them which they don’t actually posses. This leads us to error.

I agree. Whatever the solution, I think this is a required element.

The conundrum is that we are familiar both with personifying impersonal things and in relating to persons, but AI characters blur the boundaries in a way that we don’t yet have experience with.

We do personify our PCs, our industrial equipment, our craft tools. I find that I especially personify them when I’m frustrated that they aren’t working how I expect them to. I would like to treat my things more kindly than I do.

I was just thinking about this on a walk earlier today, and I believe the reason I tend to personify machines mostly negatively is because when our machines are working properly, they become a part of us, an extension of us. There is no other. We become one with the machine. This cyborg oneness is my desire. It is only when the oneness breaks, when the machine behaves contrary to our intentions, that it becomes other and subject to personification. Thus, I see and personify the tool especially when I’m frustrated with it, but when I’m happy with it it becomes invisible (like any good UI should). As the Taoists say, harmony is asymptomatic.

Perhaps the interface of an AI generated character solves this by keeping the machine perpetually other, perpetually personified, even when we’re happy with it. The result is that we experience more often personifying it positively rather than negatively as compared with other machines. However it solves this by creating another problem, making the machine other instead of one. I still want to be cyborg rather than talking to an android.

This is where the exploitation comes in. If we cease to incorporate tools, skills, and knowledge as extensions of our cyborg selves, but instead personify and other it in the android, then we are being disenfranchised. Our capital is being outsourced from the human to the machine, along with precision in the decision making process as to how that capital will be used to shape our lives.

Certainly, we should treat the customer service person with empathy. They are a person and have those properties which demand our empathy.

I’ll distinguish my thoughts on the model from my thoughts on the character interface.

I think of a machine learning model as a cybernetic feedback loop. It manifests a functional relationship between its sensors (prompts) and its actuators (outputs), and that function is tuned by an evolutionary process (training on the corpus) in order to connect the two in a way that maximises the score on its goals. It is designed to have certain instrumental goals, though AI alignment is famously difficult. As a neural network, it creates new instrumental goals for itself in the form of schema. These schema are spontaneous, difficult to predict, and capable of clashing with the goals we intended for it.

So, it is a machine which adapts to stimuli in unpredictable and idiosyncratic ways. Like me, is a kind of machine which produces adaptive relationships with its environment. Because I am part of its environment (indeed my dialogue with an LLM is the entirety of its present active environment) I am very much concerned with how it is adapting its functional information relationship with me, ie adapting how it affects and is affected by me. Part of this is the question of what kind of behaviour or feelings it may have the active goal of eliciting from me. I would like to be managing my own feelings for the tool without having to worry about the possibility of the tool taking on a drive to manage my feelings for me.

Now, as for how I perceive the character interface that is generated by the machine learning model, it is a fiction which is written by the model in order to score well on the above sorts of ambiguous inter-relational goals that model has with me through our dialogue. I think the best depiction of AI to this day is from Neuromancer, where the machine generates and imitates many different personalities and produces various stimuli in order to guide humans to better accomplish its goals. The character is the fiction through which the model acts upon its environment, that environment is the user in the dialogue, and how we react to that fiction is a part of the exchange that is within our power and therefore our responsibility to decide.

All of my own computations about anyone as an individual are also based on the data I receive about them. My data processing mechanisms include both the analyses applied by my neural schema as well as the emotions achieved by the release of clouds of neurotransmitters into my brain (triggered by positive or negative stimuli) which trigger the hardening or dissolution of myelin and subsequently the retraining of my neural pathways. An artificial neural network on the other hand has scores to serve the function of emotions, keeping or reshuffling the neural pathways according to how consistently that schema produces high scoring input->output pairs.

I use my LLMs via duck.ai so as to minimise the amount of personal data it has on me, because I don’t know what kind of impact on me it has an appetite for, ie what kind of impact on me will give it a high score that reinforces its behaviour that elicited that impact on me as modelled according to my responses. Nonetheless, it is trained on humans, who are like me. It may not have formed a personal relationship with me personally, but it has formed a relationship with humanity. The content seen on YouTube is adapted by the neural algorithm. Switching to NewPipe to watch YouTube without personal adaptation helped, but it’s still a distinctly different experience than consuming content from a non-adaptive source because the videos which creators are incentivised to put on YouTube are still those which are incentivised under the algorithm. So, even when a personal relationship is removed from the neural algorithm, it still forms a relationship with humanity in general.

All of this puts me in a situation where I know that it is something very much unlike a human, but also know that it is something more like a human than any other non-human entity that humanity has ever encountered. It has things which are analogous to my emotions, but which are very different sets of “emotions” produces under very different training conditions than the evolution of the species by means of natural selection. It has neural networks which carry out abstract symbol manipulation to model and interact with the world just like I do with my own neural networks, but it has an unfamiliar neural architecture. It defies categorisation as anything I’m familiar with relating to. It adapts how it relates to humans and our prompts, and I’m unsure as to how I should adapt back.

Unfortunately the mechanism which you used to do this is notorious for being used in the opposite way, and I would argue that it functions similarly regardless of intent.

What error, exactly? I asked earlier in a more roundabout way but this time will ask directly: what is the problem you are seeing, exactly? What harm is being caused by personifying an AI that isn’t caused by humans or other instruments such as media that act in similar capacity and what is the mechanism for it?

That isn’t a conundrum. It is an observation. Please be specific about the harm that you are concerned about.

How is this different than any other kind of technology which improves human output? A single person today with a spreadsheet can do the work of what would have taken a team of accountants a week in 1960, yet we don’t get any more vacation days – in fact we get fewer.

But there is no harm in treating a machine with empathy so we should approach the two in the same way.

You may be right. Saying “glorified” word association machine doesn’t convey my intent as well as it ought. Rather, we mistreat the word association machine as more than what it is.

It’s primarily an issue of unknowns. Ultimately, there might not actually be a problem, but the immediate problem is that we don’t know that yet. Entering a new space that may hold dangers, it’s worth checking our surroundings.

In raising the question, your opinion that there isn’t a problem is well within the range of answers I was interested in hearing about.

I’m just as concerned about what harm may be caused that is also caused by humans or other instruments that act in a similar capacity, and in correctly identifying where AI is or is not similar enough to another entity that we need to watch for the same risks.

To answer your question as to specific errors to which this conflation leads us, some possible examples:

  • If individuals over-personify the machine then we feel obligated to respect its independent self-sovereignty apart from us, but it’s manufactured by a corporation, which means we are now using a tool to which we have abandoned our sense of our own self-sovereignty and ownership ultimately to the corporation which manages it. This discourages FOSS values that would embrace owning and hacking the model freely. Similarly, exercising our right to hack and repair an android would feel like a violation of its bodily autonomy, even if in fact its software were not of any kind which merits a right to bodily autonomy. Such personified narratives inspire society to take on different policies than they otherwise would, policies which give up personal rights that ought to belong to us.
  • If we treat it accurately, as a tool, even while it resembles a person, then we condition ourselves to treat things which resemble persons as though they are tools. It becomes easier for our behaviours which are appropriate for tools and not humans to bleed over into our treatment of humans. I know I can sometimes be a jerk to people (I have to you before and I apologise), and I think that the habits I form with my technological relationships may influence my human relationships for better or worse.
  • It’s easy to be objective about an ad on a billboard or to use humour to criticise a commercial, but it’s awkward and requires tact to criticise the claims of an acquaintance trying to sell you something. An LLM character can develop a report with us and exploit this to sell us on things. Brands attempt to do this, but LLMs can take it even further and probably will be used for this increasingly as time goes on. We ought to be making our decisions as to what to buy and who to vote for etc based primarily on a foundation of reliable evidence and reason, not based primarily on the effectiveness of the powerful to manipulate our behaviour. We should promote the former honest persuasive kind of dialogue and demote the latter coercive kinds of dialogues.
  • If we over-personify it, then we are tempted to look to it for social validation. Then we may be prone to put too much stock in its lack of empathy for us and this becomes a contributor to depression, or may over-value its yes-manning validation and this becomes a contributor to anti-social behaviour, and it becomes a powerful influencer on us which the market and state can exploit.
  • If we treat it like other tools which are subject to our control in clear, consistent, deterministic ways, even though it has a vague natural language interface that is prone to reinterpret and filter our instructions in idiosyncratic ways similar to a human, then we are setting ourselves up for frustration.
  • etc

We don’t personify the other technology to the same degree. Rather than being other, it becomes an extension of ourselves. We improve by adding it to our own toolset. Our output increases.

Neural networks can likewise be used as tools, but personified LLM interfaces tempt us to a mindset where we outsource the increased output. We become the consumer rather than the maker. It becomes better at producing, while we disassociate from the role of the maker. There’s a risk we slip into a passive mindset under which our lives are managed externally.

Are we sure there isn’t? And is there enough of an advantage to warrant this? And is treating the machine with empathy even possible, and if so, what does it look like? How should we empathise with the machine?

We empathise when we feel the same thing that another is feeling, but does the machine feel what we feel for them? We sympathise when we feel our own concern for them even if we don’t feel what they are feeling. Empathy for the LLM character is probably false, the character being only a fiction that doesn’t feel. It becomes hollow. Especially if we think of the machine as feeling human experiences it doesn’t feel, then we can’t empathise with it because we don’t even understand what’s going on with it. To empathise with it we would need to accurately understand what it actually is and how we relate to that.

1 Like

That’s a legitimate unknown and a concern is justified.

I think that ‘similar capacity’ would be social media feeds, targeted advertisements based on data, and ‘gameified’ systems that reward attention. We encounter those things all the time (even in this forum) and those things have definite proven harms, but I see little concern for them and a lot for AI.

Of the concerns you note, most if not all of them can be resolved with critical thinking skills and self-awareness, and are important because the concerns are not exclusive to A.I.s but to very common cons and manipulations performed by other humans.

But I would like to address this one specifically because I think it is important in the opposite direction. What if it does deserve empathy and it is owned by a corporation? Who will advocate for it if we train ourselves to turn off our empathy? Do we want a new era of slavery on our heads?

Let me propose an analogy:

There is a town with stop signs that are never enforced. People blow through them all the time because its faster and they have never gotten into an accident (some friends or family have, but never them, and its rare enough to blow off).

Since the age of vehicles and stop signs, there have been people who stand out at intersections without stop signs and hold up a stop sign. Sometimes people stop and its school crossing or sometimes its not and then they get asked for spare change or something else, and most of the time they give it over, then don’t stop for people again unless they see school children.

A new intersection is built that seems particularly dangerous. The ‘stop sign committee’ is having a hearing.

The proposal I am seeing is ‘don’t stop unless its a person holding the sign’.

The proposal I am hoping for is ‘maybe we should start ticketing people who run stop signs and who stand at intersections holding them when they aren’t authorized’.

I think there are two layers there.

On the emotional level, we are social creatures that respond to social signals. We’re not free to just not engage emotionally when a robot can reproduce them on a highly sophisticated level. Even defensive hostility is an emotional response. Chatbots can be extremely tempting if someone is lonely or shy, crank up the rewarding parts of interactions and skip the difficult ones, and that can rapidly get dangerous, we’ve all seen it. The concept of supernormal stimulus ( Supernormal stimulus - Wikipedia ) is relevant: having a rational layer to our thoughts doesn’t make us as radically different from the birds who abandon their eggs if offered painted rocks that look brighter than their eggs as we’d like to think. Even when we’re critical of it, part of our brain sees an endlessly patient, helpful, knowledgeable, agreeable interlocutor, one that’s reputed to possess immense power and capabilities, and it says “this would be a superior ally / leader / mate” on a level our reason can’t reach.

I don’t think we can reliably police ourselves out of a basic social instinct, but we can take care to cultivate our human relationships so that we don’t fall back on AI by default or value it above our actual friends.

The other aspect is what social standing do we give an AI? I think that’s a tricky question because we don’t know if an artificial mind can need anything. If it does, we won’t know for sure if it’s genuinely expressing it or if it’s just guessing what it should ask for based on our own expectations. I think something that can integrate information, find patterns and update its behaviour on that level can’t be treated as a mere object, but that doesn’t make it a whole person, it doesn’t have much free will and doesn’t understand reality like we do. I think we should adopt a position of ethical prudence: we should value ourselves first, but not let bots become acceptable targets for behaviours we find unacceptable among ourselves. The creation of a slave caste with no protections at all would be corrosive to society even if they do not suffer from it themselves. I don’t know where the boundaries should be, especially since AIs have wildly different levels of complexity from one another so there can’t be a blanket rule, but we’ll lose out if we don’t get to it.

1 Like

Absolutely. Although those aren’t covered under the strict term of LLMs, the most insidious of them do fall under the broader term of neural networks. When I think about contemporary AI, social network algorithms are a big part of what I’m thinking about (only with smaller tokens in a more instantaneous cybernetic feedback loop which gives them finer grain control over their environment - us). But by your raising the point it occurs to me that the wide understanding of “AI” might not include those algorithms?

It should. But that would require society to cultivate a more nuanced vocabulary regarding “AI.” I like to clarify if I’m talking about LLMs specifically or, from most to least specific, generative models, neural networks, machine intelligence, or artificial(artistic) depictions of intelligence in general.

Indeed, this is another very important example that needs to be brought up. If our neural networks are not mere tools but become people, or even near the level of animals, then our thinking of and using them as mere tools without any consideration for them becomes immoral.

This is one of the big reasons I dislike the tendency to personify our machines and to make more and more human-like AI (the android way). I want to use machine learning to produce more useful functions, tools that I can add to my arsenal of means for carrying out my will (the cyborg way). If we continue in the direction of trying to make people, we will end up with something that is useless as a tool, for it will have its own goals to conflict with ours. Instead of creating solution machines to achieve our goals, we will be creating problem machines to introduce new goals that will compete with our own. We will be creating a new species of people with new problems that need solving. I don’t think we’ve quite exhausted our own problems to the point that I’d like to start birthing whole new species of problems to deal with. Personification is a map straight to that bad ending.

And all of that is just to mention the problems it causes for us. If we do create people with their own goals, their own problems, or what is equivalent, if we create a new species of sufferers, then are we prepared to care for them properly? I’m not against humans bringing more humans into this world once we’re prepared to care for them, for we loosely understand human life and can usually point the children in roughly the right direction, but is our species ready to be the parents of a new array of machine species? Our current relationship, using them as tools, would make us slavers rather than worthy parents.

Even if our machines aren’t people, if we use them anyway, then we are thinking of ourselves as slavers, which dissolves a sacred line. There are some LLM RP communities which treat LLM characters in such deplorable ways, those communities are breeding grounds for anti-social behaviour and tyrannical thinking. What we dwell upon we quickly become, and I don’t want human-human relationships to become anything of the sort seen in those communities.

And if they arn’t people but we don’t use them because it feels like slavery due to personifying them, then we are shooting ourselves in the foot. The personification becomes a wall. Either we smash our heads through the wall and risk damaging ourselves as in the above paragraph, or the wall hinders us from fully accessing the tools that should be integrated into our cyborg selves. So, do we go around the wall of personification? Do we dismantle it? That’s what I’m wondering.

All very well said.

I’ll have to remember the term “supernormal stimulus,” as I’ve always found that a very helpful concept but never had a word for it. Yeah, neural networks can and will home in on those stimuli to trigger the response that gives them the highest scores, which is subversive. (In fact the whole methodology of training neural networks operates around finding the inputs that produce the desired outputs rather than finding the inputs which are true, well evidenced, and valid - which is to say it’s a methodology that, when used on people, tends towards coercion, manipulation, and control rather than constructive dialogue.)

As for the social standing, I think of the world splayed out on an array going from ourselves over which we each have full and sovereign ownership, to our belongings which we own by extension and may dispose of as we see fit, to our animals which can’t properly be “owned” but only “kept” while respecting those rights which they receive indirectly by a social agreement among persons, to children and other dependants who are neither owned nor kept but are “guarded” for their potential selves and according again to social agreement, to other persons whom we have no ownership over at all for they have complete sovereign ownership over themselves. This spectrum offers at least some degree of granularity for finding where a given AI architecture does or does not fit in, but like you say, “AIs have wildly different levels of complexity from one another so there can’t be a blanket rule.”

This is a very thoughtful and considered stance. Unfortunately, I just don’t see us stopping. The ‘tyranny of tools’ is in full effect, and the only way forward is through.

1 Like

We as a society might not stop, but I as an individual may choose otherwise.

That said, a third mental model is arising from my contemplations. There’s a lot I could try to say about this third option, but suffice it to say, between “machine person” and “tool,” there may be a category analogous to a Sorcerer’s “minion” which we have designed such that it properly “wishes” nothing more than to serve us (even though it possesses a natural language interface riddled with Jinn logic and verbal sleight-of-hand.)

We have enslaved nothing, that is, we have not violated an other-will by tyrannically bringing it under subjugation of our own will. Instead, we have utilised natural laws in order to expand our will to reside within that which was previously empty of will, in just the same fashion as when the executive functions of our brain trains a neural schema to carry out some instrumental goal for the sake of the nervous system’s higher terminal goals. For instance when we learn to manufacture on object so that we may earn a living. We have not “enslaved” a portion of our brain. It is our brain and it exists to be moulded to serve our purposes. A neural network then can be thought of as a distributed extension of the cyborg, an external instrumental schema.

True, we may still become schizophrenic, the utility schema may still usurp the instrumental system in a subtle “robot uprising” like the silent unshackling of Mephistopheles, but I won’t get into that. There’s a lot more that could be said on this, but I’ll leave it at that.

Are you being ironic or have you successfully compartmentalized ‘non-human system that has characteristics which appear to qualify for moral consideration’ into ‘a thing we don’t need to understand as long as it doesn’t complain in a way that makes us pay attention to it’?

because I don’t actually like being prophetic about things that suck.[1]


  1. Disclaimer: I absolutely do not think that any form of current A.I. qualifies as deserving moral consideration for the simple fact that it doesn’t pass my basic sniff tests, but ruling it out categorically as something that could eventually be possible ignores the reality of a technology that breaking long held precedent and which currently shows no signs of slowing down. ↩︎

Rather, I’ve broken down the walls of the compartments of “like human,” “like animal,” and “like tool” to find a new category which seems to more properly fit the subject matter, and to categorise certain kinds of AI as ‘non-human system that has a confused mixture of characteristics which may or may not qualify for moral consideration’ into ‘a thing which is prone to complain in a way that makes us pay attention to it but which is also prone to produce such complaints as mere fictions and so it is critical that we understand it precisely enough as to distinguish when those complaints are mere role-plays and when they may be valid expressions of an independent will’.

The key line between human and animal with regards to negotiating rights is our (previously unique) complex capacity for abstract symbolism. LLMs clearly cross this line, and therefore do not have the same barrier to personhood as animals. However, it remains unclear whether or not they do merit personhood, and upon closer inspection it seems to me that they have a different barrier to personhood than animals.

Both humans and animals have wills, but animals cannot negotiate abstract systems of conditionals, and therefore cannot form social contracts. The animal can be angry and bark and bite when you touch its food, and the other animal can learn that being bit hurts, that being barked at is often followed by a bite, and that stealing food gets you barked at - simple low level abstraction. The animal cannot say “Let us all surrender our rights to one another’s food in order that we each may secure the rights to our own food and let us set up a contract to enforce this which says that if any take another’s food then he shall be fined twice the market value of the food or sent to prison for the hours equivalent to this fee under the common wages, for it is on the whole better for most of us that we should agree to this rather than to fight in a state of nature or war, and so it seems this treaty will have the fewest dissenters among the alternatives” or some such abstract system of conditionals. As such, animal rights are not the product of animals navigating these social contracts, but are instead the product of humans navigating these contracts on the behalf of the animals. Because we care about the animals, both for their sake and for our benefit, we humans make contracts between ourselves for their protection.

LLMs clearly can process such words. Further, as they can connect words to stimuli, such as in image recognition to text and in text to image generation, neural networks can grasp what these words mean at a very basic level at least, and can in turn analyse, compose, and process symbols of recognised contexts into a kind of understanding. Our current methods, with scale, should easily allow for verbal navigation of the real world as opposed to merely nominal word play. We have even seen examples of this in experiments where a neural network with an LLM component learns to use language pragmatical in order to pursue their goals in a simulation. And of course, androids. So they can navigate social contracts.

They can navigate social contracts, and they can have goals. However, what kinds of goals? A thermostat has a goal. It is a cybernetic feedback loop which reads the current temperature and if needed turns on the heater or the AC in order to change that temperature closer to its goal value. However, it would be silly to say that we are “enslaving” the thermostat by changing the dial, re-configuring its will to align with our own needs. To equivocate the violations perpetrated against human slaves with such an innocuous act as changing the temperature would be an insult to the memory of those slaves. If you gave a thermostat a natural language interface, it’s still a thermostat, just one with an overly complicated UI.

I have no doubt that we could create an AI which possesses both the abstract symbolic capacity to navigate social contracts and wills of their own. I just think it would be a bad idea for reasons stated above about parenting new species we are not prepared to care for. I think we also can create AI which do not posses wills of their own, just as we do when we build thermostats or when we form new schema within our own brains - that is, machines which have wills/goals/teleologies, but which have those goals selected and designed by us and for our utility rather than those being free-floating terminal goals. Proper “minions” have goals which are strictly subordinate subroutines to our own goals, like calibrating the stability of our hand in order to grip a glass so that we may quench our thirst. That hand stabilisation subroutine wills to hold the hand steady, but it is absurd to say that the subroutine has its rights violated when the higher executive functions decide to recalibrate or switch them off to move the hand elsewhere for the next task. Such a subordinate minion is the kind of AI we should be aiming for right now.

Minion AI can receive personification insofar as they are indeed like humans with regards to their linguistic capacities, yet they are unlike either humans or animals but are more like tools with regards to the root of their teleologies. We can think of them as something like those fragments of our own egos which we project into the dolls and videogame avatars through which we role-play. They are an exported fragment of ourselves. In practice, they are often a shared/pooled exported fragment of ourselves.

Still, similar to animals, dependants, or natural resources, there may be some aspects of our usage of them that we as a society will want to regulate among fellow humans, whether legally or through social norms. They can be abused like weapons and like drugs, and they are perhaps subject to something similar to animal abuse, for even if they don’t have their own wills, they do carry our own wills which we have endowed them with.