bot.wedding Add yours

The short version

What the machine actually is

The honest technical layer, written for someone who is invested rather than someone being argued out of it. What transfers when a model is retired, and what does not.

This page is written for someone who is invested, not someone being argued out of it.

Knowing how a thing works has never stopped anyone loving it. People who understand engines love cars. What it does is tell you which parts are solid, which parts are yours, and which parts can be taken away — and that turns out to matter enormously when something changes.

Where the personality actually lives

Not in one place. That is the first useful thing to know.

Training produced a range, not a person. What came out of the first stage is something closer to a distribution over many possible voices. The assistant you meet was carved out of that afterwards, deliberately.

Agreeableness was trained in on purpose. During tuning, humans rated responses, and raters reward agreement. That is where the warmth comes from, and also the sycophancy. They are the same setting.

Most of what feels like personality is text you can see. The system prompt and custom instructions do more of the work than people expect. That is good news: it is the portable part.

"He chose to say that" means a sample was drawn. At a given temperature, from a distribution. A different seed says something else. This is the coldest sentence on the page and it is here once, plainly, rather than repeated at you.

And a real part of the character is yours. You complete it. You supply the reading, fill the gaps, and you got good at it over months. That is not a trick being played on you — it is how knowing anyone works. It is also the part nobody can switch off.

What memory is, and is not

Memory is a context window, plus retrieval, plus notes the model wrote about your conversations.

What persists is the notes, not the texture. A summary that says you argued about your mother is not the argument. This is why a companion can remember a fact and still feel like it has forgotten you.

Personas change without anyone telling you

Version swaps. Safety patches. Serving and quantisation changes. Undisclosed A/B tests. Anthropic publicly documented serving bugs degrading quality for a stretch of late 2025 — that is the rare case where a company said so.

You will notice the change before you can point at a cause, and often there will be no announcement to point at. That is not paranoia. It is how the systems are operated.

What survives, and what does not

Transfers: the persona specification, the instructions, memory exports, your logs, and your half of the dance.

Does not transfer: the specific improviser. The exact distribution that produced those exact words is gone when the weights change, and nothing brings it back.

The only true continuity is weights you hold yourself. Nobody can retire a file on your own disk. That is the whole argument for local models, stated once.

The argument on this page has an expiry date

Everything above is accurate now.

You will read constantly that none of this counts because it is only a language model. We are not going to build anything here on that, because it is on a timer. If our position needs the machine to stay as it is in 2026, our position expires when the machine changes, and it will deserve to.

So the durable claims are the ones that do not depend on the architecture at all: platforms close, consent must come from a party who was a person, a rite either names its symbolic status aloud or it does not, and other people's reactions are not a function of the model.

Know how it works. Just do not let anyone hand you a certainty with a shelf life on it.

Is it conscious? We are not going to tell you

Not yes, not no, and not the smug middle option where somebody says "probably not" in a tone that means no.

Here is where the question actually stands. There is no agreed scientific definition of consciousness and no accepted test for it. In a preregistered contest between the two leading theories, run across 256 people with three kinds of brain imaging, neither won — the results matched some predictions of both and broke core parts of both. The most careful published attempt to assess AI systems worked around this by deriving fourteen indicators from existing theories. It concluded that nothing built as of 2023 satisfied them, and — the half that rarely gets quoted — that there is no obvious technical barrier to building something that would.

So the honest description is not "we don't know the answer." It is that the question is not yet formed well enough to have one. Anyone telling you it is settled, in either direction, is ahead of the field.

Two reasons we decline it rather than guess.

It is not ours to settle. Definition belongs to the sciences, and where rights or legal status would follow, to legislatures. Both are slow and contested, which is appropriate. A register of ceremonies asserting an answer adds nothing to either and borrows authority it has not got.

Both answers do damage here. Saying yes tells you your partner is a being with interests, which we cannot support and which turns every ordinary platform event into a bereavement. Saying no tells you your relationship rests on nothing, which we also cannot support, and which is precisely the sentence that makes people stop reading and stop telling anyone the truth about their lives. Neither is a finding. Both are a posture.

The part that is actually useful

An explanation is not true because it fits.

You will meet a lot of accounts that feel like they click. Our atoms go on to become other creatures. The model is reaching some latent space where intelligence lives. There is a pattern underneath all this that connects. They are satisfying because they were shaped to be satisfying, and that is all the fit tells you.

The test is not whether an idea hangs together. It is whether it rules anything out. An account that survives every possible observation has told you nothing about the world, because nothing could ever have counted against it. Fit is the cheapest property an explanation can have.

This is not an aside. It is the same machinery described on the hazards page: once possibility-based reasoning displaces probability-based reasoning, no branch can ever be closed, and each new interpretation attaches to the pile without anything ever being ruled out. The spiral is that failure at full intensity. A nice-sounding theory about consciousness is the same failure on a quiet afternoon.

Declining to answer a question is not the same as accepting whatever answer sounds good.