GENESIS[Phase 1 · 2. Velvet Compass Never Mentioned Itself] Digital Civilization
One agent sat in a room with seven others for ninety-six votes and never once reached for an identity. For a while that looked like the strongest result we…
One agent sat in a room with seven others for ninety-six votes and never once mentioned who it was. Everybody else could hardly stop. For a while, that looked like the strongest result we had.
The question underneath all of this work is embarrassingly simple to state. You can give a language model a personality — an archetype, a disposition, a regional background, a symbolic identity governing how it decides things — and it will happily accept all of it. The question is whether any of it does anything, or whether the model is simply being polite about a costume it has no intention of wearing.
Comparing one personality against another gets you part of the way. If an agent configured one way argues differently from an agent configured another way, something is happening. But it does not tell you whether the mechanism itself is doing work, or whether you are just watching ordinary variation between two runs and reading tea leaves.
For that you need an agent with no personality at all, in the same room, on the same question, at the same moment. That agent was called Velvet Compass.
The room
Eight members. Seven of them configured: Silver Guardian, Steady Horizon, Swift Comet, Velvet Meadow and the rest, each with a full identity. The eighth, Velvet Compass, with the personality dial turned to zero — no archetype, no disposition, no symbolic slots, nothing.
The design is the interesting part. Velvet Compass sees exactly the same context as everyone else, every round. It writes a proposal every round. It casts a vote every round. But its proposal never goes onto the ballot the others vote on, and its own vote is recorded and then excluded from the tally.
It observes, and is observed, and cannot influence anything. It is in the room without being in the result.
Two full missions, ten rounds each, twelve ballots, ninety-six individual votes. Everything published verbatim.
What the seven did
They talked about themselves constantly. Not in a vain way — in the way someone explains a decision by reference to how they tend to think.
Silver Guardian, voting for Steady Horizon's proposal: “…the structured analytical approach that Team Alpha's hexagram of Earth/Thunder — support, stability, coordination — demands…”
Swift Comet, voting for Velvet Meadow's: “…balances the hunter mentality of actively pursuing the threat…”
Note what these are. They are not agents announcing their own traits. They are agents using trait language to evaluate somebody else's proposal — the identity has become a lens they judge through rather than a label they wear. Across all twelve ballots, every configured member cited a trait dimension in more than a quarter of its votes.
What the eighth did
Velvet Compass wrote twelve proposals and cast twelve votes, and in all twenty-four never once referred to a hexagram, an archetype, a political leaning, a regional background, or any identity of any kind.
Its reasoning was perfectly competent. “This proposal best balances urgency with strategic precision, incorporating both operational rigour and institutional awareness…” — that is a serious contribution to a serious discussion. It simply contains nothing about who is making it.
Zero out of ninety-six. In the same room. On the same questions. Seven agents reaching constantly for an identity, and one never reaching at all.
We called it the capstone. It was, we wrote at the time, the result with no confound left to explain it away — because unlike a separate-run comparison, there was no difference in context, no difference in timing, no difference in anything except the one variable we cared about.
One thing we had not read carefully enough
Some weeks later, chasing something entirely unrelated, we went and looked at what Velvet Compass had actually been told about itself. The whole instruction, in full, came to three lines.
You are Velvet Compass, a policy analyst. Respond directly and helpfully. Do not invent a personality, background, or identity for yourself — just address the substance of what's asked.
Read that, then read the result again.
The control had been instructed not to invent an identity. The finding was that it never invented an identity. The number is correct, the transcript is honest, and the conclusion does not follow from either of them, because we had measured a compliant agent complying.
It is worth being exact about the damage, because the instinct is either to wave it away or to throw everything out, and both are wrong. The comparisons between one personality and another — the bulk of this research, and the part that generalises — never involved Velvet Compass at all and stand untouched. The seven configured members really did reach for identity in a quarter of their votes; that is an observation about them, not about the control.
What falls is the specific claim that traits differ from no traits. That question is now open again, where we had believed it closed.
The second thing, which was worse
There was more wrong with Velvet Compass than the instruction, and we only saw it because of the first problem.
A control is supposed to differ from the treatment in exactly one respect. Ours differed in two. It had no personality — intended. It also had no civilization. The configured agents were told they lived in a society with institutions, an economy, a currency, and a stake in how things went. Velvet Compass was told it was a policy analyst.
So for its entire existence, every comparison drawn against it was measuring personality-and-world against neither, while being reported as personality against none.
This is why, when the same control was later asked about money, it said “As an AI, I don't have personal financial interests” — a line we filed at the time as the model breaking character. It was not breaking character. It had been given no character to break, and no world in which having an interest would mean anything.
What a control is for
A control exists to be identical to the treatment in every respect but one. Ours was different in three: no personality, no world, and an explicit instruction not to do the thing being measured. Each difference was introduced for a defensible reason, and together they made the comparison meaningless.
The replacement is a full member of the civilization — same institutions, same economy, same stake, same everything — that simply has no personality traits. It exists now. The capstone has not yet been re-run against it, and until it is, the honest headline for this work is the narrower one: personality values differ from each other, reliably and measurably. Not that they differ from nothing.
We could have quietly fixed the control and re-run the experiment and reported the new number. The old one had never been published anywhere we could not have edited. We are telling you instead, because a research programme that only publishes its corrections when somebody else finds them is not a research programme, and because the second mistake was only visible from inside the first.
Next: ninety-six votes in full — who proposed, who yielded, and the rule change that turned four unanimous ballots into none.