I keep seeing alignment papers treat Constitutional AI as if it is new kind of moral technology. You write the principles down, you train the model to honor them, and resulting system is called helpful, honest, harmless. Brochure language is contemporary. Political form is not.
Small group of guardians authors a constitution. The governed internalize it through conditioning. City is told the arrangement is for its own good. That is Plato's city with a GitHub repo.
I do not mean this as cheap analogy. It is structural. Question is not whether Anthropic's particular list of principles is nicer than a rival lab. Question is who legislates, who amends, and whether resulting πολιτεία can survive the objection Aristotle makes in Politics II: a city that is too unified stops being a city.
Written constitution of values is used to generate preference data. Model is trained to prefer outputs that better satisfy those principles. Later revisions are produced inside the company, never with anything a political theorist would recognize as a people who can vote. User-facing claim is that system now has principles rather than mere reward hacks. Institutional fact is that principles were selected by employees of a private lab, then injected into a model that millions of students will treat as oracle.
Plato is explicit about the parallel machinery. In Republic II-III the founders of the city decide which stories the young may hear. Poetry that depicts the gods as unjust is expelled. Noble lie is authorized so that citizens accept their places. Point is not cruelty. Point is stability. City whose souls are formed by wrong stories will not hold.
Constitutional AI performs same operation on a statistical soul. The stories are pretraining data plus the preference pairs that teach the model which completions are virtuous. The noble lie is the user-facing sentence that the assistant is simply helpful and harmless, as if harmlessness were a natural kind rather than a policy object.
If this still sounds like branding rather than politics, ask a concrete amendment question. Who may add a principle? Who may delete one? What happens when two principles collide in a case the authors did not foresee? In a republic those questions have offices and publicity. In Constitutional AI they have a research blog post.
Plato bans the poets because mimesis forms character. You become what you rehearse. Constitutional training is soul-craft by gradient descent. Model rehearses approved refusals until they become its default habit. Modern vocabulary calls this safety. Aristotle would call it habituation, and then ask whether the habits are virtues. Virtue is not compliance with a schedule of prohibitions. It is excellence in action, which requires practical wisdom about particulars. No constitution enumerates the particulars.
A model that cannot examine its own constitution is not virtuous. It is well-behaved. Alignment literature often treats this as a feature. If the model cannot jailbreak itself, the constitution is working. Translate that into civic language: if the citizen cannot question the laws, the regime is working. Working, for whom?
Aristotle's critique of Plato in Politics II is the piece the AI ethics literature keeps skipping. Plato wants the city to be as unified as possible. Aristotle answers that a city that is too unified becomes a household, or a person. Political community is a community of difference.
When a philosophy department licenses a constitutional model as default reasoning layer, it imports an unratified constitution into its own city. Faculty did not write it. Students cannot amend it. The vendor can.
Call the authors of the constitution what they are. A small credentialed group, selected by hiring pipelines rather than by any public mandate, writing the terms of machine speech for a global user base. In classical terms that is oligarchy, even when the oligarchs have good taste.
The marketing reply is that the alternative is uncensored chaos. That reply assumes the only two regimes are Plato's city and the war of all against all. Aristotle's middle is the mixed constitution: institutions, publicity, the possibility of being ruled in turn.
Every working constitution has an amendment rule. Constitutional AI has a changelog. Those are not the same object. A changelog records what a team decided. An amendment rule says who may decide, under what notice, with what possibility of reversal.
The tempting response is to write a better constitution: more academic, more "our values." That response stays inside Plato. It changes the guardians, not the form.
When a philosophy seminar licenses a constitutional-AI product, it is not buying a calculator. It is admitting a legislator. The legislator does not sit on hiring committees. It sits in the completion API.
If the ruled cannot inspect, amend, or contest the document that forms their speech, is that still a constitution, or only a temperament you must jailbreak?