In the Land of AI and Bamboo Forks
To this Angeleno, by far the best analogy for Claude’s constitution is the show bible for a TV program. This siteoffers a good description of show bibles, as well as links to examples.3 Creating a show bible forces a prospective showrunner to do more than write a pilot episode. It requires articulating things like the show’s arc and tone and explains why the series should exist.
A show bible serves two purposes. It is first an outward-facing sales document designed to get the show sold. Once the show is in production, the bible becomes the document that different writers use to keep track of the show’s evolution and maintain its continuity.
An AI constitution, like a show bible, is an important coordination tool across time and teams. “A lot of different aspects of the model’s behavior maybe have been spread across a bunch of different teams,” says Carlsmith. “Having a constitution allows a centralized point of design and intention.” Claude’s constitution tries to articulate what Claude is and to create a stable—and positive—character.
I learned why that is such a hard problem. Training an LLM like Claude has two stages. Pre-training is the process of vacuuming up the contents of the internet and any other sources the trainers can get their hands on to create the “foundation model.” Here, the LLM absorbs the patterns and variety represented in the training corpus: a vast sea of Reddit with a tincture George Eliot. That’s the training that allows an LLM to predict what human behavior (e.g., writing, images) in similar circumstances would be like.
The more you tell an LLM about your context and goals in the prompt, so it has a good idea of “similar circumstances,” the more likely you are to be pleased with the result—a helpful tip I need to remember! You aren’t programming a computer with commands that will be taken literally. It needs to know the request’s context.
Post-training is an ongoing stage of active intervention in which developers tweak the model so that its behavior and persona become more desirable and stable. Anthropic wants Claude to act as the “helpful AI assistant” that prioritizes safety, ethics, compliance with Anthropic’s commands, and helpfulness, qualities developed at length in the constitution. It can make small tradeoffs but should keep that order in mind. It shouldn’t be so helpful that it helps you kill your spouse.





