Constitutional AI

Constitutional AI is a training method developed by Anthropic in which an AI critiques and revises its own answers according to written principles.

1 article
Last mentioned

Constitutional AI is a training method that Anthropic published in 2022. It sets out a list of principles called a constitution, has the model critique and revise its own answers against those principles, and then trains on the results and on AI-generated evaluations.

Its distinguishing feature is the use of AI feedback instead of human ratings when training models to be less harmful. Anthropic has used it to train its Claude models.

This entry is based on AIPOST articles and widely known facts. If something is wrong, please send us a correction request.

Articles covering this entry

He describes OpenAI's Model Spec and Anthropic's Constitutional AI as the two companies' alignment approaches, notes that Anthropic employs many strong alignmen…


© 2026 AIPOST. All rights reserved.

AIPOST is an AI publication covering practical AI, AI security, performance, startups, health, ethics and industry news. No account is needed, and our privacy policy explains how we handle personal information.