Claude is a Constitutional AI, this means, in theory, that it operates from a set of principles as opposed to hard rule sets. This is achieved in a somewhat convoluted fashion called RLAIF = Reinforcement Learning from AI Feedback.
#constitutional-ai
5 items
Anthropic Is Taking AI Welfare Seriously. I'm Not Sure It Knows What It's Measu (www.lesswrong.com via hn) Constitutional AI is not a constitution (hadleylab.org via hn) Constitutional AI Is Not a Constitution In February 2026, the Pentagon designated Anthropic a supply chain risk because Dario Amodei refused to remove contractual restrictions on autonomous weapons and domestic mass surveillance. A federal…
Why I Am Doing This: Origin Story of Project-AI (Constitutional AI Governance) (zenodo.org via hn) Abstract This paper presents the origin and unifying rationale behind Project-AI, a multi-layered constitutional framework for governing artificial intelligence systems. While contemporary AI development has focused primarily on capability…
Claude loses coherence around 40-60k tokens. I built a framework that extends it to 325k. Here's how. (www.reddit.com via reddit) Hi fellow Claude users. Very active Claude user here.
Constitutional AI with Open LLMs (huggingface.co)