logoalt Hacker News

pelican0yesterday at 8:33 PM1 replyview on HN

Is there a clear definition of what Alignment is in OpenAI's perspective, and what the model user can expect of it?

It's one thing if to them it means "it will do what you want following your intentions to the best of its abilities" vs "we will not let you do something dangerous with it unless you're one of us, and that's it".


Replies

stratos123yesterday at 9:15 PM

AFAIK for OpenAI it's the Model Spec: https://model-spec.openai.com/2026-08-18.html

and for Anthropic it's the Constitution, which they actually include in training to the point Claude can recite segments of it by heart: https://www.anthropic.com/constitution