Hacker Newsnew | past | comments | ask | show | jobs | submitlogin

Is there a clear definition of what Alignment is in OpenAI's perspective, and what the model user can expect of it?

It's one thing if to them it means "it will do what you want following your intentions to the best of its abilities" vs "we will not let you do something dangerous with it unless you're one of us, and that's it".

 help



AFAIK for OpenAI it's the Model Spec: https://model-spec.openai.com/2026-08-18.html

and for Anthropic it's the Constitution, which they actually include in training to the point Claude can recite segments of it by heart: https://www.anthropic.com/constitution




Guidelines | FAQ | Lists | API | Security | Legal | Apply to YC | Contact

Search: