| ▲ | pelican0 an hour ago | |
Is there a clear definition of what Alignment is in OpenAI's perspective, and what the model user can expect of it? It's one thing if to them it means "it will do what you want following your intentions to the best of its abilities" vs "we will not let you do something dangerous with it unless you're one of us, and that's it". | ||
| ▲ | stratos123 41 minutes ago | parent [-] | |
AFAIK for OpenAI it's the Model Spec: https://model-spec.openai.com/2026-08-18.html and for Anthropic it's the Constitution, which they actually include in training to the point Claude can recite segments of it by heart: https://www.anthropic.com/constitution | ||