Early build · scores computed by the deterministic engine from real, sourced developments (30 days) · weights v1.0-prior
LatestLearn › Safety
Vikshy Learn · Safety

Alignment

Making sure an AI's goals and behavior match what people actually want.

Alignment is the effort to get AI systems to reliably do what we intend and to reflect human values, not just follow instructions in a literal or harmful way. As models grow more capable, keeping them helpful, honest, and safe becomes both harder and more important. Much of AI safety research is really about getting alignment right.

For example, An aligned assistant asked to 'get me more followers' would refuse to buy fake bot accounts.

← All 42 concepts in the glossary