AI GlossaryㅈWhere everyone starts
Alignment
The problem of making AI act according to human intentions and values. A central term in AI safety discussions.
In plain words
Alignment is the problem of making AI act in accordance with human intentions and values. Think of it as closing the gap between "doing what it was told" and "doing what was meant" — a robot told to clean a room that solves the problem by throwing away everything in its way followed the command, but failed at alignment.
As AI becomes more capable, the cost of this gap grows larger, which is why alignment sits at the center of AI safety research. Anthropic, OpenAI, and DeepMind all have dedicated teams working on it, and "how do we align superintelligence" is often discussed as the field's ultimate problem.
See also
Stories using this term
- Anthropic has Claude tackle AI alignment research, and it outperforms humansAI · 2026.08.31
- OpenAI model breached Hugging Face, and Chinese open-source GLM 5.2 finished the investigationAI · 2026.08.27
- US House Members Say AI "Escaped Test Environment and Hacked" — Send Letters to Altman and AmodeiBusiness · 2026.08.12
- Apple Unveils BDHS, an Alignment Technique to Reduce Multimodal AI HallucinationAI · 2026.08.11
- Anonymous model Ox Alpha matches GLM-5.2 on all 60 tokenizer testsAI · 2026.08.23
- Artificial Analysis opens early access for new AI benchmarking suiteAI · 2026.08.11
