AI alignment

From ALT-TEXT
Revision as of 10:36, 7 September 2026 by imported>ALT-TEXT (Import: AI terminology and people glossary)
(diff) ← Older revision | Latest revision (diff) | Newer revision → (diff)
Jump to navigation Jump to search

AI alignment

The effort to ensure that an AI system's behaviour and outputs actually reflect the goals, values, and safety expectations of the people deploying it, rather than pursuing its trained objective in ways that produce harmful or unintended side effects. Alignment is an active area of research and remains an unresolved problem as AI systems grow more capable and autonomous. (See also: Reinforcement learning from human feedback, Agentic AI)