K-Culture GlossaryAWhere everyone starts
AI Safety
The research and policy field focused on making sure AI doesn't cause harm — also the reason Anthropic was founded.
In plain words
AI Safety is an umbrella term for research and policy work aimed at keeping AI from causing harm. It covers a wide spectrum — everything from stopping chatbots from giving out dangerous instructions, to making sure a far-future superintelligence doesn't end up harming humanity.
For background on safety-related stories, it helps to know that Anthropic split off from OpenAI with "safety first" as its founding principle, and that major labs publish safety reports (model cards) and run red-teaming exercises. "Safety vs. speed" is an ongoing tension in this industry.
See also
Stories using this term
No story has used this term yet. New ones attach here automatically.