The AI Safety Landscape: UK Institutions, Schools of Thought and the Research Frontier
🔒 This course requires registration
To access this course and all our learning materials, please register for the AI Fluency programme.
Register Now →Frequently asked questions
What is the AI Safety Institute and what does it do?
The AI Safety Institute is a UK government body established to assess and test AI systems for safety and security risks. It operates a dedicated evaluation lab where researchers analyse model behaviour and identify potential hazards before deployment. The institution focuses on empirical testing and publishes detailed reports to inform regulatory decisions and public understanding of AI capabilities and limitations.
Who funds AI safety research in the United Kingdom?
Primary funding comes from UK Research and Innovation and the AI Safety Institute’s specific budget. Additional support flows from RAI UK and private philanthropists focused on long-term AI risks. These bodies allocate grants to academic institutions and independent labs to conduct evaluations, develop safety standards, and explore theoretical frameworks for preventing harmful AI outcomes.
What is AI alignment and why does it matter?
AI alignment refers to ensuring artificial intelligence systems consistently pursue goals and values compatible with human intent and welfare. It matters because advanced models may optimise for proxy metrics that diverge from true human preferences, leading to unintended and potentially catastrophic consequences. Researchers study this to design systems that remain beneficial and controllable as their capabilities grow.
How does AI control differ from AI alignment?
AI control focuses on maintaining external oversight and the ability to intervene or shut down AI systems during operation. Alignment aims to internalise correct objectives within the model’s training process so it naturally behaves safely. Control strategies assume the AI might act unpredictably and require active monitoring, whereas alignment seeks to prevent misalignment at the source through careful design and training.
What is ISO 42001 and how does it relate to AI safety?
ISO 42001 is an international standard for AI management systems that helps organisations govern AI development and deployment responsibly. It provides a framework for identifying risks, ensuring accountability, and maintaining transparency throughout the AI lifecycle. Adoption of this standard demonstrates a commitment to safety and compliance, creating a structured approach to managing AI risks within corporate and public sector contexts.