AI Safety

Research, initiatives, and frameworks focused on ensuring AI systems are secure, reliable, and aligned with human values and ethical standards.

Lockton and Nexar Introduce Human Benchmark for Autonomous Vehicle Safety

Lockton and Nexar have launched a new human benchmark framework to evaluate autonomous vehicle safety against real-world human driving, designed to aid insurers, regulators, and developers.

June 02, 2026

Replica and Arity Launch Safety Hub for Roadway Risk Analysis

Replica and Arity have launched Safety Hub, a platform that integrates driving behavior and mobility data to help public agencies identify and reduce roadway risk in near real time.

June 02, 2026

GRAIL Reports NHS Galleri Trial Results Showing Fewer Stage IV Cancer Diagnoses

GRAIL presented full results from the NHS-Galleri trial at the 2026 ASCO Annual Meeting, showing a reduction in Stage IV cancer diagnoses and increased detection rates when the Galleri test was added to standard screening.

June 01, 2026

DMind AI Study Finds No AI Model Ready for Web3 Safety Tasks

DMind AI, working with Zhejiang University and Nanyang Technological University, tested 31 major AI models and found none suitable for safety-critical Web3 use cases. The results will be presented at KDD 2026 in Korea.

June 01, 2026

NIST Expands AI Consortium and Invites New Members

The National Institute of Standards and Technology has renamed and expanded its AI consortium to focus on AI measurement, innovation, and adoption, while inviting new organizations to join.

May 30, 2026

Einride Partners with TUV SUD for Independent Verification of Autonomous Safety Governance

Einride has partnered with TUV SUD to conduct an independent assessment of its Safety Management System for autonomous freight operations. The review will evaluate process robustness and alignment with global standards and regulations.

May 27, 2026

Crew Scaler Publishes Comprehensive Study on Multi-Agent AI Security

Crew Scaler has released a detailed 120-page analysis on the security of multi-agent AI systems, evaluating 16 frameworks and identifying major gaps in current safety practices.

May 26, 2026

TELUS Digital Publishes Benchmark on Generative AI Safety Risks

TELUS Digital has released a benchmark study analyzing the safety of 34 AI models through more than 620,000 adversarial tests, showing that smaller models are more vulnerable and reasoning models are harder to exploit.

May 26, 2026

OpenAI Offers $445,000 Research Role Focused on Self-Improving AI Risks

OpenAI has posted a new research position on its Preparedness safety team, offering up to $445,000 to study potential risks from self-improving AI systems, including data poisoning and automation threats.

May 26, 2026

Australia and UK Sign Agreement to Strengthen AI Safety Cooperation

The Australian and UK governments have signed a Memorandum of Understanding to deepen collaboration on AI safety, security, and governance through their national AI institutes.

May 25, 2026

Paragon Health Institute Proposes Framework to Address AI Medical Device Safety

Paragon Health Institute has published a research paper introducing a voluntary framework called Digital Similarity Analysis to improve the safety of medical devices that use artificial intelligence.

May 22, 2026

Microsoft Releases Open Source AI Safety Tools RAMPART and Clarity

Microsoft has released two open source tools, RAMPART and Clarity, designed to help developers test and validate AI agent safety during development. RAMPART converts red team findings into repeatable safety tests, while Clarity assists teams in analyzing design assumptions before implementation.

May 21, 2026

FlexRule Introduces Decision Governance Platform for Enterprise Accountability

FlexRule has launched its Decision Governance platform to help enterprises make decisions explicit, owned, and explainable across manual, automated, and AI operations. The platform aims to close long-standing governance gaps in organizational decision-making.

May 19, 2026

SCRT Labs Integrates Intel Trust Authority into SecretVM

SCRT Labs has integrated Intel Trust Authority into every SecretVM, enabling default attestation and independent verification for confidential computing workloads.

May 19, 2026

Pearl Finds AI Models Match Expert Judgment Only 70% of the Time

Pearl Enterprise evaluated 25 AI models from OpenAI, Anthropic, Google DeepMind, Microsoft, and others, finding that even top systems align with expert judgment only about 70% of the time, with performance dropping to as low as 20% in some domains.

May 14, 2026

EquiLend Acquires Finadium to Expand Securities Finance Research and Consulting

EquiLend has acquired Finadium, a research and consultancy firm in securities finance and capital markets. Finadium will operate as an independent subsidiary under current leadership, maintaining editorial independence while expanding its services.

May 12, 2026

Common Sense Media Launches Youth AI Safety Institute

Common Sense Media has launched the Youth AI Safety Institute, an independent organization that will test AI products used by children, publish the results, and establish safety standards to ensure they are developmentally appropriate.

May 07, 2026

OpenBox AI and Mastra Add Default Runtime Governance to TypeScript Agents

OpenBox AI and Mastra have partnered to integrate runtime governance directly into the Mastra TypeScript agent framework, enabling compliance-ready oversight and audit capabilities by default.

May 05, 2026

CollectivIQ Expands Platform to Address AI Hallucination and Bias

CollectivIQ has announced major upgrades to its AI consensus platform, adding multimodal image generation, integrated payment capture, and expanded retrieval capabilities to enhance accuracy, collaboration, and governance.

May 05, 2026

CISA and International Partners Publish Guidance on Secure Agentic AI Adoption

The Cybersecurity and Infrastructure Security Agency (CISA), along with international partners, has released a guide outlining cybersecurity risks and best practices for adopting agentic AI systems in critical sectors.

May 04, 2026

Subscribe to AI Policy Brief

Weekly report on AI regulations, safety standards, government policies, and compliance requirements worldwide.