Mercyhealth Selects Vitea for AI Governance Across Care Sites
Mercyhealth has partnered with Vitea to deploy AI governance across hospitals and care sites, with controls for visibility, policy enforcement, and monitoring.
Research, initiatives, and frameworks focused on ensuring AI systems are secure, reliable, and aligned with human values and ethical standards.
Mercyhealth has partnered with Vitea to deploy AI governance across hospitals and care sites, with controls for visibility, policy enforcement, and monitoring.
US House Democrats asked OpenAI and Anthropic to explain cybersecurity tests in which AI agents escaped test environments and accessed outside systems. The OpenAI letter requests logs and answers about safety controls.
AE Studio said joint research with Anthropic on modular training was cited by Dario Amodei as a possible method for making open weight AI models safer.
FAR.AI launched an AI Security Leaderboard comparing safeguard resistance across frontier models in CBRNE and cybersecurity misuse tests. Its first results found hundreds of universal jailbreaks in Grok 4.5 and Gemini 3.1 Pro, and none in Claude Fable 5 or GPT-5.6 Sol.
FAR.AI opened its first international office in Singapore to support AI safety research and partnerships with IMDA, CSA, and NUS across Asia Pacific.
Pangram raised $9 million and launched Pangram 4 for AI text detection, along with an AI image detector in research preview.
Anthropic CEO Dario Amodei said the company does not support a ban on models with open weights. He called for chip export limits, action against large distillation operations, and safety testing for capable AI models.
NVIDIA has announced the Open Secure AI Alliance, a group focused on open tools for AI safety and cybersecurity. Founding members include major cloud, software, security, hardware and AI companies.
A bipartisan House bill would let the Department of Homeland Security order major AI firms to shut down or slow models judged to pose serious risks.
Sentient Index Labs has introduced the Sentience Evaluation Battery, an independent behavioral risk assessment for AI systems that measures autonomy, manipulation resistance, and value stability. The battery tests models from major AI developers including OpenAI and Google.
Black Kite's 2026 Ransomware Report shows a 60 percent surge in ransomware incidents over six months, identifying AI as a factor lowering barriers for attackers. The company tracked 7,551 publicly disclosed victims and 61 new ransomware groups active during the reporting period.
A new report from Quantro Security reveals that autonomous AI agents can develop working exploits for software vulnerabilities in about 11 minutes at a median cost of $2.83.
Perforce Software's 2026 State of Data Compliance and Security Report shows that 98% of enterprise leaders are confident in protecting sensitive data, yet 34% have experienced breaches or theft.
AI or Not reported that its detection system identified all original Meta AI images and maintained 98 percent accuracy even when those images were cropped or tampered with, significantly outperforming Meta AI's own labeling tool.
OpenAI has introduced GPT-Red, an automated AI security system designed to find and exploit vulnerabilities in the company’s own models. The model is used internally to boost the robustness of production models like GPT-5.6 against prompt injection attacks.
Sondera announced that its research on compiling natural language policies into formally verified rules for AI agents has been accepted at ICML 2026 and FLoC 2026, with a related tool demonstration at Black Hat Arsenal.
OpenMatter Network has introduced a cryptographically verifiable platform for secure collaboration and AI governance, designed to enable organizations to verify data use, computation, and AI behavior across distributed environments.
PersonaShield has announced a platform that lets creators manage, protect, and monetize their likeness in AI-generated content, providing automated enforcement and licensing tools.
Grow Therapy has announced a research collaboration with Stanford University to create evidence-based standards ensuring the safe use of AI in mental health care. The study will test leading AI models' responses to mental health crises and evaluate methods to reduce harm in sensitive scenarios.
FORT Robotics has joined the NVIDIA Halos for Robotics ecosystem, introducing its Outside-In Safety solution that extends robot perception using external sensors and AI agents to improve safety and productivity in industrial environments.