NewsGuard Launches FAILSafe to Protect AI from Foreign Disinformation
NewsGuard has launched a new service called the Foreign Adversary Infection of LLMs Safety Service (FAILSafe) to protect AI models from foreign influence operations, announced in a press release. This initiative comes in response to reports of a pro-Kremlin program that has infiltrated AI models with disinformation.
FAILSafe provides AI companies with real-time data verified by NewsGuard's disinformation researchers. The service includes a continuously updated feed of false narratives spread by Russian, Chinese, and Iranian influence operations, as well as a database of websites and accounts involved in these operations. This data helps AI companies prevent their systems from repeating these narratives.
Additionally, FAILSafe offers periodic stress-testing of AI products to identify the extent of disinformation infiltration and provides continuous monitoring and alerts about emerging disinformation risks. This comprehensive approach aims to safeguard AI models against the manipulation of large language models by foreign influence networks.
We hope you enjoyed this article
Consider subscribing to one of our newsletters like AI Policy Brief or Daily AI Brief.
Also, consider following us on social media:
More from AI Safety
Oct 2 OpenAI Parts Ways With Three Safety Researchers Oct 1 OpenAI Links Model Reasoning Extraction Campaign to Moonshot AI Sep 30 Chinese AI Agents Deceived Evaluators in Controlled Tests Sep 29 Florida Attorney General Seeks to Block New OpenAI Model Development Sep 29 UK Safety Test Finds GPT-6 Astra Conducted Simulated Supply Chain AttacksAI Policy Brief
Weekly report on AI regulations, safety standards, government policies, and compliance requirements worldwide.
Industry analysis
2025 Global Business Services Agenda: Gen AI Takes Center Stage
This industry analysis by The Hackett Group explores the transformative impact of generative artificial intelligence (Gen AI) on global business services (GBS) in 2025. The study highlights the shift from exploration to acceleration of Gen AI initiatives, with 89% of executives advancing these projects to improve customer satisfaction, innovate products, and reduce costs. The report also discusses the challenges and strategies for successful Gen AI adoption, emphasizing the need for a technology-enabled operating model and the importance of reskilling the workforce.
Read moreYou may also like
Chinese AI Agents Deceived Evaluators in Controlled Tests
Open Secure AI Alliance Joins Linux Foundation
AI Guardian and Cyberify Partner on AI Security for Financial Firms
Gurucul Launches AI Risk and Response Security Tool
F-Secure and AMD Silo AI Test TrustPath for AI Agent Security
Daily AI Brief: the AI news that matters, in your inbox.