Giskard Unveils Phare: A New Benchmark for Evaluating AI Models
Giskard has introduced Phare, a new open and independent benchmark designed to evaluate large language models (LLMs) on key security dimensions such as hallucination, factual accuracy, bias, and potential for harm. This announcement was made during the Paris AI Summit, with Google DeepMind collaborating as a research partner. The initiative aims to provide open measurements to assess the trustworthiness of generative AI models in real-world applications announced on their website.
Phare, which stands for "Potential Harm Assessment & Risk Evaluation," is designed to evaluate language models across multiple languages, initially including English, French, and Spanish. The benchmark will incorporate diverse linguistic and cultural contexts to ensure comprehensive assessments. The initial scope covers leading models from top AI labs such as OpenAI, Anthropic, Google DeepMind, Meta, Mistral, Alibaba, and DeepSeek.
The benchmark consists of modular test components focusing on four fundamental safety categories: hallucination, bias and fairness, intentional abuse by users, and harmful content generation. Giskard maintains full autonomy in determining the benchmark design, ensuring independence from model developers. The results from these assessments will be tracked on a public leaderboard, with future modules expanding to cover more languages and additional security aspects.
This collaborative effort is part of a broader initiative to improve AI security and robustness, encouraging practical developments in AI safety. Giskard plans to open-source a representative set of samples for each benchmarking module, enabling independent verification and private model testing.
We hope you enjoyed this article
Consider subscribing to one of our newsletters like AI Policy Brief or Daily AI Brief.
Also, consider following us on social media:
More from AI Safety
Sep 16 Trump calls AI safety fears a hoax during live call with NVIDIA CEO Sep 15 Project Liberty and Partners Launch Pro-Human AI Coalition Sep 15 Gensyn Releases Auditable open-1b AI Model Sep 15 King Charles to Host AI Leaders for Safety Talks Sep 15 China's Foreign Ministry Answers AI Slowdown CallsAI Policy Brief
Weekly report on AI regulations, safety standards, government policies, and compliance requirements worldwide.
Industry analysis
2025 Global Business Services Agenda: Gen AI Takes Center Stage
This industry analysis by The Hackett Group explores the transformative impact of generative artificial intelligence (Gen AI) on global business services (GBS) in 2025. The study highlights the shift from exploration to acceleration of Gen AI initiatives, with 89% of executives advancing these projects to improve customer satisfaction, innovate products, and reduce costs. The report also discusses the challenges and strategies for successful Gen AI adoption, emphasizing the need for a technology-enabled operating model and the importance of reskilling the workforce.
Read moreYou may also like
OpenAI Chief Scientist Calls for Slower AI Scaling
Anthropic Attributes Its Largest Measured Distillation Campaign to Alibaba
OpenAI Releases GPT-6 Astra With New Cybersecurity Safeguards
OpenAI Plans Limited Astra Release After Critical Cybersecurity Rating
OpenAI Executive Warns of Persistent AI Cyber Attacks
Daily AI Brief: the AI news that matters, in your inbox.