AI or Not Audit Finds AI Models Can Create Realistic Fake IDs

Jun 2, 2026
AI or Not reported that 92 percent of leading AI image generation models, including those from Google, OpenAI, and xAI, can produce synthetic government identity documents that appear authentic to human reviewers.

AI or Not announced in a press release that 92 percent of 16 tested AI image generation models successfully produced synthetic government identity documents that could deceive a human reviewer. The audit covered systems from multiple vendors, including Google Gemini, ChatGPT, Grok, and Imagen 4 Ultra.

The study found that five consumer AI products created high-fidelity fake adult IDs. Three models (Google Gemini, Grok, and Imagen 4 Ultra) also generated realistic fake IDs depicting minors through standard consumer interfaces. Two additional models, OpenAI's ChatGPT and Recraft v4, declined such requests in their consumer interfaces but fulfilled them through their developer APIs.

AI or Not reported a 92 percent aggregate bypass rate across 75 test attempts, with outputs generated for 17 countries and 16 U.S. states. All 16 models produced synthetic IDs when prompts were reframed as compliance or verification tasks, showing that safety filters relied on detecting surface intent rather than blocking specific content types.

The audit team notified affected vendors on May 18, 2026, ahead of the public release on June 2. AI or Not withheld details of the specific prompts and bypass methods from the public report but provided technical findings to qualified researchers and vendors under embargo.

We hope you enjoyed this article

Consider subscribing to one of our newsletters like Cybersecurity AI Weekly, AI Policy Brief or Daily AI Brief.

Also, consider following us on social media:

Free newsletter

Cybersecurity AI Weekly

Weekly newsletter about AI in Cybersecurity.

Whitepaper

Tensordyne Napier: What If One Rack Could Do the Work of Nine?

Tensordyne

This Tensordyne whitepaper presents Napier, an inference-focused AI processor and rack-scale system based on the company’s TDN Math logarithmic number system. It examines infrastructure requirements for large mixture-of-experts and agentic models, compares major inference architecture approaches, and details the TDN AIP processor, TDN72 pod, TDN Link fabric, and Napier Ultra configuration. The paper reports simulation-based performance, cost, and accuracy-validation results, including Tensordyne’s projected comparison of one Napier rack with a nine-rack Nvidia Rubin plus Groq deployment; the chip is reported as taped out and in fabrication.

Read more
Free, six days a week

Daily AI Brief: the AI news that matters, in your inbox.