Catalogue of Tools & Metrics for Trustworthy AI

These tools and metrics are designed to help AI actors develop and use trustworthy AI systems and applications that respect human rights and are fair, transparent, explainable, robust, secure and safe.

AI-generated Images Challenge Visual Trust in High-risk Scenarios



SafeIMG is a safety-oriented benchmark for evaluating AI-generated (synthetic) image detectors in high-stakes, real-world contexts. Unlike existing detection benchmarks that focus on generic imagery and simple image-level authenticity labels, SafeIMG targets 12 public- and individual-safety scenarios, including situations where a convincing fake image could cause real-world harm to public safety or personal reputation (e.g., fabricated evidence, impersonation, staged incidents).

The benchmark uses images generated with GPT Image 2 and goes beyond a binary "real vs. fake" label. It includes human annotations that localise the specific suspicious regions within each image and explain the underlying issue, whether a local visual artifact (e.g., distorted hands, text, or faces) or a higher-level inconsistency (a commonsense or physical implausibility that a casual viewer might miss but an expert would catch).

Auto-discovered on 2026-08-19 by OECD Catalogue Automation

About the tool



Objective(s):





Type of approach:







Risk management stage(s):


Technology platforms:


Modify this tool

Use Cases

There is no use cases for this tool yet.

Would you like to submit a use case for this tool?

If you have used this tool, we would love to know more about your experience.

Add use case
Partnership on AI

Disclaimer: The tools and metrics featured herein are solely those of the originating authors and are not vetted or endorsed by the OECD or its member countries. The Organisation cannot be held responsible for possible issues resulting from the posting of links to third parties' tools and metrics on this catalogue. More on the methodology can be found at https://oecd.ai/catalogue/faq.