AI Radar
Support
LiveUpdated 2026-09-19 18:08 UTC

Most safety classifiers break on AI-generated images. That is the central finding of…

Most safety classifiers break on AI-generated images. That is the central finding of UnsafeBench, the ACM CCS 2025…

This is a AI post classified by Jev as Safety & policy (a tool drop), kept by the AI Radar because it carries real work, not commentary.

Most safety classifiers break on AI-generated images. That is the central finding of UnsafeBench, the ACM CCS 2025 benchmark. Ours does not. NSFW Checker is the image-safety API on http://eachlabs.ai. You send image URLs, it tells you in about a second whether each one is safe to show, at three strictness levels. We ran it on the full UnsafeBench test split. 2,037 images, every image under every mode. Highest reported F1 on the Sexual category: 0.896 real-world, 0.879 AI-generated, above GPT-4V and every dedicated NSFW classifier in the paper. Zero degradation on AI-generated images: a di

Posted by each::labs (6.3k followers) 2 days ago · 11 likes · 653 views · view the original post on X. Kept by the AI Radar as Safety & policy. Tools mentioned: Eachlabs.

More AI work like this

Every post is read and classified by Jev (TypeSafe): what it is, which market it belongs to, and whether the link is a real tool. 27.2k posts from 4.8k X accounts over the last 14 days, 1.2k tools, 19 markets. Collected every 5 minutes, fully re-ranked every hour — last update 2026-09-19 18:08 UTC. Full method.