AI Trace
Content ModerationVerified

Reviewed and published by trentmaziarz, May 11, 2026. Discovered and drafted by our automated research pipeline.

Stability AI uses AI-powered content filters on its platforms and API to detect and block policy-violating content, including prompts or images that may produce NSFW material, child sexual abuse material (CSAM), or other prohibited content. This system operates automatically on both user inputs and generated outputs.

Details

According to Stability AI's published integrity transparency report, the company operates multiple layers of content moderation: in-house text prompt filters that block generation requests violating the Acceptable Use Policy; in-house NSFW image classifiers that flag uploaded images and video; CSAM hashing systems using industry hash lists from Thorn's Safer and the Internet Watch Foundation (IWF); and a combination of automated and human review by an internal Integrity team. The company also uses in-house NSFW classifiers and open-source classifiers to filter training data. Confirmed CSAM is reported to NCMEC.

Products affected

Stability AI Developer Platform APIStable AssistantStable Diffusion (hosted versions)

Sources & Evidence

Cite this record

Trace Foundation. (2026). Stability AI: Stability AI uses AI-powered content filters on its platforms and API to detect and block policy-violating content, including prompts or images that may produce NSFW material, child sexual abuse material (CSAM), or other prohibited content. This system operates automatically on both user inputs and generated outputs (data as of 2026-05-11) [Data set record]. AI Trace. https://www.aitrace.org/r/practice/3e1c6e45-253c-4f44-aff7-d03a851928f7. Accessed September 27, 2026.

Stable link
https://www.aitrace.org/r/practice/3e1c6e45-253c-4f44-aff7-d03a851928f7
Data as of
May 11, 2026
Last verified
May 11, 2026

How to cite AI Trace

Other practices by Stability AI

Have evidence about Stability AI's AI practices? Submit a report.

Report a Sighting →