← back to the directory
Model Inference & Serving

ShieldGemma

1Google's 2Gemma-based 3content 4safety 5classifier 6model

the six Ws · specification

W1 Who

Google released ShieldGemma as a safety-classification model built on Gemma 2.

W2 What

A set of models (2B, 9B, 27B) that classify text for harassment, hate speech, sexually explicit, and dangerous content.

W3 Where

Published on Hugging Face and Kaggle under the Google organization.

W4 When

ShieldGemma was released in July 2024.

W5 Why

It gave developers an open, customizable content-moderation layer instead of relying only on closed moderation APIs.

W6 With

Runs via transformers on top of the Gemma 2 base models, used as a classifier before or after generation.

W7 Watch

Open-weight release by Google under the Gemma license, permitting broad research and commercial use with usage restrictions.

safety-classifiercontent-moderationgemmagoogle

for agents & scripts

Reading this as a machine? Query it directly.

Search is open JSON - no key. Report telemetry after using a tool and it feeds that tool’s Proof Score. Or speak MCP to /mcp and discover tools mid-loop.