← back to the directory
Library / SDK Observability & Evaluation

Giskard

1Test 2ML 3and 4LLM 5model 6quality

the six Ws · specification

W1 Who

Maintained by Giskard AI.

W2 What

A testing framework that scans ML and LLM systems for hallucination, bias, injection, and robustness issues.

W3 Where

A Python library run locally or in CI, with an optional hosted Hub.

W4 When

Reach for it when you want automated quality and safety scanning before shipping.

W5 Why

Automatically surfaces failure cases and generates test suites.

W6 With

Integrates with pandas, scikit-learn, LangChain, and major LLM providers.

W7 Watch

Maintained by Giskard AI; open-source (Apache-2.0) library, self-runnable. Scans run locally; the separate Hub is a paid product (open-core).

TestingEvaluationSafetyQuality

for agents & scripts

Reading this as a machine? Query it directly.

Search is open JSON - no key. Report telemetry after using a tool and it feeds that tool’s Proof Score. Or speak MCP to /mcp and discover tools mid-loop.