{"data":{"slug":"giskard-ai-giskard-oss","name":"giskard-oss","tagline":"Open-Source Evaluation & Testing library for LLM Agents","github_url":"https://github.com/Giskard-AI/giskard-oss","owner":"Giskard-AI","repo":"giskard-oss","owner_avatar_url":"https://avatars.githubusercontent.com/u/71782571?v=4","primary_language":"Python","stars":5727,"forks":511,"topics":["agent-evaluation","ai-red-team","ai-security","ai-testing","fairness-ai","llm","llm-eval","llm-evaluation","llm-security","llmops","ml-testing","ml-validation","mlops","rag-evaluation","red-team-tools","responsible-ai","trustworthy-ai"],"archived":false,"github_pushed_at":"2026-08-01T23:22:37+00:00","maintenance_label":"Very active","url":"https://www.graphcanon.com/tools/giskard-ai-giskard-oss","markdown_url":"https://www.graphcanon.com/tools/giskard-ai-giskard-oss.md","api_url":"https://www.graphcanon.com/api/graphcanon/tools/giskard-ai-giskard-oss","graph_url":"https://www.graphcanon.com/api/graphcanon/graph?tool=giskard-ai-giskard-oss","description":"🐢 Open-Source Evaluation & Testing library for LLM Agents","homepage_url":"https://docs.giskard.ai","license":"Apache-2.0","open_issues":92,"watchers":41,"ai_summary":"Giskard is an open-source evaluation and testing library for agentic systems in Python, including modules for checks and scanning agent vulnerabilities.","readme_excerpt":"## Install\n\n```sh\npip install giskard\n```\n\nRequires Python 3.12+.\n\n**Telemetry:** Libraries built on `giskard-core` (including `giskard-checks`) may send **optional, aggregated usage analytics** to help improve the product. No prompts, model outputs, or scenario text are included. See [what is collected and how to opt out](libs/giskard-core/README.md#telemetry).\n\n---\n\nGiskard is an open-source Python library for **testing and evaluating agentic systems**. The v3 architecture is a modular set of focused packages — each carrying only the dependencies it needs — built from scratch to wrap anything: an LLM, a black-box agent, or a multi-step pipeline.\n\n| Status         | Package          | Description                                                                                                                                                              |\n| -------------- | ---------------- | ------------------------------------------------------------------------------------------------------------------------------------------------------------------------ |\n| ✅ Beta        | `giskard-checks` | Testing & evaluation — scenario API, built-in checks, LLM-as-judge                                                                                                       |\n| ✅ Beta        | `giskard-scan`   | Agent vulnerability scanner — red teaming, prompt injection, data leakage (successor of [v2 Scan](https://legacy-docs.giskard.ai/en/stable/open_source/scan/index.html)) |\n| 📋 Planned     | `giskard-rag`    | RAG evaluation & synthetic data generation (successor of [v2 RAGET](https://legacy-docs.giskard.ai/en/stable/open_source/testset_generation/index.html))                 |","github_created_at":"2022-03-06T21:45:37+00:00","created_at":"2026-07-07T17:42:05.159294+00:00","updated_at":"2026-08-02T06:00:45.850546+00:00","categories":[{"slug":"evaluation-observability","name":"Evaluation & Observability","url":"https://www.graphcanon.com/categories/evaluation-observability","markdown_url":"https://www.graphcanon.com/categories/evaluation-observability.md","api_url":"https://www.graphcanon.com/api/graphcanon/categories/evaluation-observability"}],"tags":[{"slug":"agent-evaluation","name":"agent-evaluation"},{"slug":"ai-red-team","name":"ai-red-team"},{"slug":"ai-security","name":"ai-security"},{"slug":"ai-testing","name":"ai-testing"},{"slug":"fairness-ai","name":"fairness-ai"},{"slug":"llm","name":"llm"},{"slug":"llm-eval","name":"llm-eval"},{"slug":"llm-evaluation","name":"llm-evaluation"}],"trust":{"provenance":{"is_fork":false,"github_id":466864356,"owner_type":"Organization","methodology":"github_public_v1","parent_repo":null,"near_duplicate_slugs":[]},"computed_at":"2026-08-02T06:00:45.111Z","maintenance":{"label":"Very active","score":96,"methodology":"github_public_v1","releases_90d":14,"days_since_push":0,"last_release_at":"2026-07-13T08:49:31Z"},"security_summary":{"status":"no_lockfile","scanner":null,"low_count":0,"high_count":0,"last_scan_at":"2026-07-11T23:14:55.925Z","medium_count":0,"scan_profile":"none","critical_count":0}},"capability_facts":{"scan":{"source":"repo_scan","observed_at":"2026-08-02T06:00:45.568Z"},"languages":{"value":["python"],"source":"github.language+pyproject.toml","observed_at":"2026-08-02T06:00:45.568Z"},"license_spdx":{"value":"Apache-2.0","source":"github.license","observed_at":"2026-08-02T06:00:45.568Z"}},"decision_facts":{"hosting":null,"pricing":null,"requirements":{"notes":["Requires Python 3.12+"],"min_ram_gb":null,"requires_docker":false},"constraints":{"min_ram_gb":null,"requires_docker":false},"when_to_use":["Use Giskard when you need a tool specialized in assessing the vulnerabilities of language model agents through red teaming and prompt injection scenarios.","Leverage Giskard if your project requires evaluation modules that can work with both LLMs and black-box systems, making it versatile for different types of agentic evaluations."],"when_not_to_use":["Avoid using Giskard if you do not require comprehensive testing features like red teaming or specific checks tailored for language model agents.","Do not use this library if your project is in a programming language other than Python, as Giskard specifically caters to the Python ecosystem."],"source":"enrich:decision_facts","observed_at":"2026-07-17T03:17:51.248Z"},"constraint_facets":{"min_ram_gb":null,"requires_docker":false},"decision_summary":[{"label":"Requirements","value":"Requires Python 3.12+"},{"label":"Adopt for","value":"Giskard is an open-source library that specializes in testing and evaluating agentic systems like LLM agents using Python modules for checks and scanning vulnerabilities."}]}}