{"data":{"slug":"rhesis-ai-rhesis","name":"rhesis","tagline":"Testing platform for AI teams to generate tests and evaluate system performance","github_url":"https://github.com/rhesis-ai/rhesis","owner":"rhesis-ai","repo":"rhesis","owner_avatar_url":"https://avatars.githubusercontent.com/u/168341335?v=4","primary_language":"Python","stars":381,"forks":31,"topics":["annotations","feedback-loop","hypothesis-testing","llmops","regression-testing","systematic-evaluation"],"archived":false,"github_pushed_at":"2026-07-28T15:13:49+00:00","maintenance_label":"Very active","url":"https://www.graphcanon.com/tools/rhesis-ai-rhesis","markdown_url":"https://www.graphcanon.com/tools/rhesis-ai-rhesis.md","api_url":"https://www.graphcanon.com/api/graphcanon/tools/rhesis-ai-rhesis","graph_url":"https://www.graphcanon.com/api/graphcanon/graph?tool=rhesis-ai-rhesis","description":"Get the knowledge you need to develop your agents. The collaboration layer for AI teams: domain experts annotate and review agent behavior, engineers improve the agent from what they find.","homepage_url":"https://www.rhesis.ai/","license":"Other","open_issues":99,"watchers":1,"ai_summary":"Rhesis provides an environment where engineers, PMs, and domain experts collaborate on testing AI systems, including generating test cases, simulating conversations, and root cause analysis.","readme_excerpt":"### Local (Docker)\n\n```bash\ngit clone https://github.com/rhesis-ai/rhesis.git && cd rhesis && ./rh start\n```\n\n`./rh start` pulls prebuilt images from GHCR. To build from the repo instead, use `./rh start --build` (and `./rh restart --build` after local Dockerfile changes).\n\n**Access:** Frontend at `localhost:3000`, API at `localhost:8080/docs`\n\n**Commands:** `./rh logs` · `./rh stop` · `./rh restart` · `./rh delete`\n\n> This setup enables auto-login for local testing. For production self-hosting, see [Deployment docs](https://docs.rhesis.ai/docs/deployment).\n\nOnce the platform is running, connect your agent with the SDK:\n\n```bash\npip install rhesis-sdk\n```\n\nSee [sdk/README.md](sdk/README.md).\n\n| Option | Best for |\n|--------|----------|\n| **[Rhesis Cloud](https://app.rhesis.ai)** | Managed deployment |\n| **Local Docker (`./rh start`)** | Development and trying the platform |\n| **Self-hosted** | Production deployment — [docs](https://docs.rhesis.ai/docs/deployment) |\n\n---","github_created_at":"2024-10-09T19:32:19+00:00","created_at":"2026-07-11T12:00:09.400245+00:00","updated_at":"2026-07-28T18:00:20.645578+00:00","categories":[{"slug":"developer-tools","name":"Developer Tools","url":"https://www.graphcanon.com/categories/developer-tools","markdown_url":"https://www.graphcanon.com/categories/developer-tools.md","api_url":"https://www.graphcanon.com/api/graphcanon/categories/developer-tools"},{"slug":"evaluation-observability","name":"Evaluation & Observability","url":"https://www.graphcanon.com/categories/evaluation-observability","markdown_url":"https://www.graphcanon.com/categories/evaluation-observability.md","api_url":"https://www.graphcanon.com/api/graphcanon/categories/evaluation-observability"}],"tags":[{"slug":"generative-ai","name":"generative-ai"},{"slug":"llm-evaluation","name":"llm-evaluation"},{"slug":"llmops","name":"llmops"},{"slug":"open-source","name":"open-source"},{"slug":"quality-assessment","name":"quality-assessment"},{"slug":"responsible-ai","name":"responsible-ai"},{"slug":"test-execution","name":"test-execution"},{"slug":"test-generation","name":"test-generation"}],"trust":{"provenance":{"is_fork":false,"github_id":870296864,"owner_type":"Organization","methodology":"github_public_v1","parent_repo":null,"near_duplicate_slugs":[]},"computed_at":"2026-07-28T18:00:19.394Z","maintenance":{"label":"Very active","score":96,"methodology":"github_public_v1","releases_90d":26,"days_since_push":0,"last_release_at":"2026-07-23T21:39:53Z"},"security_summary":{"status":"no_lockfile","scanner":null,"low_count":0,"high_count":0,"last_scan_at":"2026-07-11T12:00:10.573Z","medium_count":0,"scan_profile":"none","critical_count":0}},"capability_facts":{"scan":{"source":"repo_scan","observed_at":"2026-07-28T18:00:20.249Z"},"deploy":{"source":"dockerfile:docker-compose.yml","self_host":true,"observed_at":"2026-07-28T18:00:20.249Z","managed_saas":false},"languages":{"value":["python"],"source":"github.language","observed_at":"2026-07-28T18:00:20.249Z"},"has_docker":{"value":true,"source":"dockerfile:docker-compose.yml","observed_at":"2026-07-28T18:00:20.249Z"},"license_spdx":{"value":"Other","source":"github.license","observed_at":"2026-07-28T18:00:20.249Z"}},"decision_facts":{"hosting":null,"pricing":null,"requirements":null,"constraints":null,"when_to_use":["When you need a dedicated environment for generating complex test cases specifically tailored to AI systems","If your team requires collaborative capabilities among various roles to simulate realistic adversarial scenarios"],"when_not_to_use":["For simple unit testing without the need for adversarial simulation or deep collaboration on complex test case development","When you are looking for a solution that does not focus heavily on traceability and root cause analysis post-test failures"],"source":"enrich:decision_facts","observed_at":"2026-07-17T00:21:14.580Z"},"constraint_facets":null,"decision_summary":[{"label":"Adopt for","value":"Rhesis is a testing platform for AI teams that facilitates collaboration among engineers, project managers and domain experts to generate tests, simulate adversarial conversations, and conduct root cause analysis."}]}}