---
title: "future-agi vs myclaw-bench"
type: "comparison"
canonical_url: "https://www.graphcanon.com/compare/future-agi-future-agi-vs-leoyeai-myclaw-bench"
tools: ["future-agi-future-agi", "leoyeai-myclaw-bench"]
---

# future-agi vs myclaw-bench

*GraphCanon updated Aug 2, 2026*

## Verdict

Pick future-agi if future-AGI is an open-source toolkit for evaluating and improving LLMs and AI agents. It includes features like tracing, evaluations, simulations, datasets, gateway operations, and guardrails; pick myclaw-bench if myclaw-bench is a benchmark suite comprising 45 tasks across four tiers designed for evaluating AI agents within the OpenClaw platform.

[future-agi](https://futureagi.com) reports 1.6k GitHub stars, 449 forks, and 596 open issues, last pushed Aug 1, 2026. [myclaw-bench](https://myclaw.ai) has 227 stars, 38 forks, and 2 open issues, last pushed Jul 20, 2026. Figures are from public GitHub metadata via [future-agi's repository](https://github.com/future-agi/future-agi) and [myclaw-bench's repository](https://github.com/LeoYeAI/myclaw-bench).

| | [future-agi](/tools/future-agi-future-agi.md) | [myclaw-bench](/tools/leoyeai-myclaw-bench.md) |
| --- | --- | --- |
| Tagline | End-to-end platform for evaluating, observing, and improving LLM and AI agent applications | Benchmark for AI agents on OpenClaw |
| Stars | 1,559 | 227 |
| Forks | 449 | 38 |
| Open issues | 596 | 2 |
| Language | Python | Python |
| Adopt for | Future-AGI is an open-source toolkit for evaluating and improving LLMs and AI agents. It includes features like tracing, evaluations, simulations, datasets, gateway operations, and guardrails. | myclaw-bench is a benchmark suite comprising 45 tasks across four tiers designed for evaluating AI agents within the OpenClaw platform. |
| Persona | - | - |
| Runtime | - | - |
| License | Apache-2.0 | MIT |
| Categories | AI Agents, Evaluation & Observability | AI Agents, Evaluation & Observability |

## Trust and health

_Sourced signals - not a safety guarantee. No winner column._

| | [future-agi](/tools/future-agi-future-agi.md) | [myclaw-bench](/tools/leoyeai-myclaw-bench.md) |
| --- | --- | --- |
| Maintenance | Very active (96%) | Active (82%) |
| Days since push | 1d | 8d |
| Open issues (now) | 596 | 2 |
| Owner type | Organization | User |
| Full report | [trust report](/tools/future-agi-future-agi/trust.md) | [trust report](/tools/leoyeai-myclaw-bench/trust.md) |

## Decision facts: future-agi

- **Pricing:** freemium - Future-AGI is open-source under the Apache 2.0 license, allowing for free use but with potential paid services through deployment and support channels.
- **Requirements:** Min 4 GB RAM; Requires Docker
- **Adopt for:** Future-AGI is an open-source toolkit for evaluating and improving LLMs and AI agents. It includes features like tracing, evaluations, simulations, datasets, gateway operations, and guardrails.

## Decision facts: myclaw-bench

- **Requirements:** Requires Python version 3.10 or higher to execute the benchmark tasks.; Necessitates installation of the 'uv' package manager from Astral for dependencies management.
- **Adopt for:** myclaw-bench is a benchmark suite comprising 45 tasks across four tiers designed for evaluating AI agents within the OpenClaw platform.

## Choose when

### Choose future-agi if…

- License: future-agi is Apache-2.0, myclaw-bench is MIT.
- Pricing: Future-AGI is open-source under the Apache 2.0 license, allowing for free use but with potential paid services through deployment and support channels..
- Requirements: Min 4 GB RAM; Requires Docker.
- Tags unique to future-agi: ai-gateway, docker-compose, evals, llm.
- future-agi ships Docker support for self-hosted deployment.
- - Use Future-AGI when you require an end-to-end evaluation platform that supports self-hosting through Docker Compose or VM-based services on public clouds.

### Choose myclaw-bench if…

- License: myclaw-bench is MIT, future-agi is Apache-2.0.
- Requirements: Requires Python version 3.10 or higher to execute the benchmark tasks.; Necessitates installation of the 'uv' package manager from Astral for dependencies management..
- Tags unique to myclaw-bench: ai-agent-evaluation, benchmarking-tools, openclaw.
- Use myclaw-bench if you are developing AI agents specifically for deployment on the OpenClaw platform, as it offers a precise evaluation tailored to this ecosystem.

## When NOT to use future-agi

- - Avoid using Future-AGI if you require Kubernetes or Helm support as of the current state; though these are planned for future release, they are not yet available.
- - If your deployment strategy relies on a managed service like AWS Marketplace, consider other options since it is currently 'Coming Soon'.

## When NOT to use myclaw-bench

- Avoid using myclaw-bench if your AI agents will not be deployed on the OpenClaw platform, as its benchmarks are specifically designed to test within this framework.
- Do not use if you require synthetic tests for controlling variables in a highly abstracted scenario, since myclaw-bench exclusively leverages real agent session data.

## Common questions

### What is the difference between future-agi and myclaw-bench?

future-agi: End-to-end platform for evaluating, observing, and improving LLM and AI agent applications. myclaw-bench: Benchmark for AI agents on OpenClaw. See the comparison table for live GitHub stats and shared categories.

### When should I choose future-agi over myclaw-bench?

Choose future-agi over myclaw-bench when License: future-agi is Apache-2.0, myclaw-bench is MIT; Pricing: Future-AGI is open-source under the Apache 2.0 license, allowing for free use but with potential paid services through deployment and support channels.; Requirements: Min 4 GB RAM; Requires Docker; Tags unique to future-agi: ai-gateway, docker-compose, evals, llm; future-agi ships Docker support for self-hosted deployment; - Use Future-AGI when you require an end-to-end evaluation platform that supports self-hosting through Docker Compose or VM-based services on public clouds.

### When should I choose myclaw-bench over future-agi?

Choose myclaw-bench over future-agi when License: myclaw-bench is MIT, future-agi is Apache-2.0; Requirements: Requires Python version 3.10 or higher to execute the benchmark tasks.; Necessitates installation of the 'uv' package manager from Astral for dependencies management.; Tags unique to myclaw-bench: ai-agent-evaluation, benchmarking-tools, openclaw; Use myclaw-bench if you are developing AI agents specifically for deployment on the OpenClaw platform, as it offers a precise evaluation tailored to this ecosystem.

### When should I avoid future-agi?

- Avoid using Future-AGI if you require Kubernetes or Helm support as of the current state; though these are planned for future release, they are not yet available. - If your deployment strategy relies on a managed service like AWS Marketplace, consider other options since it is currently 'Coming Soon'.

### When should I avoid myclaw-bench?

Avoid using myclaw-bench if your AI agents will not be deployed on the OpenClaw platform, as its benchmarks are specifically designed to test within this framework. Do not use if you require synthetic tests for controlling variables in a highly abstracted scenario, since myclaw-bench exclusively leverages real agent session data.

### Is future-agi or myclaw-bench more popular on GitHub?

future-agi has more GitHub stars (1,559 vs 227). Stars measure visibility, not whether either tool fits your constraints.

### Are future-agi and myclaw-bench open source?

Yes - both are open-source projects on GitHub (future-agi: Apache-2.0, myclaw-bench: MIT).

### Where can I find alternatives to future-agi or myclaw-bench?

GraphCanon lists graph-backed alternatives at [future-agi alternatives](/tools/future-agi-future-agi/alternatives) and [myclaw-bench alternatives](/tools/leoyeai-myclaw-bench/alternatives) ([future-agi markdown twin](/tools/future-agi-future-agi/alternatives.md), [myclaw-bench markdown twin](/tools/leoyeai-myclaw-bench/alternatives.md)), ranked by typed relationship edges rather than popularity votes.

### Is there a machine-readable version of this comparison?

Yes. The markdown twin at [this comparison](/compare/future-agi-future-agi-vs-leoyeai-myclaw-bench.md) mirrors this page for agents and LLM crawlers, with the same stats table and FAQ answers.

### Which is better maintained, future-agi or myclaw-bench?

future-agi: Very active. myclaw-bench: Active. Compare maintenance labels, days since push, and release cadence in the trust section below - stars alone do not measure maintenance.

### Where are the full trust reports for future-agi and myclaw-bench?

GraphCanon publishes per-repo trust reports with dated maintenance, provenance, and scan summaries: [future-agi trust report](/tools/future-agi-future-agi/trust); [myclaw-bench trust report](/tools/leoyeai-myclaw-bench/trust).

---

**Machine-readable endpoints**

- JSON: [`/api/graphcanon/graph?tool=future-agi-future-agi`](/api/graphcanon/graph?tool=future-agi-future-agi)
- LLM index: [/llms.txt](/llms.txt)
- Full corpus: [/llms-full.txt](/llms-full.txt)

_GraphCanon - The knowledge graph for AI development. https://www.graphcanon.com/_
