{"data":{"node":{"slug":"bentoml-bentoml","name":"BentoML","tagline":"The easiest way to serve AI apps and models","github_url":"https://github.com/bentoml/BentoML","owner":"bentoml","repo":"BentoML","owner_avatar_url":"https://avatars.githubusercontent.com/u/49176046?v=4","primary_language":"Python","stars":8793,"forks":1010,"topics":["ai-inference","deep-learning","generative-ai","inference-platform","llm","llm-inference","llm-serving","llmops","machine-learning","ml-engineering","mlops","model-inference-service","model-serving","multimodal","python"],"archived":false,"github_pushed_at":"2026-08-03T17:00:21+00:00","maintenance_label":"Active","stars_delta_30d":65,"url":"https://www.graphcanon.com/tools/bentoml-bentoml","markdown_url":"https://www.graphcanon.com/tools/bentoml-bentoml.md","api_url":"https://www.graphcanon.com/api/graphcanon/tools/bentoml-bentoml","graph_url":"https://www.graphcanon.com/api/graphcanon/graph?tool=bentoml-bentoml"},"categories":[{"slug":"inference-serving","name":"Inference & Serving","url":"https://www.graphcanon.com/categories/inference-serving","markdown_url":"https://www.graphcanon.com/categories/inference-serving.md","api_url":"https://www.graphcanon.com/api/graphcanon/categories/inference-serving"},{"slug":"model-training","name":"Model Training","url":"https://www.graphcanon.com/categories/model-training","markdown_url":"https://www.graphcanon.com/categories/model-training.md","api_url":"https://www.graphcanon.com/api/graphcanon/categories/model-training"}],"tags":[{"slug":"ai-inference","name":"ai-inference"},{"slug":"deep-learning","name":"deep-learning"},{"slug":"generative-ai","name":"generative-ai"},{"slug":"inference-platform","name":"inference-platform"},{"slug":"llm","name":"llm"},{"slug":"llm-inference","name":"llm-inference"},{"slug":"llm-serving","name":"llm-serving"},{"slug":"mlops","name":"mlops"}],"edges":[{"type":"integrates_with","direction":"out","explanation":"BentoML, which serves as a unified framework for model serving in AI applications, integrates with Haystack to enable the deployment of Haystack's modular pipelines and agent workflows, thus facilitating the scalable execution of tasks such as retrieval, routing, memory management, and generation within production-ready LLM applications.","successor_context":null,"tool":{"slug":"deepset-ai-haystack","name":"haystack","tagline":"Open-source AI orchestration framework for building context-engineered LLM applications.","github_url":"https://github.com/deepset-ai/haystack","owner":"deepset-ai","repo":"haystack","owner_avatar_url":"https://avatars.githubusercontent.com/u/51827949?v=4","primary_language":"Python","stars":26073,"forks":2972,"topics":["agent","agents","ai","gemini","generative-ai","gpt-4","information-retrieval","large-language-models","llm","machine-learning","nlp","orchestration","python","pytorch","question-answering","rag","retrieval-augmented-generation","semantic-search","summarization","transformers"],"archived":false,"github_pushed_at":"2026-08-01T03:06:32+00:00","maintenance_label":"Very active","url":"https://www.graphcanon.com/tools/deepset-ai-haystack","markdown_url":"https://www.graphcanon.com/tools/deepset-ai-haystack.md","api_url":"https://www.graphcanon.com/api/graphcanon/tools/deepset-ai-haystack","graph_url":"https://www.graphcanon.com/api/graphcanon/graph?tool=deepset-ai-haystack"}},{"type":"integrates_with","direction":"out","explanation":"BentoML, as a framework for building online serving systems optimized for AI models, can integrate with Jina-Serve (referred to here as 'serve') to expand its service capabilities by leveraging Jina-Serve's support for gRPC, HTTP, and WebSockets protocols, thus enabling more versatile deployment options.","successor_context":null,"tool":{"slug":"jina-ai-serve","name":"serve","tagline":"Build multimodal AI applications with cloud-native stack","github_url":"https://github.com/jina-ai/serve","owner":"jina-ai","repo":"serve","owner_avatar_url":"https://avatars.githubusercontent.com/u/60539444?v=4","primary_language":"Python","stars":21863,"forks":2243,"topics":["cloud-native","cncf","deep-learning","docker","fastapi","framework","generative-ai","grpc","jaeger","kubernetes","llmops","machine-learning","microservice","mlops","multimodal","neural-search","opentelemetry","orchestration","pipeline","prometheus"],"archived":false,"github_pushed_at":"2025-03-24T13:59:54+00:00","maintenance_label":"Dormant","url":"https://www.graphcanon.com/tools/jina-ai-serve","markdown_url":"https://www.graphcanon.com/tools/jina-ai-serve.md","api_url":"https://www.graphcanon.com/api/graphcanon/tools/jina-ai-serve","graph_url":"https://www.graphcanon.com/api/graphcanon/graph?tool=jina-ai-serve"}},{"type":"related","direction":"out","explanation":"BentoML and Serve both cater to building multimodal AI applications, but they do so differently. BentoML focuses on serving models via APIs and Docker containers, while Serve emphasizes cloud-native infrastructure for deployment.","successor_context":null,"tool":{"slug":"jina-ai-serve","name":"serve","tagline":"Build multimodal AI applications with cloud-native stack","github_url":"https://github.com/jina-ai/serve","owner":"jina-ai","repo":"serve","owner_avatar_url":"https://avatars.githubusercontent.com/u/60539444?v=4","primary_language":"Python","stars":21863,"forks":2243,"topics":["cloud-native","cncf","deep-learning","docker","fastapi","framework","generative-ai","grpc","jaeger","kubernetes","llmops","machine-learning","microservice","mlops","multimodal","neural-search","opentelemetry","orchestration","pipeline","prometheus"],"archived":false,"github_pushed_at":"2025-03-24T13:59:54+00:00","maintenance_label":"Dormant","url":"https://www.graphcanon.com/tools/jina-ai-serve","markdown_url":"https://www.graphcanon.com/tools/jina-ai-serve.md","api_url":"https://www.graphcanon.com/api/graphcanon/tools/jina-ai-serve","graph_url":"https://www.graphcanon.com/api/graphcanon/graph?tool=jina-ai-serve"}},{"type":"related","direction":"out","explanation":"Both BentoML and MLflow are involved in the broader area of model serving, inference, and deployment of AI models. However, they serve different specific purposes with MLflow focusing on tracking experiments and managing the lifecycle of models while BentoML focuses more on deploying and serving models efficiently.","successor_context":null,"tool":{"slug":"mlflow-mlflow","name":"mlflow","tagline":"AI engineering platform for debugging, evaluating, monitoring, and optimizing AI applications","github_url":"https://github.com/mlflow/mlflow","owner":"mlflow","repo":"mlflow","owner_avatar_url":"https://avatars.githubusercontent.com/u/39938107?v=4","primary_language":"Python","stars":27591,"forks":6189,"topics":["agentops","agents","ai","ai-governance","apache-spark","evaluation","langchain","llm-evaluation","llmops","machine-learning","ml","mlflow","mlops","model-management","observability","open-source","openai","prompt-engineering"],"archived":false,"github_pushed_at":"2026-08-20T00:54:28+00:00","maintenance_label":"Very active","stars_delta_30d":476,"url":"https://www.graphcanon.com/tools/mlflow-mlflow","markdown_url":"https://www.graphcanon.com/tools/mlflow-mlflow.md","api_url":"https://www.graphcanon.com/api/graphcanon/tools/mlflow-mlflow","graph_url":"https://www.graphcanon.com/api/graphcanon/graph?tool=mlflow-mlflow"}},{"type":"related","direction":"out","explanation":"Pixeltable and BentoML are both used in multimodal AI data applications. Pixeltable provides a backend to handle data from different modalities, whereas BentoML focuses on the deployment side by serving these models efficiently.","successor_context":null,"tool":{"slug":"pixeltable-pixeltable","name":"pixeltable","tagline":"Unified multimodal backend for AI data apps","github_url":"https://github.com/pixeltable/pixeltable","owner":"pixeltable","repo":"pixeltable","owner_avatar_url":"https://avatars.githubusercontent.com/u/160283145?v=4","primary_language":"Python","stars":1613,"forks":219,"topics":["ai","computer-vision","data-science","database","feature-engineering","feature-store","genai","llm","machine-learning","ml","multimodal","vector-database"],"archived":false,"github_pushed_at":"2026-08-21T06:36:51+00:00","maintenance_label":"Very active","stars_delta_30d":9,"url":"https://www.graphcanon.com/tools/pixeltable-pixeltable","markdown_url":"https://www.graphcanon.com/tools/pixeltable-pixeltable.md","api_url":"https://www.graphcanon.com/api/graphcanon/tools/pixeltable-pixeltable","graph_url":"https://www.graphcanon.com/api/graphcanon/graph?tool=pixeltable-pixeltable"}},{"type":"integrates_with","direction":"out","explanation":"BentoML, which serves as a framework for deploying and managing machine learning models, integrates with vLLM, an efficient model serving engine for large language models (LLMs). This integration allows BentoML to leverage vLLM's capabilities for fast and memory-efficient inference of LLMs when building online serving systems.","successor_context":null,"tool":{"slug":"vllm-project-vllm","name":"vllm","tagline":"A high-throughput and memory-efficient inference and serving engine for LLMs","github_url":"https://github.com/vllm-project/vllm","owner":"vllm-project","repo":"vllm","owner_avatar_url":"https://avatars.githubusercontent.com/u/136984999?v=4","primary_language":"Python","stars":87847,"forks":20135,"topics":["amd","blackwell","cuda","deepseek","deepseek-v3","gpt","gpt-oss","inference","kimi","llama","llm","llm-serving","model-serving","moe","openai","pytorch","qwen","qwen3","tpu","transformer"],"archived":false,"github_pushed_at":"2026-08-01T11:55:36+00:00","maintenance_label":"Very active","url":"https://www.graphcanon.com/tools/vllm-project-vllm","markdown_url":"https://www.graphcanon.com/tools/vllm-project-vllm.md","api_url":"https://www.graphcanon.com/api/graphcanon/tools/vllm-project-vllm","graph_url":"https://www.graphcanon.com/api/graphcanon/graph?tool=vllm-project-vllm"}},{"type":"integrates_with","direction":"out","explanation":"BentoML integrates with Ray to leverage Ray's distributed computing capabilities, specifically its Serve library for scalable model serving, which enhances BentoML's ability to deploy and scale AI models across a cluster.","successor_context":null,"tool":{"slug":"ray-project-ray","name":"ray","tagline":"Ray is an AI compute engine with a core distributed runtime and AI Libraries for accelerating ML workloads.","github_url":"https://github.com/ray-project/ray","owner":"ray-project","repo":"ray","owner_avatar_url":"https://avatars.githubusercontent.com/u/22125274?v=4","primary_language":"Python","stars":43526,"forks":7929,"topics":["data-science","deep-learning","deployment","distributed","hyperparameter-optimization","hyperparameter-search","large-language-models","llm","llm-inference","llm-serving","machine-learning","optimization","parallel","python","pytorch","ray","reinforcement-learning","rllib","serving","tensorflow"],"archived":false,"github_pushed_at":"2026-08-16T00:26:16+00:00","maintenance_label":"Very active","stars_delta_30d":270,"url":"https://www.graphcanon.com/tools/ray-project-ray","markdown_url":"https://www.graphcanon.com/tools/ray-project-ray.md","api_url":"https://www.graphcanon.com/api/graphcanon/tools/ray-project-ray","graph_url":"https://www.graphcanon.com/api/graphcanon/graph?tool=ray-project-ray"}},{"type":"alternative","direction":"in","explanation":"Both Jina-Serve and BentoML are frameworks for building and deploying AI services, but they differ in their architectural approaches and the protocols they support (e.g., gRPC vs HTTP/REST).","successor_context":null,"tool":{"slug":"jina-ai-serve","name":"serve","tagline":"Build multimodal AI applications with cloud-native stack","github_url":"https://github.com/jina-ai/serve","owner":"jina-ai","repo":"serve","owner_avatar_url":"https://avatars.githubusercontent.com/u/60539444?v=4","primary_language":"Python","stars":21863,"forks":2243,"topics":["cloud-native","cncf","deep-learning","docker","fastapi","framework","generative-ai","grpc","jaeger","kubernetes","llmops","machine-learning","microservice","mlops","multimodal","neural-search","opentelemetry","orchestration","pipeline","prometheus"],"archived":false,"github_pushed_at":"2025-03-24T13:59:54+00:00","maintenance_label":"Dormant","url":"https://www.graphcanon.com/tools/jina-ai-serve","markdown_url":"https://www.graphcanon.com/tools/jina-ai-serve.md","api_url":"https://www.graphcanon.com/api/graphcanon/tools/jina-ai-serve","graph_url":"https://www.graphcanon.com/api/graphcanon/graph?tool=jina-ai-serve"}}],"neighbours":[{"slug":"mlflow-mlflow","name":"mlflow","tagline":"AI engineering platform for debugging, evaluating, monitoring, and optimizing AI applications","github_url":"https://github.com/mlflow/mlflow","owner":"mlflow","repo":"mlflow","owner_avatar_url":"https://avatars.githubusercontent.com/u/39938107?v=4","primary_language":"Python","stars":27591,"forks":6189,"topics":["agentops","agents","ai","ai-governance","apache-spark","evaluation","langchain","llm-evaluation","llmops","machine-learning","ml","mlflow","mlops","model-management","observability","open-source","openai","prompt-engineering"],"archived":false,"github_pushed_at":"2026-08-20T00:54:28+00:00","maintenance_label":"Very active","url":"https://www.graphcanon.com/tools/mlflow-mlflow","markdown_url":"https://www.graphcanon.com/tools/mlflow-mlflow.md","api_url":"https://www.graphcanon.com/api/graphcanon/tools/mlflow-mlflow","graph_url":"https://www.graphcanon.com/api/graphcanon/graph?tool=mlflow-mlflow","shared_categories":["model-training","inference-serving"]},{"slug":"ai-dynamo-dynamo","name":"dynamo","tagline":"A Datacenter Scale Distributed Inference Serving Framework","github_url":"https://github.com/ai-dynamo/dynamo","owner":"ai-dynamo","repo":"dynamo","owner_avatar_url":"https://avatars.githubusercontent.com/u/201626793?v=4","primary_language":"Rust","stars":7575,"forks":1368,"topics":["diffusion","disaggregated-serving","kubernetes","llm-inference","omni","routing-engine","rust","sglang","tensorrt-llm","vllm"],"archived":false,"github_pushed_at":"2026-07-25T05:43:22+00:00","maintenance_label":"Active","url":"https://www.graphcanon.com/tools/ai-dynamo-dynamo","markdown_url":"https://www.graphcanon.com/tools/ai-dynamo-dynamo.md","api_url":"https://www.graphcanon.com/api/graphcanon/tools/ai-dynamo-dynamo","graph_url":"https://www.graphcanon.com/api/graphcanon/graph?tool=ai-dynamo-dynamo","shared_categories":["inference-serving"]},{"slug":"tensorflow-serving","name":"serving","tagline":"A flexible, high-performance serving system for machine learning models","github_url":"https://github.com/tensorflow/serving","owner":"tensorflow","repo":"serving","owner_avatar_url":"https://avatars.githubusercontent.com/u/15658638?v=4","primary_language":"C++","stars":6359,"forks":2204,"topics":["cpp","deep-learning","deep-neural-networks","machine-learning","ml","neural-network","python","serving","tensorflow"],"archived":false,"github_pushed_at":"2026-07-30T07:02:43+00:00","maintenance_label":"Active","url":"https://www.graphcanon.com/tools/tensorflow-serving","markdown_url":"https://www.graphcanon.com/tools/tensorflow-serving.md","api_url":"https://www.graphcanon.com/api/graphcanon/tools/tensorflow-serving","graph_url":"https://www.graphcanon.com/api/graphcanon/graph?tool=tensorflow-serving","shared_categories":["inference-serving"]},{"slug":"kserve-kserve","name":"kserve","tagline":"Standardized Distributed Generative and Predictive AI Inference Platform for Scalable, Multi-Framework Deployment on Kubernetes","github_url":"https://github.com/kserve/kserve","owner":"kserve","repo":"kserve","owner_avatar_url":"https://avatars.githubusercontent.com/u/83512434?v=4","primary_language":"Go","stars":5731,"forks":1591,"topics":["artificial-intelligence","cncf","genai","hacktoberfest","istio","k8s","knative","kserve","kubeflow","kubernetes","llm-inference","machine-learning","mlops","model-interpretability","model-serving","pytorch","service-mesh","tensorflow","vllm","xgboost"],"archived":false,"github_pushed_at":"2026-07-24T14:21:37+00:00","maintenance_label":"Active","url":"https://www.graphcanon.com/tools/kserve-kserve","markdown_url":"https://www.graphcanon.com/tools/kserve-kserve.md","api_url":"https://www.graphcanon.com/api/graphcanon/tools/kserve-kserve","graph_url":"https://www.graphcanon.com/api/graphcanon/graph?tool=kserve-kserve","shared_categories":["inference-serving"]},{"slug":"seldonio-seldon-core","name":"seldon-core","tagline":"An MLOps framework to package, deploy, monitor and manage thousands of production machine learning models","github_url":"https://github.com/SeldonIO/seldon-core","owner":"SeldonIO","repo":"seldon-core","owner_avatar_url":"https://avatars.githubusercontent.com/u/10297834?v=4","primary_language":"Go","stars":4765,"forks":867,"topics":["aiops","deployment","kubernetes","machine-learning","machine-learning-operations","mlops","production-machine-learning","serving"],"archived":false,"github_pushed_at":"2026-03-23T11:39:54+00:00","maintenance_label":"Slowing","url":"https://www.graphcanon.com/tools/seldonio-seldon-core","markdown_url":"https://www.graphcanon.com/tools/seldonio-seldon-core.md","api_url":"https://www.graphcanon.com/api/graphcanon/tools/seldonio-seldon-core","graph_url":"https://www.graphcanon.com/api/graphcanon/graph?tool=seldonio-seldon-core","shared_categories":["inference-serving"]},{"slug":"pytorch-serve","name":"serve","tagline":"Serve, optimize and scale PyTorch models in production","github_url":"https://github.com/pytorch/serve","owner":"pytorch","repo":"serve","owner_avatar_url":"https://avatars.githubusercontent.com/u/21003710?v=4","primary_language":"Java","stars":4350,"forks":882,"topics":["cpu","deep-learning","docker","gpu","kubernetes","machine-learning","metrics","mlops","optimization","pytorch","serving"],"archived":true,"github_pushed_at":"2025-08-06T19:17:08+00:00","maintenance_label":"Archived","url":"https://www.graphcanon.com/tools/pytorch-serve","markdown_url":"https://www.graphcanon.com/tools/pytorch-serve.md","api_url":"https://www.graphcanon.com/api/graphcanon/tools/pytorch-serve","graph_url":"https://www.graphcanon.com/api/graphcanon/graph?tool=pytorch-serve","shared_categories":["inference-serving"]},{"slug":"superlinked-sie","name":"sie","tagline":"Open-source inference server and production cluster for all the models your agent needs.","github_url":"https://github.com/superlinked/sie","owner":"superlinked","repo":"sie","owner_avatar_url":"https://avatars.githubusercontent.com/u/94243920?v=4","primary_language":"Python","stars":2297,"forks":215,"topics":["bge","colbert","data-pipeline","deep-learning","embeddings","inference","inference-server","information-retrieval","llm","ml","mlops","natural-language-processing","nlp","python","reranking","retrieval","retrieval-augmented-generation","semantic-search","splade","vector-search"],"archived":false,"github_pushed_at":"2026-07-22T08:53:55+00:00","maintenance_label":"Steady","url":"https://www.graphcanon.com/tools/superlinked-sie","markdown_url":"https://www.graphcanon.com/tools/superlinked-sie.md","api_url":"https://www.graphcanon.com/api/graphcanon/tools/superlinked-sie","graph_url":"https://www.graphcanon.com/api/graphcanon/graph?tool=superlinked-sie","shared_categories":["inference-serving"]},{"slug":"vertaai-modeldb","name":"modeldb","tagline":"Open Source ML Model Versioning Metadata and Experiment Management","github_url":"https://github.com/VertaAI/modeldb","owner":"VertaAI","repo":"modeldb","owner_avatar_url":"https://avatars.githubusercontent.com/u/45020510?v=4","primary_language":"Java","stars":1749,"forks":289,"topics":["machine-learning","mit","model-management","model-versioning","modeldb","verta"],"archived":false,"github_pushed_at":"2024-07-23T17:06:34+00:00","maintenance_label":"Dormant","url":"https://www.graphcanon.com/tools/vertaai-modeldb","markdown_url":"https://www.graphcanon.com/tools/vertaai-modeldb.md","api_url":"https://www.graphcanon.com/api/graphcanon/tools/vertaai-modeldb","graph_url":"https://www.graphcanon.com/api/graphcanon/graph?tool=vertaai-modeldb","shared_categories":["model-training"]},{"slug":"modelfoxdotdev-modelfox","name":"modelfox","tagline":"ModelFox simplifies machine learning model training and deployment.","github_url":"https://github.com/modelfoxdotdev/modelfox","owner":"modelfoxdotdev","repo":"modelfox","owner_avatar_url":"https://avatars.githubusercontent.com/u/102552266?v=4","primary_language":"Rust","stars":1467,"forks":64,"topics":["automl","developer-tools","elixir","elixir-lang","go","golang","javascript","js","machine-learning","mlops","python","python3","ruby","ruby-on-rails","rust","rust-crate","rust-lang","rust-library","rustlang"],"archived":false,"github_pushed_at":"2024-08-02T17:23:15+00:00","maintenance_label":"Dormant","url":"https://www.graphcanon.com/tools/modelfoxdotdev-modelfox","markdown_url":"https://www.graphcanon.com/tools/modelfoxdotdev-modelfox.md","api_url":"https://www.graphcanon.com/api/graphcanon/tools/modelfoxdotdev-modelfox","graph_url":"https://www.graphcanon.com/api/graphcanon/graph?tool=modelfoxdotdev-modelfox","shared_categories":["model-training"]},{"slug":"ebhy-budgetml","name":"budgetml","tagline":"Deploys ML inference service economically","github_url":"https://github.com/ebhy/budgetml","owner":"ebhy","repo":"budgetml","owner_avatar_url":"https://avatars.githubusercontent.com/u/76654256?v=4","primary_language":"Python","stars":1343,"forks":65,"topics":["api","data-science","deployment","fastapi","inference","machine-learning","mlops"],"archived":false,"github_pushed_at":"2024-02-12T17:29:24+00:00","maintenance_label":"Dormant","url":"https://www.graphcanon.com/tools/ebhy-budgetml","markdown_url":"https://www.graphcanon.com/tools/ebhy-budgetml.md","api_url":"https://www.graphcanon.com/api/graphcanon/tools/ebhy-budgetml","graph_url":"https://www.graphcanon.com/api/graphcanon/graph?tool=ebhy-budgetml","shared_categories":["inference-serving"]},{"slug":"vercel-modelfusion","name":"modelfusion","tagline":"TypeScript library for building AI applications","github_url":"https://github.com/vercel/modelfusion","owner":"vercel","repo":"modelfusion","owner_avatar_url":"https://avatars.githubusercontent.com/u/14985020?v=4","primary_language":"TypeScript","stars":1319,"forks":95,"topics":["ai","artificial-intelligence","chatbot","claude","dall-e","embedding","gpt-3","huggingface","javascript","js","llamacpp","llm","mistral","multi-modal","ollama","openai","stable-diffusion","ts","typescript","whisper"],"archived":true,"github_pushed_at":"2024-07-19T15:17:19+00:00","maintenance_label":"Archived","url":"https://www.graphcanon.com/tools/vercel-modelfusion","markdown_url":"https://www.graphcanon.com/tools/vercel-modelfusion.md","api_url":"https://www.graphcanon.com/api/graphcanon/tools/vercel-modelfusion","graph_url":"https://www.graphcanon.com/api/graphcanon/graph?tool=vercel-modelfusion","shared_categories":[]},{"slug":"kubeai-project-kubeai","name":"kubeai","tagline":"AI Inference Operator for Kubernetes","github_url":"https://github.com/kubeai-project/kubeai","owner":"kubeai-project","repo":"kubeai","owner_avatar_url":"https://avatars.githubusercontent.com/u/232319222?v=4","primary_language":"Go","stars":1237,"forks":131,"topics":["ai","autoscaler","faster-whisper","inference-operator","k8s","kubernetes","llm","ollama","ollama-operator","openai-api","vllm","vllm-operator","whisper"],"archived":false,"github_pushed_at":"2026-07-31T01:04:47+00:00","maintenance_label":"Active","url":"https://www.graphcanon.com/tools/kubeai-project-kubeai","markdown_url":"https://www.graphcanon.com/tools/kubeai-project-kubeai.md","api_url":"https://www.graphcanon.com/api/graphcanon/tools/kubeai-project-kubeai","graph_url":"https://www.graphcanon.com/api/graphcanon/graph?tool=kubeai-project-kubeai","shared_categories":["inference-serving"]}]}}