{"data":{"node":{"slug":"writesonic-gptrouter","name":"GPTRouter","tagline":"Manage multiple LLMs and image models for reliable and fast responses","github_url":"https://github.com/Writesonic/GPTRouter","owner":"Writesonic","repo":"GPTRouter","owner_avatar_url":"https://avatars.githubusercontent.com/u/83169781?v=4","primary_language":"TypeScript","stars":455,"forks":38,"topics":["anthropic","azure-openai","cohere","google-gemini","langchain","llama-index","llm","llmops","llms","mlops","openai","palm-api"],"archived":false,"github_pushed_at":"2024-04-10T11:14:59+00:00","maintenance_label":"Dormant","stars_delta_30d":0,"url":"https://www.graphcanon.com/tools/writesonic-gptrouter","markdown_url":"https://www.graphcanon.com/tools/writesonic-gptrouter.md","api_url":"https://www.graphcanon.com/api/graphcanon/tools/writesonic-gptrouter","graph_url":"https://www.graphcanon.com/api/graphcanon/graph?tool=writesonic-gptrouter"},"categories":[{"slug":"inference-serving","name":"Inference & Serving","url":"https://www.graphcanon.com/categories/inference-serving","markdown_url":"https://www.graphcanon.com/categories/inference-serving.md","api_url":"https://www.graphcanon.com/api/graphcanon/categories/inference-serving"},{"slug":"llm-frameworks","name":"LLM Frameworks","url":"https://www.graphcanon.com/categories/llm-frameworks","markdown_url":"https://www.graphcanon.com/categories/llm-frameworks.md","api_url":"https://www.graphcanon.com/api/graphcanon/categories/llm-frameworks"},{"slug":"model-training","name":"Model Training","url":"https://www.graphcanon.com/categories/model-training","markdown_url":"https://www.graphcanon.com/categories/model-training.md","api_url":"https://www.graphcanon.com/api/graphcanon/categories/model-training"}],"tags":[{"slug":"anthropic","name":"anthropic"},{"slug":"azure-openai","name":"azure-openai"},{"slug":"cohere","name":"cohere"},{"slug":"google-gemini","name":"google-gemini"},{"slug":"langchain","name":"langchain"},{"slug":"llama-index","name":"llama-index"},{"slug":"llmops","name":"llmops"},{"slug":"mlops","name":"mlops"}],"edges":[{"type":"alternative","direction":"out","explanation":"Both GPTRouter and adaline-gateway aim to facilitate calling multiple LLMs by providing an integrated interface or SDK.","successor_context":null,"tool":{"slug":"adaline-gateway","name":"gateway","tagline":"Unified SDK for calling over 200 LLMs","github_url":"https://github.com/adaline/gateway","owner":"adaline","repo":"gateway","owner_avatar_url":"https://avatars.githubusercontent.com/u/382430?v=4","primary_language":"TypeScript","stars":605,"forks":26,"topics":["ai","ai-agents","anthropic","language-model","llm","llmops","openai","prompt-engineering","togetherai","typescript"],"archived":false,"github_pushed_at":"2026-07-29T19:11:29+00:00","maintenance_label":"Active","stars_delta_30d":4,"url":"https://www.graphcanon.com/tools/adaline-gateway","markdown_url":"https://www.graphcanon.com/tools/adaline-gateway.md","api_url":"https://www.graphcanon.com/api/graphcanon/tools/adaline-gateway","graph_url":"https://www.graphcanon.com/api/graphcanon/graph?tool=adaline-gateway"}},{"type":"integrates_with","direction":"out","explanation":"GPTRouter can integrate with LangChain's agent-engineering capabilities, providing a robust platform for managing multiple LLMs and ensuring reliability.","successor_context":null,"tool":{"slug":"langchain-ai-langchain","name":"langchain","tagline":"The agent engineering platform.","github_url":"https://github.com/langchain-ai/langchain","owner":"langchain-ai","repo":"langchain","owner_avatar_url":"https://avatars.githubusercontent.com/u/126733545?v=4","primary_language":"Python","stars":143615,"forks":23930,"topics":["agents","ai","ai-agents","anthropic","chatgpt","deepagents","enterprise","framework","gemini","generative-ai","langchain","langgraph","llm","multiagent","open-source","openai","pydantic","python","rag","typescript"],"archived":false,"github_pushed_at":"2026-08-07T08:27:07+00:00","maintenance_label":"Very active","stars_delta_30d":2337,"url":"https://www.graphcanon.com/tools/langchain-ai-langchain","markdown_url":"https://www.graphcanon.com/tools/langchain-ai-langchain.md","api_url":"https://www.graphcanon.com/api/graphcanon/tools/langchain-ai-langchain","graph_url":"https://www.graphcanon.com/api/graphcanon/graph?tool=langchain-ai-langchain"}},{"type":"integrates_with","direction":"out","explanation":"GPTRouter integrates with headroom to optimize the data flow and efficiency between GPTRouter's managed language models and the actual AI processing. Headroom compresses inputs from various sources, which helps in reducing token consumption without affecting output quality, thereby improving performance and cost-effectiveness for services handled by GPTRouter.","successor_context":null,"tool":{"slug":"headroomlabs-ai-headroom","name":"headroom","tagline":"Compress tool outputs and data to reduce tokens before reaching the LLM.","github_url":"https://github.com/headroomlabs-ai/headroom","owner":"headroomlabs-ai","repo":"headroom","owner_avatar_url":"https://avatars.githubusercontent.com/u/294291659?v=4","primary_language":"Python","stars":66470,"forks":5103,"topics":["agent","ai","anthropic","claude-code","compression","context-engineering","context-window","cursor","fastapi","langchain","llm","mcp","openai","prompt-engineering","proxy","python","rag","token-optimization","tokens","typescript"],"archived":false,"github_pushed_at":"2026-08-16T00:57:42+00:00","maintenance_label":"Very active","stars_delta_30d":6941,"url":"https://www.graphcanon.com/tools/headroomlabs-ai-headroom","markdown_url":"https://www.graphcanon.com/tools/headroomlabs-ai-headroom.md","api_url":"https://www.graphcanon.com/api/graphcanon/tools/headroomlabs-ai-headroom","graph_url":"https://www.graphcanon.com/api/graphcanon/graph?tool=headroomlabs-ai-headroom"}},{"type":"alternative","direction":"out","explanation":"GPTRouter and LiteLLM both serve as gateways for connecting to various language models (LLMs), offering fallbacks and performance optimizations.","successor_context":null,"tool":{"slug":"berriai-litellm","name":"litellm","tagline":"Python SDK and Proxy Server for calling multiple LLM APIs","github_url":"https://github.com/BerriAI/litellm","owner":"BerriAI","repo":"litellm","owner_avatar_url":"https://avatars.githubusercontent.com/u/121462774?v=4","primary_language":"Python","stars":55221,"forks":10231,"topics":["ai-gateway","anthropic","azure-openai","bedrock","gateway","langchain","litellm","llm","llm-gateway","llmops","mcp-gateway","openai","openai-proxy","rust","rust-ai","vertex-ai"],"archived":false,"github_pushed_at":"2026-08-01T05:53:28+00:00","maintenance_label":"Very active","url":"https://www.graphcanon.com/tools/berriai-litellm","markdown_url":"https://www.graphcanon.com/tools/berriai-litellm.md","api_url":"https://www.graphcanon.com/api/graphcanon/tools/berriai-litellm","graph_url":"https://www.graphcanon.com/api/graphcanon/graph?tool=berriai-litellm"}},{"type":"integrates_with","direction":"out","explanation":"Quivr focuses on integrating GenAI into applications through RAG (Retrieval-Augmented Generation), which can be complemented by GPTRouter's management of multiple LLMs to ensure reliability and speed.","successor_context":null,"tool":{"slug":"quivrhq-quivr","name":"quivr","tagline":"Opiniated RAG for integrating GenAI in your apps 🧠","github_url":"https://github.com/QuivrHQ/quivr","owner":"QuivrHQ","repo":"quivr","owner_avatar_url":"https://avatars.githubusercontent.com/u/159330290?v=4","primary_language":"Python","stars":39401,"forks":3727,"topics":["ai","api","chatbot","chatgpt","database","docker","framework","frontend","groq","html","javascript","llm","openai","postgresql","privacy","rag","react","security","typescript","vector"],"archived":false,"github_pushed_at":"2025-07-09T12:55:23+00:00","maintenance_label":"Dormant","stars_delta_30d":188,"url":"https://www.graphcanon.com/tools/quivrhq-quivr","markdown_url":"https://www.graphcanon.com/tools/quivrhq-quivr.md","api_url":"https://www.graphcanon.com/api/graphcanon/tools/quivrhq-quivr","graph_url":"https://www.graphcanon.com/api/graphcanon/graph?tool=quivrhq-quivr"}},{"type":"alternative","direction":"out","explanation":"Both GPTRouter and 9router provide a routing service for managing different AI models to improve the reliability and efficiency of LLM-based applications.","successor_context":null,"tool":{"slug":"decolua-9router","name":"9router","tagline":"Unlimited FREE AI coding with auto-fallback and token savings","github_url":"https://github.com/decolua/9router","owner":"decolua","repo":"9router","owner_avatar_url":"https://avatars.githubusercontent.com/u/8282593?v=4","primary_language":"JavaScript","stars":25841,"forks":4635,"topics":["ai-agents","ai-gateway","anthropic","chatgpt","claude","claude-code","cline","codex","copilot","cursor","deepseek","free-ai","gemini","gemini-cli","llm","llm-gateway","openai","openai-proxy","qwen","token-saver"],"archived":false,"github_pushed_at":"2026-08-14T10:08:34+00:00","maintenance_label":"Very active","stars_delta_30d":3072,"url":"https://www.graphcanon.com/tools/decolua-9router","markdown_url":"https://www.graphcanon.com/tools/decolua-9router.md","api_url":"https://www.graphcanon.com/api/graphcanon/tools/decolua-9router","graph_url":"https://www.graphcanon.com/api/graphcanon/graph?tool=decolua-9router"}},{"type":"integrates_with","direction":"out","explanation":"Autogen creates multi-agent AI applications, and integrating with GPTRouter can provide enhanced reliability by managing fallback models for seamless operation of agents.","successor_context":null,"tool":{"slug":"microsoft-autogen","name":"autogen","tagline":"A programming framework for agentic AI","github_url":"https://github.com/microsoft/autogen","owner":"microsoft","repo":"autogen","owner_avatar_url":"https://avatars.githubusercontent.com/u/6154722?v=4","primary_language":"Python","stars":60139,"forks":9059,"topics":["agentic","agentic-agi","agents","ai","autogen","autogen-ecosystem","chatgpt","framework","llm-agent","llm-framework"],"archived":false,"github_pushed_at":"2026-04-15T11:59:09+00:00","maintenance_label":"Slowing","url":"https://www.graphcanon.com/tools/microsoft-autogen","markdown_url":"https://www.graphcanon.com/tools/microsoft-autogen.md","api_url":"https://www.graphcanon.com/api/graphcanon/tools/microsoft-autogen","graph_url":"https://www.graphcanon.com/api/graphcanon/graph?tool=microsoft-autogen"}},{"type":"related","direction":"out","explanation":null,"successor_context":null,"tool":{"slug":"open-webui-open-webui","name":"open-webui","tagline":"User-friendly AI Interface (Supports Ollama, OpenAI API, ...)","github_url":"https://github.com/open-webui/open-webui","owner":"open-webui","repo":"open-webui","owner_avatar_url":"https://avatars.githubusercontent.com/u/158137808?v=4","primary_language":"Python","stars":148875,"forks":21676,"topics":["ai","llm","llm-ui","llm-webui","llms","mcp","ollama","ollama-webui","open-webui","openai","openapi","rag","self-hosted","ui","webui"],"archived":false,"github_pushed_at":"2026-08-15T07:10:16+00:00","maintenance_label":"Very active","stars_delta_30d":3224,"url":"https://www.graphcanon.com/tools/open-webui-open-webui","markdown_url":"https://www.graphcanon.com/tools/open-webui-open-webui.md","api_url":"https://www.graphcanon.com/api/graphcanon/tools/open-webui-open-webui","graph_url":"https://www.graphcanon.com/api/graphcanon/graph?tool=open-webui-open-webui"}},{"type":"alternative","direction":"in","explanation":"Both Portkey Gateway and GPTRouter provide similar functionality as gateways enabling interaction with multiple LLMs through a single API or interface.","successor_context":null,"tool":{"slug":"portkey-ai-gateway","name":"gateway","tagline":"A high-performance AI Gateway connecting to over 1,600 LLMs with guardrails.","github_url":"https://github.com/Portkey-AI/gateway","owner":"Portkey-AI","repo":"gateway","owner_avatar_url":"https://avatars.githubusercontent.com/u/131141116?v=4","primary_language":"TypeScript","stars":12668,"forks":1236,"topics":["ai-gateway","gateway","generative-ai","hacktoberfest","langchain","llm","llm-gateway","llmops","llms","mcp","mcp-client","mcp-gateway","mcp-servers","model-router","openai"],"archived":false,"github_pushed_at":"2026-05-25T13:54:51+00:00","maintenance_label":"Steady","stars_delta_30d":314,"url":"https://www.graphcanon.com/tools/portkey-ai-gateway","markdown_url":"https://www.graphcanon.com/tools/portkey-ai-gateway.md","api_url":"https://www.graphcanon.com/api/graphcanon/tools/portkey-ai-gateway","graph_url":"https://www.graphcanon.com/api/graphcanon/graph?tool=portkey-ai-gateway"}},{"type":"alternative","direction":"in","explanation":"Both 9Router and GPTRouter act as gateways to multiple LLM providers, offering services such as optimizing costs and managing fallbacks between different models or providers.","successor_context":null,"tool":{"slug":"decolua-9router","name":"9router","tagline":"Unlimited FREE AI coding with auto-fallback and token savings","github_url":"https://github.com/decolua/9router","owner":"decolua","repo":"9router","owner_avatar_url":"https://avatars.githubusercontent.com/u/8282593?v=4","primary_language":"JavaScript","stars":25841,"forks":4635,"topics":["ai-agents","ai-gateway","anthropic","chatgpt","claude","claude-code","cline","codex","copilot","cursor","deepseek","free-ai","gemini","gemini-cli","llm","llm-gateway","openai","openai-proxy","qwen","token-saver"],"archived":false,"github_pushed_at":"2026-08-14T10:08:34+00:00","maintenance_label":"Very active","stars_delta_30d":3072,"url":"https://www.graphcanon.com/tools/decolua-9router","markdown_url":"https://www.graphcanon.com/tools/decolua-9router.md","api_url":"https://www.graphcanon.com/api/graphcanon/tools/decolua-9router","graph_url":"https://www.graphcanon.com/api/graphcanon/graph?tool=decolua-9router"}}],"neighbours":[{"slug":"patchy631-ai-engineering-hub","name":"ai-engineering-hub","tagline":"Tutorials on LLMs, RAGs, and real-world AI agent applications","github_url":"https://github.com/patchy631/ai-engineering-hub","owner":"patchy631","repo":"ai-engineering-hub","owner_avatar_url":"https://avatars.githubusercontent.com/u/38653995?v=4","primary_language":"Jupyter Notebook","stars":37020,"forks":6107,"topics":["agents","ai","llms","machine-learning","mcp","rag"],"archived":false,"github_pushed_at":"2026-07-27T18:43:06+00:00","maintenance_label":"Active","url":"https://www.graphcanon.com/tools/patchy631-ai-engineering-hub","markdown_url":"https://www.graphcanon.com/tools/patchy631-ai-engineering-hub.md","api_url":"https://www.graphcanon.com/api/graphcanon/tools/patchy631-ai-engineering-hub","graph_url":"https://www.graphcanon.com/api/graphcanon/graph?tool=patchy631-ai-engineering-hub","shared_categories":["llm-frameworks"]},{"slug":"eugeneyan-open-llms","name":"open-llms","tagline":"A list of open LLMs available for commercial use.","github_url":"https://github.com/eugeneyan/open-llms","owner":"eugeneyan","repo":"open-llms","owner_avatar_url":"https://avatars.githubusercontent.com/u/6831355?v=4","primary_language":null,"stars":12849,"forks":985,"topics":["commercial","large-language-models","llm","llms"],"archived":false,"github_pushed_at":"2025-02-13T06:37:12+00:00","maintenance_label":"Dormant","url":"https://www.graphcanon.com/tools/eugeneyan-open-llms","markdown_url":"https://www.graphcanon.com/tools/eugeneyan-open-llms.md","api_url":"https://www.graphcanon.com/api/graphcanon/tools/eugeneyan-open-llms","graph_url":"https://www.graphcanon.com/api/graphcanon/graph?tool=eugeneyan-open-llms","shared_categories":["llm-frameworks"]},{"slug":"steven2358-awesome-generative-ai","name":"awesome-generative-ai","tagline":"A curated list of modern Generative Artificial Intelligence projects and services","github_url":"https://github.com/steven2358/awesome-generative-ai","owner":"steven2358","repo":"awesome-generative-ai","owner_avatar_url":"https://avatars.githubusercontent.com/u/164072?v=4","primary_language":null,"stars":12501,"forks":1990,"topics":["ai","artificial-intelligence","awesome","awesome-list","generative-ai","generative-art","large-language-models","llm"],"archived":false,"github_pushed_at":"2026-08-03T10:58:05+00:00","maintenance_label":"Active","url":"https://www.graphcanon.com/tools/steven2358-awesome-generative-ai","markdown_url":"https://www.graphcanon.com/tools/steven2358-awesome-generative-ai.md","api_url":"https://www.graphcanon.com/api/graphcanon/tools/steven2358-awesome-generative-ai","graph_url":"https://www.graphcanon.com/api/graphcanon/graph?tool=steven2358-awesome-generative-ai","shared_categories":["llm-frameworks","inference-serving"]},{"slug":"simonw-llm","name":"llm","tagline":"Access large language models from the command-line","github_url":"https://github.com/simonw/llm","owner":"simonw","repo":"llm","owner_avatar_url":"https://avatars.githubusercontent.com/u/9599?v=4","primary_language":"Python","stars":12324,"forks":939,"topics":["ai","llms","openai"],"archived":false,"github_pushed_at":"2026-08-05T14:29:09+00:00","maintenance_label":"Active","url":"https://www.graphcanon.com/tools/simonw-llm","markdown_url":"https://www.graphcanon.com/tools/simonw-llm.md","api_url":"https://www.graphcanon.com/api/graphcanon/tools/simonw-llm","graph_url":"https://www.graphcanon.com/api/graphcanon/graph?tool=simonw-llm","shared_categories":["llm-frameworks","inference-serving"]},{"slug":"oumi-ai-oumi","name":"oumi","tagline":"Easily fine-tune, evaluate and deploy open source LLMs/VLMs","github_url":"https://github.com/oumi-ai/oumi","owner":"oumi-ai","repo":"oumi","owner_avatar_url":"https://avatars.githubusercontent.com/u/167452922?v=4","primary_language":"Python","stars":9359,"forks":780,"topics":["dpo","evaluation","fine-tuning","gpt-oss","gpt-oss-120b","gpt-oss-20b","inference","llama","llms","sft","slms","vlms"],"archived":false,"github_pushed_at":"2026-07-24T05:44:23+00:00","maintenance_label":"Active","url":"https://www.graphcanon.com/tools/oumi-ai-oumi","markdown_url":"https://www.graphcanon.com/tools/oumi-ai-oumi.md","api_url":"https://www.graphcanon.com/api/graphcanon/tools/oumi-ai-oumi","graph_url":"https://www.graphcanon.com/api/graphcanon/graph?tool=oumi-ai-oumi","shared_categories":["model-training","inference-serving"]},{"slug":"mnfst-awesome-free-llm-apis","name":"awesome-free-llm-apis","tagline":"List of Permanent Free LLM API","github_url":"https://github.com/mnfst/awesome-free-llm-apis","owner":"mnfst","repo":"awesome-free-llm-apis","owner_avatar_url":"https://avatars.githubusercontent.com/u/8403534?v=4","primary_language":"JavaScript","stars":6532,"forks":630,"topics":["ai-agents","anthropic","awesome","awesome-list","gemini","llm","llm-router","llm-routing","ollama","openai","openclaw","openclaw-plugin","router"],"archived":false,"github_pushed_at":"2026-07-30T15:05:10+00:00","maintenance_label":"Active","url":"https://www.graphcanon.com/tools/mnfst-awesome-free-llm-apis","markdown_url":"https://www.graphcanon.com/tools/mnfst-awesome-free-llm-apis.md","api_url":"https://www.graphcanon.com/api/graphcanon/tools/mnfst-awesome-free-llm-apis","graph_url":"https://www.graphcanon.com/api/graphcanon/graph?tool=mnfst-awesome-free-llm-apis","shared_categories":["inference-serving"]},{"slug":"andyyyy64-whichllm","name":"whichllm","tagline":"Command-line tool to find and benchmark local LLM performance","github_url":"https://github.com/Andyyyy64/whichllm","owner":"Andyyyy64","repo":"whichllm","owner_avatar_url":"https://avatars.githubusercontent.com/u/105579829?v=4","primary_language":"Python","stars":6225,"forks":330,"topics":["ai","apple-silicon","benchmarks","cli","command-line-tool","gguf","gpu","huggingface","inference","llm","local-llm","ollama","python","vram"],"archived":false,"github_pushed_at":"2026-08-05T07:15:32+00:00","maintenance_label":"Active","url":"https://www.graphcanon.com/tools/andyyyy64-whichllm","markdown_url":"https://www.graphcanon.com/tools/andyyyy64-whichllm.md","api_url":"https://www.graphcanon.com/api/graphcanon/tools/andyyyy64-whichllm","graph_url":"https://www.graphcanon.com/api/graphcanon/graph?tool=andyyyy64-whichllm","shared_categories":["inference-serving"]},{"slug":"nyldn-claude-octopus","name":"claude-octopus","tagline":"Surface AI blindspots before you ship","github_url":"https://github.com/nyldn/claude-octopus","owner":"nyldn","repo":"claude-octopus","owner_avatar_url":"https://avatars.githubusercontent.com/u/4805949?v=4","primary_language":"Shell","stars":3962,"forks":374,"topics":["ai-agents","ai-orchestration","claude-code","claude-code-plugin","codex","copilot","developer-tools","double-diamond","gemini","multi-ai","multi-llm","ollama"],"archived":false,"github_pushed_at":"2026-08-13T23:28:23+00:00","maintenance_label":"Active","url":"https://www.graphcanon.com/tools/nyldn-claude-octopus","markdown_url":"https://www.graphcanon.com/tools/nyldn-claude-octopus.md","api_url":"https://www.graphcanon.com/api/graphcanon/tools/nyldn-claude-octopus","graph_url":"https://www.graphcanon.com/api/graphcanon/graph?tool=nyldn-claude-octopus","shared_categories":[]},{"slug":"superlinked-sie","name":"sie","tagline":"Open-source inference server and production cluster for all the models your agent needs.","github_url":"https://github.com/superlinked/sie","owner":"superlinked","repo":"sie","owner_avatar_url":"https://avatars.githubusercontent.com/u/94243920?v=4","primary_language":"Python","stars":2804,"forks":272,"topics":["bge","colbert","data-pipeline","deep-learning","embeddings","inference","inference-server","information-retrieval","llm","ml","mlops","natural-language-processing","nlp","python","reranking","retrieval","retrieval-augmented-generation","semantic-search","splade","vector-search"],"archived":false,"github_pushed_at":"2026-08-21T20:28:04+00:00","maintenance_label":"Very active","url":"https://www.graphcanon.com/tools/superlinked-sie","markdown_url":"https://www.graphcanon.com/tools/superlinked-sie.md","api_url":"https://www.graphcanon.com/api/graphcanon/tools/superlinked-sie","graph_url":"https://www.graphcanon.com/api/graphcanon/graph?tool=superlinked-sie","shared_categories":["inference-serving"]},{"slug":"stochasticai-xturing","name":"xTuring","tagline":"Personalize and control open-source LLMs with ease","github_url":"https://github.com/stochasticai/xTuring","owner":"stochasticai","repo":"xTuring","owner_avatar_url":"https://avatars.githubusercontent.com/u/66399337?v=4","primary_language":"Python","stars":2670,"forks":210,"topics":["adapter","deep-learning","fine-tuning","finetuning","gen-ai","generative-ai","gpt-2","gpt-j","language-model","llama","llm","lora","mistral","mixed-precision","peft","quantization"],"archived":false,"github_pushed_at":"2026-03-04T23:07:06+00:00","maintenance_label":"Slowing","url":"https://www.graphcanon.com/tools/stochasticai-xturing","markdown_url":"https://www.graphcanon.com/tools/stochasticai-xturing.md","api_url":"https://www.graphcanon.com/api/graphcanon/tools/stochasticai-xturing","graph_url":"https://www.graphcanon.com/api/graphcanon/graph?tool=stochasticai-xturing","shared_categories":["model-training","llm-frameworks"]},{"slug":"theopenco-llmgateway","name":"llmgateway","tagline":"Route and manage LLM requests via unified API interface","github_url":"https://github.com/theopenco/llmgateway","owner":"theopenco","repo":"llmgateway","owner_avatar_url":"https://avatars.githubusercontent.com/u/211671860?v=4","primary_language":"TypeScript","stars":1518,"forks":170,"topics":["ai","ai-gateway","analytics","anthropic","api-key-management","claude","codex","enterprise","guardrails","inference","llm","llm-gateway","llm-proxy","llms","observability","openai","opencode","rate-limiting","typescript"],"archived":false,"github_pushed_at":"2026-08-09T00:38:36+00:00","maintenance_label":"Active","url":"https://www.graphcanon.com/tools/theopenco-llmgateway","markdown_url":"https://www.graphcanon.com/tools/theopenco-llmgateway.md","api_url":"https://www.graphcanon.com/api/graphcanon/tools/theopenco-llmgateway","graph_url":"https://www.graphcanon.com/api/graphcanon/graph?tool=theopenco-llmgateway","shared_categories":["inference-serving"]},{"slug":"openinfer-project-openinfer","name":"openinfer","tagline":"Pure Rust CUDA LLM inference engine serving multiple models including Qwen3 and Kimi-K2","github_url":"https://github.com/openinfer-project/openinfer","owner":"openinfer-project","repo":"openinfer","owner_avatar_url":"https://avatars.githubusercontent.com/u/292134277?v=4","primary_language":"Rust","stars":585,"forks":89,"topics":["cuda","cuda-kernels","deepseek","gpu","inference","inference-engine","kimi","kimi-k2","kv-cache","llm","llm-inference","llm-serving","model-serving","moe","openai-api","paged-attention","qwen","qwen3","rust","vllm"],"archived":false,"github_pushed_at":"2026-07-25T14:08:34+00:00","maintenance_label":"Active","url":"https://www.graphcanon.com/tools/openinfer-project-openinfer","markdown_url":"https://www.graphcanon.com/tools/openinfer-project-openinfer.md","api_url":"https://www.graphcanon.com/api/graphcanon/tools/openinfer-project-openinfer","graph_url":"https://www.graphcanon.com/api/graphcanon/graph?tool=openinfer-project-openinfer","shared_categories":["inference-serving"]}]}}