{"data":{"node":{"slug":"mohitsoni48-turbollm","name":"TurboLLM","tagline":"Run any local LLM engine auto-tuned to your GPU with polished web UI and OpenAI/Anthropic-compatible API","github_url":"https://github.com/mohitsoni48/TurboLLM","owner":"mohitsoni48","repo":"TurboLLM","owner_avatar_url":"https://avatars.githubusercontent.com/u/63787789?v=4","primary_language":"TypeScript","stars":274,"forks":38,"topics":["ai","anthropic-api","claude-code","gguf","gpu","inference","llama-cpp","llama-server","llm","local-llm","offline","openai-api","self-hosted"],"archived":false,"github_pushed_at":"2026-09-19T13:11:45+00:00","maintenance_label":"Very active","stars_delta_30d":49,"url":"https://www.graphcanon.com/tools/mohitsoni48-turbollm","markdown_url":"https://www.graphcanon.com/tools/mohitsoni48-turbollm.md","api_url":"https://www.graphcanon.com/api/graphcanon/tools/mohitsoni48-turbollm","graph_url":"https://www.graphcanon.com/api/graphcanon/graph?tool=mohitsoni48-turbollm"},"categories":[{"slug":"inference-serving","name":"Inference & Serving","url":"https://www.graphcanon.com/categories/inference-serving","markdown_url":"https://www.graphcanon.com/categories/inference-serving.md","api_url":"https://www.graphcanon.com/api/graphcanon/categories/inference-serving"},{"slug":"model-training","name":"Model Training","url":"https://www.graphcanon.com/categories/model-training","markdown_url":"https://www.graphcanon.com/categories/model-training.md","api_url":"https://www.graphcanon.com/api/graphcanon/categories/model-training"}],"tags":[{"slug":"ai","name":"ai"},{"slug":"anthropic-api","name":"anthropic-api"},{"slug":"claude-code","name":"claude-code"},{"slug":"gpu","name":"gpu"},{"slug":"inference","name":"inference"},{"slug":"llama-cpp","name":"llama-cpp"},{"slug":"local-llm","name":"local-llm"},{"slug":"openai-api","name":"openai-api"}],"edges":[{"type":"integrates_with","direction":"in","explanation":"`TurboLLM` can use `llama.cpp` as one of its engines for running LLM models, showing integration potential.","successor_context":null,"tool":{"slug":"ggml-org-llama-cpp","name":"llama.cpp","tagline":"LLM inference in C/C++","github_url":"https://github.com/ggml-org/llama.cpp","owner":"ggml-org","repo":"llama.cpp","owner_avatar_url":"https://avatars.githubusercontent.com/u/134263123?v=4","primary_language":"C++","stars":128639,"forks":23359,"topics":["ggml"],"archived":false,"github_pushed_at":"2026-09-18T08:13:37+00:00","maintenance_label":"Very active","stars_delta_30d":5698,"url":"https://www.graphcanon.com/tools/ggml-org-llama-cpp","markdown_url":"https://www.graphcanon.com/tools/ggml-org-llama-cpp.md","api_url":"https://www.graphcanon.com/api/graphcanon/tools/ggml-org-llama-cpp","graph_url":"https://www.graphcanon.com/api/graphcanon/graph?tool=ggml-org-llama-cpp"}},{"type":"alternative","direction":"in","explanation":"TurboLLM provides a polished web UI for running local LLMs with auto-tuning, similar to Open WebUI.","successor_context":null,"tool":{"slug":"open-webui-open-webui","name":"open-webui","tagline":"User-friendly AI Interface","github_url":"https://github.com/open-webui/open-webui","owner":"open-webui","repo":"open-webui","owner_avatar_url":"https://avatars.githubusercontent.com/u/158137808?v=4","primary_language":"Python","stars":152445,"forks":22306,"topics":["ai","llm","llm-ui","llm-webui","llms","mcp","ollama","ollama-webui","open-webui","openai","openapi","rag","self-hosted","ui","webui"],"archived":false,"github_pushed_at":"2026-09-18T00:11:06+00:00","maintenance_label":"Very active","stars_delta_30d":3570,"url":"https://www.graphcanon.com/tools/open-webui-open-webui","markdown_url":"https://www.graphcanon.com/tools/open-webui-open-webui.md","api_url":"https://www.graphcanon.com/api/graphcanon/tools/open-webui-open-webui","graph_url":"https://www.graphcanon.com/api/graphcanon/graph?tool=open-webui-open-webui"}},{"type":"alternative","direction":"in","explanation":"Both GPT4All and TurboLLM are focused on running LLMs locally with optimized API compatibility.","successor_context":null,"tool":{"slug":"nomic-ai-gpt4all","name":"gpt4all","tagline":"Run Local LLMs on Any Device","github_url":"https://github.com/nomic-ai/gpt4all","owner":"nomic-ai","repo":"gpt4all","owner_avatar_url":"https://avatars.githubusercontent.com/u/102670180?v=4","primary_language":"C++","stars":77390,"forks":8288,"topics":["ai-chat","llm-inference"],"archived":false,"github_pushed_at":"2025-05-27T20:05:19+00:00","maintenance_label":"Dormant","stars_delta_30d":-6,"url":"https://www.graphcanon.com/tools/nomic-ai-gpt4all","markdown_url":"https://www.graphcanon.com/tools/nomic-ai-gpt4all.md","api_url":"https://www.graphcanon.com/api/graphcanon/tools/nomic-ai-gpt4all","graph_url":"https://www.graphcanon.com/api/graphcanon/graph?tool=nomic-ai-gpt4all"}},{"type":"alternative","direction":"in","explanation":"Both PrivateGPT and TurboLLM offer solutions to run local LLM engines on auto-tuned GPUs with web UIs, placing them as alternatives for similar needs.","successor_context":null,"tool":{"slug":"zylon-ai-private-gpt","name":"private-gpt","tagline":"Complete API layer for private AI applications on local models","github_url":"https://github.com/zylon-ai/private-gpt","owner":"zylon-ai","repo":"private-gpt","owner_avatar_url":"https://avatars.githubusercontent.com/u/143802295?v=4","primary_language":"Python","stars":57497,"forks":7615,"topics":["ai","ai-tools","on-premise"],"archived":false,"github_pushed_at":"2026-09-07T15:58:50+00:00","maintenance_label":"Very active","stars_delta_30d":82,"url":"https://www.graphcanon.com/tools/zylon-ai-private-gpt","markdown_url":"https://www.graphcanon.com/tools/zylon-ai-private-gpt.md","api_url":"https://www.graphcanon.com/api/graphcanon/tools/zylon-ai-private-gpt","graph_url":"https://www.graphcanon.com/api/graphcanon/graph?tool=zylon-ai-private-gpt"}}],"neighbours":[{"slug":"gaizhenbiao-chuanhuchatgpt","name":"ChuanhuChatGPT","tagline":"GUI for ChatGPT and LLMs with features like agents and file-based QA.","github_url":"https://github.com/GaiZhenbiao/ChuanhuChatGPT","owner":"GaiZhenbiao","repo":"ChuanhuChatGPT","owner_avatar_url":"https://avatars.githubusercontent.com/u/51039745?v=4","primary_language":"Python","stars":15272,"forks":2195,"topics":["chatbot","chatglm","chatgpt-api","claude","dalle3","ernie","gemini","gemma","inspurai","llama","midjourney","minimax","moss","ollama","qwen","spark","stablelm"],"archived":false,"github_pushed_at":"2026-09-16T11:44:01+00:00","maintenance_label":"Very active","url":"https://www.graphcanon.com/tools/gaizhenbiao-chuanhuchatgpt","markdown_url":"https://www.graphcanon.com/tools/gaizhenbiao-chuanhuchatgpt.md","api_url":"https://www.graphcanon.com/api/graphcanon/tools/gaizhenbiao-chuanhuchatgpt","graph_url":"https://www.graphcanon.com/api/graphcanon/graph?tool=gaizhenbiao-chuanhuchatgpt","shared_categories":["inference-serving"]},{"slug":"simonw-llm","name":"llm","tagline":"Access large language models from the command-line","github_url":"https://github.com/simonw/llm","owner":"simonw","repo":"llm","owner_avatar_url":"https://avatars.githubusercontent.com/u/9599?v=4","primary_language":"Python","stars":12473,"forks":978,"topics":["ai","llms","openai"],"archived":false,"github_pushed_at":"2026-09-02T19:24:35+00:00","maintenance_label":"Active","url":"https://www.graphcanon.com/tools/simonw-llm","markdown_url":"https://www.graphcanon.com/tools/simonw-llm.md","api_url":"https://www.graphcanon.com/api/graphcanon/tools/simonw-llm","graph_url":"https://www.graphcanon.com/api/graphcanon/graph?tool=simonw-llm","shared_categories":["inference-serving"]},{"slug":"oumi-ai-oumi","name":"oumi","tagline":"Easily fine-tune, evaluate and deploy open source LLMs/VLMs","github_url":"https://github.com/oumi-ai/oumi","owner":"oumi-ai","repo":"oumi","owner_avatar_url":"https://avatars.githubusercontent.com/u/167452922?v=4","primary_language":"Python","stars":9387,"forks":790,"topics":["dpo","evaluation","fine-tuning","gpt-oss","gpt-oss-120b","gpt-oss-20b","inference","llama","llms","open-weight","open-weight-models","open-weights","sft","slms","vlms"],"archived":false,"github_pushed_at":"2026-09-20T05:45:37+00:00","maintenance_label":"Very active","url":"https://www.graphcanon.com/tools/oumi-ai-oumi","markdown_url":"https://www.graphcanon.com/tools/oumi-ai-oumi.md","api_url":"https://www.graphcanon.com/api/graphcanon/tools/oumi-ai-oumi","graph_url":"https://www.graphcanon.com/api/graphcanon/graph?tool=oumi-ai-oumi","shared_categories":["model-training","inference-serving"]},{"slug":"andyyyy64-whichllm","name":"whichllm","tagline":"Command-line tool to find and benchmark local LLM performance","github_url":"https://github.com/Andyyyy64/whichllm","owner":"Andyyyy64","repo":"whichllm","owner_avatar_url":"https://avatars.githubusercontent.com/u/105579829?v=4","primary_language":"Python","stars":6666,"forks":368,"topics":["localllm"],"archived":false,"github_pushed_at":"2026-09-19T16:22:49+00:00","maintenance_label":"Very active","url":"https://www.graphcanon.com/tools/andyyyy64-whichllm","markdown_url":"https://www.graphcanon.com/tools/andyyyy64-whichllm.md","api_url":"https://www.graphcanon.com/api/graphcanon/tools/andyyyy64-whichllm","graph_url":"https://www.graphcanon.com/api/graphcanon/graph?tool=andyyyy64-whichllm","shared_categories":["inference-serving"]},{"slug":"b4rtaz-distributed-llama","name":"distributed-llama","tagline":"Distributed LLM inference using home devices cluster","github_url":"https://github.com/b4rtaz/distributed-llama","owner":"b4rtaz","repo":"distributed-llama","owner_avatar_url":"https://avatars.githubusercontent.com/u/12797776?v=4","primary_language":"C++","stars":3060,"forks":250,"topics":["distributed-computing","distributed-llm","llama2","llama3","llm","llm-inference","llms","neural-network","open-llm"],"archived":false,"github_pushed_at":"2026-07-05T16:47:20+00:00","maintenance_label":"Steady","url":"https://www.graphcanon.com/tools/b4rtaz-distributed-llama","markdown_url":"https://www.graphcanon.com/tools/b4rtaz-distributed-llama.md","api_url":"https://www.graphcanon.com/api/graphcanon/tools/b4rtaz-distributed-llama","graph_url":"https://www.graphcanon.com/api/graphcanon/graph?tool=b4rtaz-distributed-llama","shared_categories":["inference-serving"]},{"slug":"stochasticai-xturing","name":"xTuring","tagline":"Build, personalize and control your own LLMs","github_url":"https://github.com/stochasticai/xTuring","owner":"stochasticai","repo":"xTuring","owner_avatar_url":"https://avatars.githubusercontent.com/u/66399337?v=4","primary_language":"Python","stars":2674,"forks":211,"topics":["adapter","deep-learning","fine-tuning","finetuning","gen-ai","generative-ai","gpt-2","gpt-j","language-model","llama","llm","lora","mistral","mixed-precision","peft","quantization"],"archived":false,"github_pushed_at":"2026-09-12T23:50:56+00:00","maintenance_label":"Active","url":"https://www.graphcanon.com/tools/stochasticai-xturing","markdown_url":"https://www.graphcanon.com/tools/stochasticai-xturing.md","api_url":"https://www.graphcanon.com/api/graphcanon/tools/stochasticai-xturing","graph_url":"https://www.graphcanon.com/api/graphcanon/graph?tool=stochasticai-xturing","shared_categories":["model-training"]},{"slug":"theopenco-llmgateway","name":"llmgateway","tagline":"Route and manage LLM requests via unified API interface","github_url":"https://github.com/theopenco/llmgateway","owner":"theopenco","repo":"llmgateway","owner_avatar_url":"https://avatars.githubusercontent.com/u/211671860?v=4","primary_language":"TypeScript","stars":1624,"forks":181,"topics":["ai","ai-gateway","analytics","anthropic","api-key-management","claude","codex","enterprise","guardrails","inference","llm","llm-gateway","llm-proxy","llms","observability","openai","opencode","rate-limiting","typescript"],"archived":false,"github_pushed_at":"2026-09-11T06:00:04+00:00","maintenance_label":"Active","url":"https://www.graphcanon.com/tools/theopenco-llmgateway","markdown_url":"https://www.graphcanon.com/tools/theopenco-llmgateway.md","api_url":"https://www.graphcanon.com/api/graphcanon/tools/theopenco-llmgateway","graph_url":"https://www.graphcanon.com/api/graphcanon/graph?tool=theopenco-llmgateway","shared_categories":["inference-serving"]},{"slug":"waybarrios-vllm-mlx","name":"vllm-mlx","tagline":"Server for LLMs and vision-language models compatible with Apple Silicon","github_url":"https://github.com/waybarrios/vllm-mlx","owner":"waybarrios","repo":"vllm-mlx","owner_avatar_url":"https://avatars.githubusercontent.com/u/6794828?v=4","primary_language":"Python","stars":1588,"forks":222,"topics":["anthropic","anthropic-api","apple-silicon","claude-code","continuous-batching","inference-server","llm","local-llm","macos","mcp","mlx","multimodal-ai","openai","openai-api","openai-compatible","speech-to-text","text-to-speech","tool-calling","vision-language-model","vllm"],"archived":false,"github_pushed_at":"2026-09-19T18:11:07+00:00","maintenance_label":"Very active","url":"https://www.graphcanon.com/tools/waybarrios-vllm-mlx","markdown_url":"https://www.graphcanon.com/tools/waybarrios-vllm-mlx.md","api_url":"https://www.graphcanon.com/api/graphcanon/tools/waybarrios-vllm-mlx","graph_url":"https://www.graphcanon.com/api/graphcanon/graph?tool=waybarrios-vllm-mlx","shared_categories":["model-training","inference-serving"]},{"slug":"atomicbot-ai-atomic-chat","name":"Atomic-Chat","tagline":"Local AI app and inference engine for agents","github_url":"https://github.com/AtomicBot-ai/Atomic-Chat","owner":"AtomicBot-ai","repo":"Atomic-Chat","owner_avatar_url":"https://avatars.githubusercontent.com/u/259533419?v=4","primary_language":"TypeScript","stars":1526,"forks":178,"topics":["ai-chat","ai-tools","apple-silicon","chatgpt","deepseek","desktop-app","gemma","gguf","gpt-oss","llamacpp","llm","llm-inference","local-ai","local-first","local-llm","mcp","mlx","open-source","qwen","self-hosted"],"archived":false,"github_pushed_at":"2026-09-19T07:46:33+00:00","maintenance_label":"Very active","url":"https://www.graphcanon.com/tools/atomicbot-ai-atomic-chat","markdown_url":"https://www.graphcanon.com/tools/atomicbot-ai-atomic-chat.md","api_url":"https://www.graphcanon.com/api/graphcanon/tools/atomicbot-ai-atomic-chat","graph_url":"https://www.graphcanon.com/api/graphcanon/graph?tool=atomicbot-ai-atomic-chat","shared_categories":["inference-serving"]},{"slug":"sauravpanda-browserai","name":"BrowserAI","tagline":"Run local LLMs like llama, deepseek-distill, kokoro and more inside your browser","github_url":"https://github.com/sauravpanda/BrowserAI","owner":"sauravpanda","repo":"BrowserAI","owner_avatar_url":"https://avatars.githubusercontent.com/u/12201824?v=4","primary_language":"TypeScript","stars":1451,"forks":136,"topics":["agents","ai","llama","llm","llm-inference","local","localllm","tts","webgpu"],"archived":false,"github_pushed_at":"2026-07-21T02:47:31+00:00","maintenance_label":"Steady","url":"https://www.graphcanon.com/tools/sauravpanda-browserai","markdown_url":"https://www.graphcanon.com/tools/sauravpanda-browserai.md","api_url":"https://www.graphcanon.com/api/graphcanon/tools/sauravpanda-browserai","graph_url":"https://www.graphcanon.com/api/graphcanon/graph?tool=sauravpanda-browserai","shared_categories":["inference-serving"]},{"slug":"georgian-io-llm-finetuning-toolkit","name":"LLM-Finetuning-Toolkit","tagline":"Toolkit for fine-tuning and testing open-source large language models","github_url":"https://github.com/georgian-io/LLM-Finetuning-Toolkit","owner":"georgian-io","repo":"LLM-Finetuning-Toolkit","owner_avatar_url":"https://avatars.githubusercontent.com/u/10764713?v=4","primary_language":"Python","stars":870,"forks":107,"topics":["ablation-study","classification","falcon","fine-tuning","finetuning","flan-t5","large-language-models","llama2","llm-test","lora","mistral-7b","nlp","nlp-machine-learning","qlora","redpajama","summarization","unit-testing","zephyr"],"archived":false,"github_pushed_at":"2026-05-04T16:33:40+00:00","maintenance_label":"Slowing","url":"https://www.graphcanon.com/tools/georgian-io-llm-finetuning-toolkit","markdown_url":"https://www.graphcanon.com/tools/georgian-io-llm-finetuning-toolkit.md","api_url":"https://www.graphcanon.com/api/graphcanon/tools/georgian-io-llm-finetuning-toolkit","graph_url":"https://www.graphcanon.com/api/graphcanon/graph?tool=georgian-io-llm-finetuning-toolkit","shared_categories":["model-training"]},{"slug":"openinfer-project-openinfer","name":"openinfer","tagline":"Pure Rust CUDA LLM inference engine serving multiple models including Qwen3 and Kimi-K2","github_url":"https://github.com/openinfer-project/openinfer","owner":"openinfer-project","repo":"openinfer","owner_avatar_url":"https://avatars.githubusercontent.com/u/292134277?v=4","primary_language":"Rust","stars":705,"forks":107,"topics":["cuda","cuda-kernels","deepseek","gpu","inference","inference-engine","kimi","kimi-k2","kv-cache","llm","llm-inference","llm-serving","model-serving","moe","openai-api","paged-attention","qwen","qwen3","rust","vllm"],"archived":false,"github_pushed_at":"2026-09-18T17:57:56+00:00","maintenance_label":"Very active","url":"https://www.graphcanon.com/tools/openinfer-project-openinfer","markdown_url":"https://www.graphcanon.com/tools/openinfer-project-openinfer.md","api_url":"https://www.graphcanon.com/api/graphcanon/tools/openinfer-project-openinfer","graph_url":"https://www.graphcanon.com/api/graphcanon/graph?tool=openinfer-project-openinfer","shared_categories":["inference-serving"]}]}}