{"data":{"node":{"slug":"sgl-project-sglang","name":"sglang","tagline":"High-performance serving framework for large language and multimodal models","github_url":"https://github.com/sgl-project/sglang","owner":"sgl-project","repo":"sglang","owner_avatar_url":"https://avatars.githubusercontent.com/u/147780389?v=4","primary_language":"Python","stars":31454,"forks":7720,"topics":["attention","blackwell","cuda","deepseek","diffusion","glm","gpt-oss","inference","llama","llm","minimax","moe","qwen","qwen-image","reinforcement-learning","transformer","vlm","wan"],"archived":false,"github_pushed_at":"2026-08-07T06:00:20+00:00","maintenance_label":"Very active","stars_delta_30d":1409,"url":"https://www.graphcanon.com/tools/sgl-project-sglang","markdown_url":"https://www.graphcanon.com/tools/sgl-project-sglang.md","api_url":"https://www.graphcanon.com/api/graphcanon/tools/sgl-project-sglang","graph_url":"https://www.graphcanon.com/api/graphcanon/graph?tool=sgl-project-sglang"},"categories":[{"slug":"inference-serving","name":"Inference & Serving","url":"https://www.graphcanon.com/categories/inference-serving","markdown_url":"https://www.graphcanon.com/categories/inference-serving.md","api_url":"https://www.graphcanon.com/api/graphcanon/categories/inference-serving"}],"tags":[{"slug":"attention","name":"attention"},{"slug":"cuda","name":"cuda"},{"slug":"diffusion","name":"diffusion"},{"slug":"inference","name":"inference"},{"slug":"llm","name":"llm"},{"slug":"moe","name":"moe"},{"slug":"reinforcement-learning","name":"reinforcement-learning"},{"slug":"transformer","name":"transformer"}],"edges":[{"type":"alternative","direction":"out","explanation":"SGLang and ollama both serve as frameworks for handling large language models for inference. However, they aim to solve this problem through different underlying mechanisms and optimizations.","successor_context":null,"tool":{"slug":"ollama-ollama","name":"ollama","tagline":"Get up and running with various large language models using Ollama.","github_url":"https://github.com/ollama/ollama","owner":"ollama","repo":"ollama","owner_avatar_url":"https://avatars.githubusercontent.com/u/151674099?v=4","primary_language":"Go","stars":177524,"forks":17229,"topics":["deepseek","gemma","gemma3","glm","go","golang","gpt-oss","llama","llama3","llm","llms","minimax","mistral","ollama","qwen"],"archived":false,"github_pushed_at":"2026-07-31T23:59:29+00:00","maintenance_label":"Very active","url":"https://www.graphcanon.com/tools/ollama-ollama","markdown_url":"https://www.graphcanon.com/tools/ollama-ollama.md","api_url":"https://www.graphcanon.com/api/graphcanon/tools/ollama-ollama","graph_url":"https://www.graphcanon.com/api/graphcanon/graph?tool=ollama-ollama"}},{"type":"alternative","direction":"out","explanation":"SGLang and OpenLLM both serve as frameworks for deploying and managing large language models (LLMs), with SGLang providing a high-performance serving environment particularly for multimodal models, while OpenLLM focuses on enabling the self-hosting of LLMs through an OpenAI-compatible API interface. This alternative relationship arises from their differing approaches to deployment and optimization","successor_context":null,"tool":{"slug":"bentoml-openllm","name":"OpenLLM","tagline":"Run any open-source LLMs as OpenAI compatible API endpoint in the cloud.","github_url":"https://github.com/bentoml/OpenLLM","owner":"bentoml","repo":"OpenLLM","owner_avatar_url":"https://avatars.githubusercontent.com/u/49176046?v=4","primary_language":"Python","stars":12454,"forks":828,"topics":["bentoml","fine-tuning","llama","llama2","llama3-1","llama3-2","llama3-2-vision","llm","llm-inference","llm-ops","llm-serving","llmops","mistral","mlops","model-inference","open-source-llm","openllm","vicuna"],"archived":false,"github_pushed_at":"2026-08-03T16:59:03+00:00","maintenance_label":"Very active","stars_delta_30d":66,"url":"https://www.graphcanon.com/tools/bentoml-openllm","markdown_url":"https://www.graphcanon.com/tools/bentoml-openllm.md","api_url":"https://www.graphcanon.com/api/graphcanon/tools/bentoml-openllm","graph_url":"https://www.graphcanon.com/api/graphcanon/graph?tool=bentoml-openllm"}},{"type":"integrates_with","direction":"out","explanation":"SGLang is a high-performance serving framework for large language models and multimodal models, which can integrate with the huggingface transformers library to support various state-of-the-art machine learning models.","successor_context":null,"tool":{"slug":"huggingface-transformers","name":"transformers","tagline":"Transformers: the model-definition framework for state-of-the-art machine learning models in text, vision, audio, and multimodal models","github_url":"https://github.com/huggingface/transformers","owner":"huggingface","repo":"transformers","owner_avatar_url":"https://avatars.githubusercontent.com/u/25720743?v=4","primary_language":"Python","stars":164121,"forks":34249,"topics":["audio","deep-learning","deepseek","gemma","glm","hacktoberfest","llm","machine-learning","model-hub","natural-language-processing","nlp","pretrained-models","python","pytorch","pytorch-transformers","qwen","speech-recognition","transformer","vlm"],"archived":false,"github_pushed_at":"2026-08-15T22:28:12+00:00","maintenance_label":"Very active","stars_delta_30d":1457,"url":"https://www.graphcanon.com/tools/huggingface-transformers","markdown_url":"https://www.graphcanon.com/tools/huggingface-transformers.md","api_url":"https://www.graphcanon.com/api/graphcanon/tools/huggingface-transformers","graph_url":"https://www.graphcanon.com/api/graphcanon/graph?tool=huggingface-transformers"}},{"type":"depends_on","direction":"out","explanation":"The repository content indicates that SGLang supports the latest open models from NVIDIA such as Nemotron 3 Ultra and Super; thus it depends on tools like NemoClaw for reference stacking and sandboxing AI agents.","successor_context":null,"tool":{"slug":"nvidia-nemoclaw","name":"NemoClaw","tagline":"Run agents like Hermes, LangChain Deep Agents, and OpenClaw securely inside NVIDIA OpenShell with managed inference","github_url":"https://github.com/NVIDIA/NemoClaw","owner":"NVIDIA","repo":"NemoClaw","owner_avatar_url":"https://avatars.githubusercontent.com/u/1728152?v=4","primary_language":"TypeScript","stars":22208,"forks":3031,"topics":["ai-agents","hermes","nvidia","openclaw","openshell","sandboxing","typescript"],"archived":false,"github_pushed_at":"2026-08-20T00:00:52+00:00","maintenance_label":"Very active","stars_delta_30d":357,"url":"https://www.graphcanon.com/tools/nvidia-nemoclaw","markdown_url":"https://www.graphcanon.com/tools/nvidia-nemoclaw.md","api_url":"https://www.graphcanon.com/api/graphcanon/tools/nvidia-nemoclaw","graph_url":"https://www.graphcanon.com/api/graphcanon/graph?tool=nvidia-nemoclaw"}},{"type":"alternative","direction":"out","explanation":"Both SGLang and RTK are designed to enhance the performance of language models during inference, with RTK specifically targeting token reduction for CLI applications.","successor_context":null,"tool":{"slug":"rtk-ai-rtk","name":"rtk","tagline":"CLI proxy reducing LLM token consumption by 60-90% on common dev commands","github_url":"https://github.com/rtk-ai/rtk","owner":"rtk-ai","repo":"rtk","owner_avatar_url":"https://avatars.githubusercontent.com/u/258253854?v=4","primary_language":"Rust","stars":76247,"forks":4795,"topics":["agentic-coding","ai-coding","anthropic","claude-code","cli","command-line-tool","cost-reduction","developer-tools","llm","open-source","productivity","rust","token-optimization"],"archived":false,"github_pushed_at":"2026-08-15T02:01:37+00:00","maintenance_label":"Very active","stars_delta_30d":4851,"url":"https://www.graphcanon.com/tools/rtk-ai-rtk","markdown_url":"https://www.graphcanon.com/tools/rtk-ai-rtk.md","api_url":"https://www.graphcanon.com/api/graphcanon/tools/rtk-ai-rtk","graph_url":"https://www.graphcanon.com/api/graphcanon/graph?tool=rtk-ai-rtk"}},{"type":"alternative","direction":"out","explanation":"SGLang and mlc-LLM both aim at deploying large language models efficiently across different hardware setups. They differ in their underlying technologies and deployment strategies, making them alternatives for model serving.","successor_context":null,"tool":{"slug":"mlc-ai-mlc-llm","name":"mlc-llm","tagline":"Universal LLM Deployment Engine with ML Compilation","github_url":"https://github.com/mlc-ai/mlc-llm","owner":"mlc-ai","repo":"mlc-llm","owner_avatar_url":"https://avatars.githubusercontent.com/u/106173866?v=4","primary_language":"Python","stars":23063,"forks":2111,"topics":["language-model","llm","machine-learning-compilation","tvm"],"archived":false,"github_pushed_at":"2026-07-31T03:03:18+00:00","maintenance_label":"Active","stars_delta_30d":103,"url":"https://www.graphcanon.com/tools/mlc-ai-mlc-llm","markdown_url":"https://www.graphcanon.com/tools/mlc-ai-mlc-llm.md","api_url":"https://www.graphcanon.com/api/graphcanon/tools/mlc-ai-mlc-llm","graph_url":"https://www.graphcanon.com/api/graphcanon/graph?tool=mlc-ai-mlc-llm"}},{"type":"integrates_with","direction":"out","explanation":"SGLang could integrate with ML Compilation to facilitate deployment and optimization of various LLMs across different hardware platforms.","successor_context":null,"tool":{"slug":"mlc-ai-mlc-llm","name":"mlc-llm","tagline":"Universal LLM Deployment Engine with ML Compilation","github_url":"https://github.com/mlc-ai/mlc-llm","owner":"mlc-ai","repo":"mlc-llm","owner_avatar_url":"https://avatars.githubusercontent.com/u/106173866?v=4","primary_language":"Python","stars":23063,"forks":2111,"topics":["language-model","llm","machine-learning-compilation","tvm"],"archived":false,"github_pushed_at":"2026-07-31T03:03:18+00:00","maintenance_label":"Active","stars_delta_30d":103,"url":"https://www.graphcanon.com/tools/mlc-ai-mlc-llm","markdown_url":"https://www.graphcanon.com/tools/mlc-ai-mlc-llm.md","api_url":"https://www.graphcanon.com/api/graphcanon/tools/mlc-ai-mlc-llm","graph_url":"https://www.graphcanon.com/api/graphcanon/graph?tool=mlc-ai-mlc-llm"}},{"type":"alternative","direction":"out","explanation":"SGLang and vllm both aim at providing easy and fast LLM serving solutions but use different approaches to achieve high performance in inference.","successor_context":null,"tool":{"slug":"vllm-project-vllm","name":"vllm","tagline":"A high-throughput and memory-efficient inference and serving engine for LLMs","github_url":"https://github.com/vllm-project/vllm","owner":"vllm-project","repo":"vllm","owner_avatar_url":"https://avatars.githubusercontent.com/u/136984999?v=4","primary_language":"Python","stars":87847,"forks":20135,"topics":["amd","blackwell","cuda","deepseek","deepseek-v3","gpt","gpt-oss","inference","kimi","llama","llm","llm-serving","model-serving","moe","openai","pytorch","qwen","qwen3","tpu","transformer"],"archived":false,"github_pushed_at":"2026-08-01T11:55:36+00:00","maintenance_label":"Very active","url":"https://www.graphcanon.com/tools/vllm-project-vllm","markdown_url":"https://www.graphcanon.com/tools/vllm-project-vllm.md","api_url":"https://www.graphcanon.com/api/graphcanon/tools/vllm-project-vllm","graph_url":"https://www.graphcanon.com/api/graphcanon/graph?tool=vllm-project-vllm"}},{"type":"integrates_with","direction":"out","explanation":"Given that AstrBot is a development framework for multiple platforms including LLM support, it can integrate with SGLang to leverage its high-performance serving capabilities for large language models and multimodal inference.","successor_context":null,"tool":{"slug":"astrbotdevs-astrbot","name":"AstrBot","tagline":"AI Agent Assistant & development framework that integrates lots of IM platforms, LLMs, plugins and AI feature","github_url":"https://github.com/AstrBotDevs/AstrBot","owner":"AstrBotDevs","repo":"AstrBot","owner_avatar_url":"https://avatars.githubusercontent.com/u/197911947?v=4","primary_language":"Python","stars":39240,"forks":2799,"topics":["agent","ai","astrbot","chatbot","chatgpt","discord","docker","gemini","gpt","llama","llm","mcp","openai","python","qq","qqbot","telegram"],"archived":false,"github_pushed_at":"2026-08-16T09:02:32+00:00","maintenance_label":"Very active","stars_delta_30d":2785,"url":"https://www.graphcanon.com/tools/astrbotdevs-astrbot","markdown_url":"https://www.graphcanon.com/tools/astrbotdevs-astrbot.md","api_url":"https://www.graphcanon.com/api/graphcanon/tools/astrbotdevs-astrbot","graph_url":"https://www.graphcanon.com/api/graphcanon/graph?tool=astrbotdevs-astrbot"}},{"type":"integrates_with","direction":"in","explanation":"sglang is a serving framework for large models which complements litellm's gateway by providing the backend infrastructure to serve large language models efficiently.","successor_context":null,"tool":{"slug":"berriai-litellm","name":"litellm","tagline":"Python SDK and Proxy Server for calling multiple LLM APIs","github_url":"https://github.com/BerriAI/litellm","owner":"BerriAI","repo":"litellm","owner_avatar_url":"https://avatars.githubusercontent.com/u/121462774?v=4","primary_language":"Python","stars":55221,"forks":10231,"topics":["ai-gateway","anthropic","azure-openai","bedrock","gateway","langchain","litellm","llm","llm-gateway","llmops","mcp-gateway","openai","openai-proxy","rust","rust-ai","vertex-ai"],"archived":false,"github_pushed_at":"2026-08-01T05:53:28+00:00","maintenance_label":"Very active","url":"https://www.graphcanon.com/tools/berriai-litellm","markdown_url":"https://www.graphcanon.com/tools/berriai-litellm.md","api_url":"https://www.graphcanon.com/api/graphcanon/tools/berriai-litellm","graph_url":"https://www.graphcanon.com/api/graphcanon/graph?tool=berriai-litellm"}},{"type":"related","direction":"in","explanation":"This entry seems duplicate and repeats the already mentioned relation with sglang, highlighting its connection in efficient model serving.","successor_context":null,"tool":{"slug":"berriai-litellm","name":"litellm","tagline":"Python SDK and Proxy Server for calling multiple LLM APIs","github_url":"https://github.com/BerriAI/litellm","owner":"BerriAI","repo":"litellm","owner_avatar_url":"https://avatars.githubusercontent.com/u/121462774?v=4","primary_language":"Python","stars":55221,"forks":10231,"topics":["ai-gateway","anthropic","azure-openai","bedrock","gateway","langchain","litellm","llm","llm-gateway","llmops","mcp-gateway","openai","openai-proxy","rust","rust-ai","vertex-ai"],"archived":false,"github_pushed_at":"2026-08-01T05:53:28+00:00","maintenance_label":"Very active","url":"https://www.graphcanon.com/tools/berriai-litellm","markdown_url":"https://www.graphcanon.com/tools/berriai-litellm.md","api_url":"https://www.graphcanon.com/api/graphcanon/tools/berriai-litellm","graph_url":"https://www.graphcanon.com/api/graphcanon/graph?tool=berriai-litellm"}},{"type":"integrates_with","direction":"in","explanation":"SGLang is a serving framework that can integrate with Ray for efficient scaling and deployment of large language models and multimodal applications.","successor_context":null,"tool":{"slug":"ray-project-ray","name":"ray","tagline":"Ray is an AI compute engine with a core distributed runtime and AI Libraries for accelerating ML workloads.","github_url":"https://github.com/ray-project/ray","owner":"ray-project","repo":"ray","owner_avatar_url":"https://avatars.githubusercontent.com/u/22125274?v=4","primary_language":"Python","stars":43526,"forks":7929,"topics":["data-science","deep-learning","deployment","distributed","hyperparameter-optimization","hyperparameter-search","large-language-models","llm","llm-inference","llm-serving","machine-learning","optimization","parallel","python","pytorch","ray","reinforcement-learning","rllib","serving","tensorflow"],"archived":false,"github_pushed_at":"2026-08-16T00:26:16+00:00","maintenance_label":"Very active","stars_delta_30d":270,"url":"https://www.graphcanon.com/tools/ray-project-ray","markdown_url":"https://www.graphcanon.com/tools/ray-project-ray.md","api_url":"https://www.graphcanon.com/api/graphcanon/tools/ray-project-ray","graph_url":"https://www.graphcanon.com/api/graphcanon/graph?tool=ray-project-ray"}},{"type":"integrates_with","direction":"in","explanation":"SG-lang is a serving framework for large language models, making it compatible with MLflow as an infrastructure to support the deployment and management of LLMs.","successor_context":null,"tool":{"slug":"mlflow-mlflow","name":"mlflow","tagline":"AI engineering platform for debugging, evaluating, monitoring, and optimizing AI applications","github_url":"https://github.com/mlflow/mlflow","owner":"mlflow","repo":"mlflow","owner_avatar_url":"https://avatars.githubusercontent.com/u/39938107?v=4","primary_language":"Python","stars":27591,"forks":6189,"topics":["agentops","agents","ai","ai-governance","apache-spark","evaluation","langchain","llm-evaluation","llmops","machine-learning","ml","mlflow","mlops","model-management","observability","open-source","openai","prompt-engineering"],"archived":false,"github_pushed_at":"2026-08-20T00:54:28+00:00","maintenance_label":"Very active","stars_delta_30d":476,"url":"https://www.graphcanon.com/tools/mlflow-mlflow","markdown_url":"https://www.graphcanon.com/tools/mlflow-mlflow.md","api_url":"https://www.graphcanon.com/api/graphcanon/tools/mlflow-mlflow","graph_url":"https://www.graphcanon.com/api/graphcanon/graph?tool=mlflow-mlflow"}},{"type":"integrates_with","direction":"in","explanation":"'SGLang' provides a serving framework for large language models which could be used alongside llm-action’s focus on LLM inference practices and optimizations.","successor_context":null,"tool":{"slug":"liguodongiot-llm-action","name":"llm-action","tagline":"Aims to share large model technology principles and practical experience (large model engineering, application implementation)","github_url":"https://github.com/liguodongiot/llm-action","owner":"liguodongiot","repo":"llm-action","owner_avatar_url":"https://avatars.githubusercontent.com/u/13220186?v=4","primary_language":"HTML","stars":24898,"forks":2842,"topics":["llm","llm-inference","llm-serving","llm-training","llmops"],"archived":false,"github_pushed_at":"2026-07-19T13:13:31+00:00","maintenance_label":"Active","stars_delta_30d":162,"url":"https://www.graphcanon.com/tools/liguodongiot-llm-action","markdown_url":"https://www.graphcanon.com/tools/liguodongiot-llm-action.md","api_url":"https://www.graphcanon.com/api/graphcanon/tools/liguodongiot-llm-action","graph_url":"https://www.graphcanon.com/api/graphcanon/graph?tool=liguodongiot-llm-action"}},{"type":"alternative","direction":"in","explanation":"Both Jina-Serve and sglang are serving frameworks for large language models and multimodal models, each offering their own approach to deployment and scalability.","successor_context":null,"tool":{"slug":"jina-ai-serve","name":"serve","tagline":"Build multimodal AI applications with cloud-native stack","github_url":"https://github.com/jina-ai/serve","owner":"jina-ai","repo":"serve","owner_avatar_url":"https://avatars.githubusercontent.com/u/60539444?v=4","primary_language":"Python","stars":21863,"forks":2243,"topics":["cloud-native","cncf","deep-learning","docker","fastapi","framework","generative-ai","grpc","jaeger","kubernetes","llmops","machine-learning","microservice","mlops","multimodal","neural-search","opentelemetry","orchestration","pipeline","prometheus"],"archived":false,"github_pushed_at":"2025-03-24T13:59:54+00:00","maintenance_label":"Dormant","url":"https://www.graphcanon.com/tools/jina-ai-serve","markdown_url":"https://www.graphcanon.com/tools/jina-ai-serve.md","api_url":"https://www.graphcanon.com/api/graphcanon/tools/jina-ai-serve","graph_url":"https://www.graphcanon.com/api/graphcanon/graph?tool=jina-ai-serve"}},{"type":"integrates_with","direction":"in","explanation":"`ml-engineering` focuses on scalable machine learning solutions including serving large language models, for which `sglang` provides a framework.","successor_context":null,"tool":{"slug":"stas00-ml-engineering","name":"ml-engineering","tagline":"Machine Learning Engineering Open Book","github_url":"https://github.com/stas00/ml-engineering","owner":"stas00","repo":"ml-engineering","owner_avatar_url":"https://avatars.githubusercontent.com/u/10676103?v=4","primary_language":"Python","stars":18632,"forks":1200,"topics":["ai","debugging","gpus","inference","large-language-models","llm","machine-learning","machine-learning-engineering","mlops","network","pytorch","scalability","slurm","storage","training","transformers"],"archived":false,"github_pushed_at":"2026-08-14T19:59:44+00:00","maintenance_label":"Very active","stars_delta_30d":216,"url":"https://www.graphcanon.com/tools/stas00-ml-engineering","markdown_url":"https://www.graphcanon.com/tools/stas00-ml-engineering.md","api_url":"https://www.graphcanon.com/api/graphcanon/tools/stas00-ml-engineering","graph_url":"https://www.graphcanon.com/api/graphcanon/graph?tool=stas00-ml-engineering"}},{"type":"alternative","direction":"in","explanation":"SGLang and Ray both target providing a serving framework for large language models and offer tools for scaling ML workloads.","successor_context":null,"tool":{"slug":"ray-project-ray","name":"ray","tagline":"Ray is an AI compute engine with a core distributed runtime and AI Libraries for accelerating ML workloads.","github_url":"https://github.com/ray-project/ray","owner":"ray-project","repo":"ray","owner_avatar_url":"https://avatars.githubusercontent.com/u/22125274?v=4","primary_language":"Python","stars":43526,"forks":7929,"topics":["data-science","deep-learning","deployment","distributed","hyperparameter-optimization","hyperparameter-search","large-language-models","llm","llm-inference","llm-serving","machine-learning","optimization","parallel","python","pytorch","ray","reinforcement-learning","rllib","serving","tensorflow"],"archived":false,"github_pushed_at":"2026-08-16T00:26:16+00:00","maintenance_label":"Very active","stars_delta_30d":270,"url":"https://www.graphcanon.com/tools/ray-project-ray","markdown_url":"https://www.graphcanon.com/tools/ray-project-ray.md","api_url":"https://www.graphcanon.com/api/graphcanon/tools/ray-project-ray","graph_url":"https://www.graphcanon.com/api/graphcanon/graph?tool=ray-project-ray"}},{"type":"integrates_with","direction":"in","explanation":"`SGLang` serves as a language model serving framework, which can potentially integrate with Petals for serving large models using distributed methods.","successor_context":null,"tool":{"slug":"bigscience-workshop-petals","name":"petals","tagline":"Run LLMs at home, BitTorrent-style. Fine-tuning and inference up to 10x faster than offloading","github_url":"https://github.com/bigscience-workshop/petals","owner":"bigscience-workshop","repo":"petals","owner_avatar_url":"https://avatars.githubusercontent.com/u/82455566?v=4","primary_language":"Python","stars":10496,"forks":642,"topics":["bloom","chatbot","deep-learning","distributed-systems","falcon","gpt","guanaco","language-models","large-language-models","llama","machine-learning","mixtral","neural-networks","nlp","pipeline-parallelism","pretrained-models","pytorch","tensor-parallelism","transformer","volunteer-computing"],"archived":false,"github_pushed_at":"2024-09-07T11:54:28+00:00","maintenance_label":"Dormant","stars_delta_30d":212,"url":"https://www.graphcanon.com/tools/bigscience-workshop-petals","markdown_url":"https://www.graphcanon.com/tools/bigscience-workshop-petals.md","api_url":"https://www.graphcanon.com/api/graphcanon/tools/bigscience-workshop-petals","graph_url":"https://www.graphcanon.com/api/graphcanon/graph?tool=bigscience-workshop-petals"}},{"type":"integrates_with","direction":"in","explanation":"PowerInfer, as a high-speed LLM serving solution, can potentially integrate with sglang's broader framework for both large language and multimodal model serving.","successor_context":null,"tool":{"slug":"tiiny-ai-powerinfer","name":"PowerInfer","tagline":"High-speed Large Language Model Serving for Local Deployment","github_url":"https://github.com/Tiiny-AI/PowerInfer","owner":"Tiiny-AI","repo":"PowerInfer","owner_avatar_url":"https://avatars.githubusercontent.com/u/256922953?v=4","primary_language":"C++","stars":9718,"forks":591,"topics":["large-language-models","llama","llm","llm-inference","local-inference"],"archived":false,"github_pushed_at":"2026-05-11T06:48:06+00:00","maintenance_label":"Slowing","stars_delta_30d":76,"url":"https://www.graphcanon.com/tools/tiiny-ai-powerinfer","markdown_url":"https://www.graphcanon.com/tools/tiiny-ai-powerinfer.md","api_url":"https://www.graphcanon.com/api/graphcanon/tools/tiiny-ai-powerinfer","graph_url":"https://www.graphcanon.com/api/graphcanon/graph?tool=tiiny-ai-powerinfer"}},{"type":"integrates_with","direction":"in","explanation":"Both projects focus on large language and multimodal models, SGLang as a serving framework could potentially integrate with VAR to handle model deployment for image generation.","successor_context":null,"tool":{"slug":"foundationvision-var","name":"VAR","tagline":"Official implementation of Visual Autoregressive Modeling for scalable image generation","github_url":"https://github.com/FoundationVision/VAR","owner":"FoundationVision","repo":"VAR","owner_avatar_url":"https://avatars.githubusercontent.com/u/151817217?v=4","primary_language":"Jupyter Notebook","stars":8727,"forks":571,"topics":["auto-regressive-model","autoregressive-models","diffusion-models","generative-ai","generative-model","gpt","gpt-2","image-generation","large-language-models","neurips","transformers","vision-transformer"],"archived":false,"github_pushed_at":"2025-11-10T21:42:29+00:00","maintenance_label":"Slowing","stars_delta_30d":19,"url":"https://www.graphcanon.com/tools/foundationvision-var","markdown_url":"https://www.graphcanon.com/tools/foundationvision-var.md","api_url":"https://www.graphcanon.com/api/graphcanon/tools/foundationvision-var","graph_url":"https://www.graphcanon.com/api/graphcanon/graph?tool=foundationvision-var"}},{"type":"alternative","direction":"in","explanation":"Bifrost and SGLang are both frameworks for serving large language models, though Bifrost focuses on performance and enterprise features.","successor_context":null,"tool":{"slug":"maximhq-bifrost","name":"bifrost","tagline":"Fast Enterprise AI Gateway with Adaptive Load Balancer and Guardrails","github_url":"https://github.com/maximhq/bifrost","owner":"maximhq","repo":"bifrost","owner_avatar_url":"https://avatars.githubusercontent.com/u/139708451?v=4","primary_language":"Go","stars":7449,"forks":1073,"topics":["ai-gateway","gateway","gateway-services","generative-ai","guardrails","llm","llm-cost","llm-gateway","llm-observability","llmops","load-balancing","mcp-client","mcp-gateway","mcp-server","model-router","token-management"],"archived":false,"github_pushed_at":"2026-08-20T11:57:33+00:00","maintenance_label":"Very active","stars_delta_30d":812,"url":"https://www.graphcanon.com/tools/maximhq-bifrost","markdown_url":"https://www.graphcanon.com/tools/maximhq-bifrost.md","api_url":"https://www.graphcanon.com/api/graphcanon/tools/maximhq-bifrost","graph_url":"https://www.graphcanon.com/api/graphcanon/graph?tool=maximhq-bifrost"}},{"type":"integrates_with","direction":"in","explanation":"OptiLLM could integrate with the sglang framework for serving large language models, providing both optimized performance and a robust serving infrastructure.","successor_context":null,"tool":{"slug":"algorithmicsuperintelligence-optillm","name":"optillm","tagline":"Optimizing inference proxy for LLMs","github_url":"https://github.com/algorithmicsuperintelligence/optillm","owner":"algorithmicsuperintelligence","repo":"optillm","owner_avatar_url":"https://avatars.githubusercontent.com/u/238764598?v=4","primary_language":"Python","stars":4244,"forks":385,"topics":["agent","agentic-ai","agentic-framework","agentic-workflow","agents","api-gateway","chain-of-thought","genai","large-language-models","llm","llm-inference","llmapi","mixture-of-experts","moa","monte-carlo-tree-search","openai","openai-api","optimization","prompt-engineering","proxy-server"],"archived":false,"github_pushed_at":"2026-07-18T12:56:27+00:00","maintenance_label":"Steady","stars_delta_30d":67,"url":"https://www.graphcanon.com/tools/algorithmicsuperintelligence-optillm","markdown_url":"https://www.graphcanon.com/tools/algorithmicsuperintelligence-optillm.md","api_url":"https://www.graphcanon.com/api/graphcanon/tools/algorithmicsuperintelligence-optillm","graph_url":"https://www.graphcanon.com/api/graphcanon/graph?tool=algorithmicsuperintelligence-optillm"}},{"type":"integrates_with","direction":"in","explanation":"\"free-llm-api-keys\" can supply API keys for LLMs that the \"sglang\" serving framework supports, making it easier to serve and deploy large language models.","successor_context":null,"tool":{"slug":"alistaitsacle-free-llm-api-keys","name":"free-llm-api-keys","tagline":"Free LLM API keys for GPT-5.5, Claude, DeepSeek, Gemini, Grok","github_url":"https://github.com/alistaitsacle/free-llm-api-keys","owner":"alistaitsacle","repo":"free-llm-api-keys","owner_avatar_url":"https://avatars.githubusercontent.com/u/94039400?v=4","primary_language":null,"stars":3176,"forks":362,"topics":["ai","api","api-key","api-keys","chatgpt","claude","deepseek","free","free-api","free-api-key","free-gpt","free-llm","gemini","gpt","gpt4","grok","large-language-models","llm","llm-api","openai"],"archived":false,"github_pushed_at":"2026-07-08T08:07:07+00:00","maintenance_label":"Very active","url":"https://www.graphcanon.com/tools/alistaitsacle-free-llm-api-keys","markdown_url":"https://www.graphcanon.com/tools/alistaitsacle-free-llm-api-keys.md","api_url":"https://www.graphcanon.com/api/graphcanon/tools/alistaitsacle-free-llm-api-keys","graph_url":"https://www.graphcanon.com/api/graphcanon/graph?tool=alistaitsacle-free-llm-api-keys"}},{"type":"depends_on","direction":"in","explanation":"If AICI needs a framework or method to deploy its controllers with various LLMs, sglang could serve as the underlying serving framework upon which AICI builds and experiments.","successor_context":null,"tool":{"slug":"microsoft-aici","name":"aici","tagline":"Builds Controllers for Constrained LLM Output in Real-time Using Wasm","github_url":"https://github.com/microsoft/aici","owner":"microsoft","repo":"aici","owner_avatar_url":"https://avatars.githubusercontent.com/u/6154722?v=4","primary_language":"Rust","stars":2075,"forks":85,"topics":["ai","inference","language-model","llm","llm-framework","llm-inference","llm-serving","llmops","model-serving","rust","transformer","wasm","wasmtime"],"archived":false,"github_pushed_at":"2025-01-22T21:14:57+00:00","maintenance_label":"Dormant","stars_delta_30d":-2,"url":"https://www.graphcanon.com/tools/microsoft-aici","markdown_url":"https://www.graphcanon.com/tools/microsoft-aici.md","api_url":"https://www.graphcanon.com/api/graphcanon/tools/microsoft-aici","graph_url":"https://www.graphcanon.com/api/graphcanon/graph?tool=microsoft-aici"}},{"type":"alternative","direction":"in","explanation":"Both sglang and Paddler offer serving frameworks for large language models with considerations for multimodal models; however, their approaches to deployment and management may differ.","successor_context":null,"tool":{"slug":"intentee-paddler","name":"paddler","tagline":"Open-source LLM/VLM load balancer and serving platform for self-hosting at scale","github_url":"https://github.com/intentee/paddler","owner":"intentee","repo":"paddler","owner_avatar_url":"https://avatars.githubusercontent.com/u/215040511?v=4","primary_language":"Rust","stars":1663,"forks":97,"topics":["ai","llamacpp","llm","llmops","load-balancer"],"archived":false,"github_pushed_at":"2026-07-19T19:36:21+00:00","maintenance_label":"Steady","stars_delta_30d":21,"url":"https://www.graphcanon.com/tools/intentee-paddler","markdown_url":"https://www.graphcanon.com/tools/intentee-paddler.md","api_url":"https://www.graphcanon.com/api/graphcanon/tools/intentee-paddler","graph_url":"https://www.graphcanon.com/api/graphcanon/graph?tool=intentee-paddler"}},{"type":"alternative","direction":"in","explanation":"Both Langcorn and sgLang are focused on the deployment of large language models, but they approach it differently. While Langcorn is tailored specifically for LangChain models with FastAPI as a backend, sgLang is more generalized in its approach.","successor_context":null,"tool":{"slug":"msoedov-langcorn","name":"langcorn","tagline":"Serving LangChain LLM apps and agents automagically with FastApi","github_url":"https://github.com/msoedov/langcorn","owner":"msoedov","repo":"langcorn","owner_avatar_url":"https://avatars.githubusercontent.com/u/1958116?v=4","primary_language":"Python","stars":938,"forks":69,"topics":["api","fastapi","langchain","langchain-python","large-language-models","llm","llmops","openai-api","rest-api","vercel","vercel-serverless-functions"],"archived":false,"github_pushed_at":"2024-07-15T18:05:52+00:00","maintenance_label":"Dormant","stars_delta_30d":0,"url":"https://www.graphcanon.com/tools/msoedov-langcorn","markdown_url":"https://www.graphcanon.com/tools/msoedov-langcorn.md","api_url":"https://www.graphcanon.com/api/graphcanon/tools/msoedov-langcorn","graph_url":"https://www.graphcanon.com/api/graphcanon/graph?tool=msoedov-langcorn"}},{"type":"alternative","direction":"in","explanation":"vllm and sglang both serve as high-performance frameworks for the efficient deployment and inference of large language models, with vllm specifically emphasizing memory efficiency and support for quantization techniques, while sglang offers broader support including multimodal models. Their alternative relationship stems from providing similar functionalities tailored to different optimization and","successor_context":null,"tool":{"slug":"vllm-project-vllm","name":"vllm","tagline":"A high-throughput and memory-efficient inference and serving engine for LLMs","github_url":"https://github.com/vllm-project/vllm","owner":"vllm-project","repo":"vllm","owner_avatar_url":"https://avatars.githubusercontent.com/u/136984999?v=4","primary_language":"Python","stars":87847,"forks":20135,"topics":["amd","blackwell","cuda","deepseek","deepseek-v3","gpt","gpt-oss","inference","kimi","llama","llm","llm-serving","model-serving","moe","openai","pytorch","qwen","qwen3","tpu","transformer"],"archived":false,"github_pushed_at":"2026-08-01T11:55:36+00:00","maintenance_label":"Very active","url":"https://www.graphcanon.com/tools/vllm-project-vllm","markdown_url":"https://www.graphcanon.com/tools/vllm-project-vllm.md","api_url":"https://www.graphcanon.com/api/graphcanon/tools/vllm-project-vllm","graph_url":"https://www.graphcanon.com/api/graphcanon/graph?tool=vllm-project-vllm"}},{"type":"alternative","direction":"in","explanation":"Both SGLang and LoRAX are serving frameworks designed for large language models, differing in their approach to handling dynamic model loadings and integration with various LLM adapters.","successor_context":null,"tool":{"slug":"predibase-lorax","name":"lorax","tagline":"Multi-LoRA inference server for scalable fine-tuned LLMs","github_url":"https://github.com/predibase/lorax","owner":"predibase","repo":"lorax","owner_avatar_url":"https://avatars.githubusercontent.com/u/75280641?v=4","primary_language":"Python","stars":3826,"forks":326,"topics":["fine-tuning","gpt","llama","llm","llm-inference","llm-serving","llmops","lora","model-serving","pytorch","transformers"],"archived":false,"github_pushed_at":"2026-05-28T18:12:20+00:00","maintenance_label":"Steady","stars_delta_30d":10,"url":"https://www.graphcanon.com/tools/predibase-lorax","markdown_url":"https://www.graphcanon.com/tools/predibase-lorax.md","api_url":"https://www.graphcanon.com/api/graphcanon/tools/predibase-lorax","graph_url":"https://www.graphcanon.com/api/graphcanon/graph?tool=predibase-lorax"}},{"type":"related","direction":"in","explanation":"Both serve large language and multimodal models, but sglang is a serving framework while OpenRLHF focuses on RL optimization.","successor_context":null,"tool":{"slug":"openrlhf-openrlhf","name":"OpenRLHF","tagline":"Easy-to-use, Scalable and High-performance Agentic RL Framework based on Ray","github_url":"https://github.com/OpenRLHF/OpenRLHF","owner":"OpenRLHF","repo":"OpenRLHF","owner_avatar_url":"https://avatars.githubusercontent.com/u/175771028?v=4","primary_language":"Python","stars":9891,"forks":996,"topics":["large-language-models","proximal-policy-optimization","raylib","reinforcement-learning","reinforcement-learning-from-human-feedback","transformers","visual-language-models","vllm"],"archived":false,"github_pushed_at":"2026-07-14T01:57:21+00:00","maintenance_label":"Active","stars_delta_30d":132,"url":"https://www.graphcanon.com/tools/openrlhf-openrlhf","markdown_url":"https://www.graphcanon.com/tools/openrlhf-openrlhf.md","api_url":"https://www.graphcanon.com/api/graphcanon/tools/openrlhf-openrlhf","graph_url":"https://www.graphcanon.com/api/graphcanon/graph?tool=openrlhf-openrlhf"}}],"neighbours":[{"slug":"langchain-ai-langchain","name":"langchain","tagline":"The agent engineering platform.","github_url":"https://github.com/langchain-ai/langchain","owner":"langchain-ai","repo":"langchain","owner_avatar_url":"https://avatars.githubusercontent.com/u/126733545?v=4","primary_language":"Python","stars":143615,"forks":23930,"topics":["agents","ai","ai-agents","anthropic","chatgpt","deepagents","enterprise","framework","gemini","generative-ai","langchain","langgraph","llm","multiagent","open-source","openai","pydantic","python","rag","typescript"],"archived":false,"github_pushed_at":"2026-08-07T08:27:07+00:00","maintenance_label":"Active","url":"https://www.graphcanon.com/tools/langchain-ai-langchain","markdown_url":"https://www.graphcanon.com/tools/langchain-ai-langchain.md","api_url":"https://www.graphcanon.com/api/graphcanon/tools/langchain-ai-langchain","graph_url":"https://www.graphcanon.com/api/graphcanon/graph?tool=langchain-ai-langchain","shared_categories":[]},{"slug":"vllm-project-vllm","name":"vllm","tagline":"A high-throughput and memory-efficient inference and serving engine for LLMs","github_url":"https://github.com/vllm-project/vllm","owner":"vllm-project","repo":"vllm","owner_avatar_url":"https://avatars.githubusercontent.com/u/136984999?v=4","primary_language":"Python","stars":87847,"forks":20135,"topics":["amd","blackwell","cuda","deepseek","deepseek-v3","gpt","gpt-oss","inference","kimi","llama","llm","llm-serving","model-serving","moe","openai","pytorch","qwen","qwen3","tpu","transformer"],"archived":false,"github_pushed_at":"2026-08-01T11:55:36+00:00","maintenance_label":"Active","url":"https://www.graphcanon.com/tools/vllm-project-vllm","markdown_url":"https://www.graphcanon.com/tools/vllm-project-vllm.md","api_url":"https://www.graphcanon.com/api/graphcanon/tools/vllm-project-vllm","graph_url":"https://www.graphcanon.com/api/graphcanon/graph?tool=vllm-project-vllm","shared_categories":["inference-serving"]},{"slug":"lm-sys-fastchat","name":"FastChat","tagline":"An open platform for training, serving, and evaluating large language models","github_url":"https://github.com/lm-sys/FastChat","owner":"lm-sys","repo":"FastChat","owner_avatar_url":"https://avatars.githubusercontent.com/u/126381704?v=4","primary_language":"Python","stars":39517,"forks":4788,"topics":[],"archived":false,"github_pushed_at":"2026-05-01T00:25:53+00:00","maintenance_label":"Slowing","url":"https://www.graphcanon.com/tools/lm-sys-fastchat","markdown_url":"https://www.graphcanon.com/tools/lm-sys-fastchat.md","api_url":"https://www.graphcanon.com/api/graphcanon/tools/lm-sys-fastchat","graph_url":"https://www.graphcanon.com/api/graphcanon/graph?tool=lm-sys-fastchat","shared_categories":["inference-serving"]},{"slug":"langchain-ai-langgraph","name":"langgraph","tagline":"Low-level orchestration framework for building stateful agents.","github_url":"https://github.com/langchain-ai/langgraph","owner":"langchain-ai","repo":"langgraph","owner_avatar_url":"https://avatars.githubusercontent.com/u/126733545?v=4","primary_language":"Python","stars":38352,"forks":6458,"topics":["agents","ai","ai-agents","chatgpt","deepagents","enterprise","framework","gemini","generative-ai","langchain","langgraph","llm","multiagent","open-source","openai","pydantic","python","rag"],"archived":false,"github_pushed_at":"2026-07-28T14:49:41+00:00","maintenance_label":"Active","url":"https://www.graphcanon.com/tools/langchain-ai-langgraph","markdown_url":"https://www.graphcanon.com/tools/langchain-ai-langgraph.md","api_url":"https://www.graphcanon.com/api/graphcanon/tools/langchain-ai-langgraph","graph_url":"https://www.graphcanon.com/api/graphcanon/graph?tool=langchain-ai-langgraph","shared_categories":[]},{"slug":"langfuse-langfuse","name":"langfuse","tagline":"Open source AI engineering platform: LLM evals, observability, metrics, prompt management, playground, datasets","github_url":"https://github.com/langfuse/langfuse","owner":"langfuse","repo":"langfuse","owner_avatar_url":"https://avatars.githubusercontent.com/u/134601687?v=4","primary_language":"TypeScript","stars":32271,"forks":3466,"topics":["analytics","autogen","evaluation","langchain","large-language-models","llama-index","llm","llm-evaluation","llm-observability","llmops","monitoring","observability","open-source","openai","playground","prompt-engineering","prompt-management","self-hosted","ycombinator"],"archived":false,"github_pushed_at":"2026-07-31T22:58:07+00:00","maintenance_label":"Active","url":"https://www.graphcanon.com/tools/langfuse-langfuse","markdown_url":"https://www.graphcanon.com/tools/langfuse-langfuse.md","api_url":"https://www.graphcanon.com/api/graphcanon/tools/langfuse-langfuse","graph_url":"https://www.graphcanon.com/api/graphcanon/graph?tool=langfuse-langfuse","shared_categories":[]},{"slug":"bradyfu-awesome-multimodal-large-language-models","name":"Awesome-Multimodal-Large-Language-Models","tagline":"Latest Advances on Multimodal Large Language Models","github_url":"https://github.com/BradyFU/Awesome-Multimodal-Large-Language-Models","owner":"BradyFU","repo":"Awesome-Multimodal-Large-Language-Models","owner_avatar_url":"https://avatars.githubusercontent.com/u/54254631?v=4","primary_language":null,"stars":17978,"forks":1133,"topics":["chain-of-thought","in-context-learning","instruction-following","instruction-tuning","large-language-models","large-vision-language-model","large-vision-language-models","multi-modality","multimodal-chain-of-thought","multimodal-in-context-learning","multimodal-instruction-tuning","multimodal-large-language-models","visual-instruction-tuning"],"archived":false,"github_pushed_at":"2026-08-14T17:17:50+00:00","maintenance_label":"Active","url":"https://www.graphcanon.com/tools/bradyfu-awesome-multimodal-large-language-models","markdown_url":"https://www.graphcanon.com/tools/bradyfu-awesome-multimodal-large-language-models.md","api_url":"https://www.graphcanon.com/api/graphcanon/tools/bradyfu-awesome-multimodal-large-language-models","graph_url":"https://www.graphcanon.com/api/graphcanon/graph?tool=bradyfu-awesome-multimodal-large-language-models","shared_categories":[]},{"slug":"lightning-ai-litgpt","name":"litgpt","tagline":"High-performance LLMs with recipes for pretraining, finetuning and deployment","github_url":"https://github.com/Lightning-AI/litgpt","owner":"Lightning-AI","repo":"litgpt","owner_avatar_url":"https://avatars.githubusercontent.com/u/58386951?v=4","primary_language":"Python","stars":13605,"forks":1483,"topics":["ai","artificial-intelligence","deep-learning","large-language-models","llm","llm-inference","llms"],"archived":false,"github_pushed_at":"2026-07-20T10:24:12+00:00","maintenance_label":"Steady","url":"https://www.graphcanon.com/tools/lightning-ai-litgpt","markdown_url":"https://www.graphcanon.com/tools/lightning-ai-litgpt.md","api_url":"https://www.graphcanon.com/api/graphcanon/tools/lightning-ai-litgpt","graph_url":"https://www.graphcanon.com/api/graphcanon/graph?tool=lightning-ai-litgpt","shared_categories":["inference-serving"]},{"slug":"fminference-flexllmgen","name":"FlexLLMGen","tagline":"Running large language models on a single GPU for throughput-oriented scenarios.","github_url":"https://github.com/FMInference/FlexLLMGen","owner":"FMInference","repo":"FlexLLMGen","owner_avatar_url":"https://avatars.githubusercontent.com/u/125944572?v=4","primary_language":"Python","stars":9361,"forks":590,"topics":["deep-learning","gpt-3","high-throughput","large-language-models","machine-learning","offloading","opt"],"archived":true,"github_pushed_at":"2024-10-28T03:05:41+00:00","maintenance_label":"Archived","url":"https://www.graphcanon.com/tools/fminference-flexllmgen","markdown_url":"https://www.graphcanon.com/tools/fminference-flexllmgen.md","api_url":"https://www.graphcanon.com/api/graphcanon/tools/fminference-flexllmgen","graph_url":"https://www.graphcanon.com/api/graphcanon/graph?tool=fminference-flexllmgen","shared_categories":["inference-serving"]},{"slug":"langroid-langroid","name":"langroid","tagline":"Harness LLMs with Multi-Agent Programming","github_url":"https://github.com/langroid/langroid","owner":"langroid","repo":"langroid","owner_avatar_url":"https://avatars.githubusercontent.com/u/130325191?v=4","primary_language":"Python","stars":4090,"forks":390,"topics":["agents","ai","chatgpt","function-calling","gpt","gpt-4","gpt4","information-retrieval","language-model","llama","llm","llm-agent","llm-framework","local-llm","multi-agent-systems","openai-api","rag","retrieval-augmented-generation"],"archived":false,"github_pushed_at":"2026-07-29T12:59:53+00:00","maintenance_label":"Active","url":"https://www.graphcanon.com/tools/langroid-langroid","markdown_url":"https://www.graphcanon.com/tools/langroid-langroid.md","api_url":"https://www.graphcanon.com/api/graphcanon/tools/langroid-langroid","graph_url":"https://www.graphcanon.com/api/graphcanon/graph?tool=langroid-langroid","shared_categories":["inference-serving"]},{"slug":"jia-lab-research-mgm","name":"MGM","tagline":"Mini-Gemini: Mining the Potential of Multi-modality Vision Language Models","github_url":"https://github.com/JIA-Lab-research/MGM","owner":"JIA-Lab-research","repo":"MGM","owner_avatar_url":"https://avatars.githubusercontent.com/u/64006090?v=4","primary_language":"Python","stars":3331,"forks":276,"topics":["generation","large-language-models","vision-language-model"],"archived":false,"github_pushed_at":"2024-05-04T14:36:51+00:00","maintenance_label":"Dormant","url":"https://www.graphcanon.com/tools/jia-lab-research-mgm","markdown_url":"https://www.graphcanon.com/tools/jia-lab-research-mgm.md","api_url":"https://www.graphcanon.com/api/graphcanon/tools/jia-lab-research-mgm","graph_url":"https://www.graphcanon.com/api/graphcanon/graph?tool=jia-lab-research-mgm","shared_categories":[]},{"slug":"langchain-ai-langserve","name":"langserve","tagline":"LangServe 🦜️🏓","github_url":"https://github.com/langchain-ai/langserve","owner":"langchain-ai","repo":"langserve","owner_avatar_url":"https://avatars.githubusercontent.com/u/126733545?v=4","primary_language":"JavaScript","stars":2332,"forks":272,"topics":["deployment","fastapi","langchain","langchain-python","llm","llms"],"archived":true,"github_pushed_at":"2026-05-05T19:49:07+00:00","maintenance_label":"Archived","url":"https://www.graphcanon.com/tools/langchain-ai-langserve","markdown_url":"https://www.graphcanon.com/tools/langchain-ai-langserve.md","api_url":"https://www.graphcanon.com/api/graphcanon/tools/langchain-ai-langserve","graph_url":"https://www.graphcanon.com/api/graphcanon/graph?tool=langchain-ai-langserve","shared_categories":["inference-serving"]},{"slug":"msoedov-langcorn","name":"langcorn","tagline":"Serving LangChain LLM apps and agents automagically with FastApi","github_url":"https://github.com/msoedov/langcorn","owner":"msoedov","repo":"langcorn","owner_avatar_url":"https://avatars.githubusercontent.com/u/1958116?v=4","primary_language":"Python","stars":938,"forks":69,"topics":["api","fastapi","langchain","langchain-python","large-language-models","llm","llmops","openai-api","rest-api","vercel","vercel-serverless-functions"],"archived":false,"github_pushed_at":"2024-07-15T18:05:52+00:00","maintenance_label":"Dormant","url":"https://www.graphcanon.com/tools/msoedov-langcorn","markdown_url":"https://www.graphcanon.com/tools/msoedov-langcorn.md","api_url":"https://www.graphcanon.com/api/graphcanon/tools/msoedov-langcorn","graph_url":"https://www.graphcanon.com/api/graphcanon/graph?tool=msoedov-langcorn","shared_categories":["inference-serving"]}]}}