{"data":{"slug":"maximhq-bifrost","name":"bifrost","tagline":"Fast Enterprise AI Gateway with Adaptive Load Balancer and Guardrails","github_url":"https://github.com/maximhq/bifrost","owner":"maximhq","repo":"bifrost","owner_avatar_url":"https://avatars.githubusercontent.com/u/139708451?v=4","primary_language":"Go","stars":7449,"forks":1073,"topics":["ai-gateway","gateway","gateway-services","generative-ai","guardrails","llm","llm-cost","llm-gateway","llm-observability","llmops","load-balancing","mcp-client","mcp-gateway","mcp-server","model-router","token-management"],"archived":false,"github_pushed_at":"2026-08-20T11:57:33+00:00","maintenance_label":"Very active","stars_delta_30d":812,"url":"https://www.graphcanon.com/tools/maximhq-bifrost","markdown_url":"https://www.graphcanon.com/tools/maximhq-bifrost.md","api_url":"https://www.graphcanon.com/api/graphcanon/tools/maximhq-bifrost","graph_url":"https://www.graphcanon.com/api/graphcanon/graph?tool=maximhq-bifrost","description":"Fastest enterprise AI gateway (50x faster than LiteLLM) with adaptive load balancer, cluster mode, guardrails, 1000+ models support & <100 µs overhead at 5k RPS.","homepage_url":"https://www.getmaxim.ai/bifrost","license":"Apache-2.0","open_issues":910,"watchers":28,"ai_summary":"Bifrost is an ultra-fast enterprise gateway for AI that supports over 1000 models with minimal overhead.","readme_excerpt":"## Quick Start\n\n\n\n**Go from zero to production-ready AI gateway in under a minute.**\n\n**Step 1:** Start Bifrost Gateway\n\n```bash\n\n---\n\n# Install and run locally\nnpx -y @maximhq/bifrost\n\n---\n\n# Or use Docker\ndocker run -p 8080:8080 maximhq/bifrost\n```\n\n**Step 2:** Configure via Web UI\n\n```bash\n\n---\n\n### Core Infrastructure\n\n- **[Unified Interface](https://docs.getbifrost.ai/providers/supported-providers/overview)** - Single OpenAI-compatible API for all providers\n- **[Multi-Provider Support](https://docs.getbifrost.ai/quickstart/gateway/provider-configuration)** - OpenAI, Anthropic, AWS Bedrock, Google Vertex, Azure, Cerebras, Cohere, Mistral, Ollama, Groq, and more\n- **[Automatic Fallbacks](https://docs.getbifrost.ai/features/retries-and-fallbacks)** - Seamless failover between providers and models with zero downtime\n- **[Load Balancing](https://docs.getbifrost.ai/features/retries-and-fallbacks)** - Intelligent request distribution across multiple API keys and providers\n\n---\n\n## Getting Started Options\n\nChoose the deployment method that fits your needs:\n\n---\n\n# Docker - Production ready\ndocker run -p 8080:8080 -v $(pwd)/data:/app/data maximhq/bifrost\n```\n\n**Features:** Web UI, real-time monitoring, multi-provider management, zero-config startup\n\n**Learn More:** [Gateway Setup Guide](https://docs.getbifrost.ai/quickstart/gateway/setting-up)\n\n---\n\n### Quick Start\n\n- [Gateway Setup](https://docs.getbifrost.ai/quickstart/gateway/setting-up) - HTTP API deployment in 30 seconds\n- [Go SDK Setup](https://docs.getbifrost.ai/quickstart/go-sdk/setting-up) - Direct Go integration\n- [Provider Configuration](https://docs.getbifrost.ai/quickstart/gateway/provider-configuration) - Multi-provider setup\n\n---\n\n## License\n\nThis project is licensed under the Apache 2.0 License - see the [LICENSE](LICENSE) file for details.\n\nBuilt with ❤️ by [Maxim](https://github.com/maximhq)","github_created_at":"2025-03-19T07:21:26+00:00","created_at":"2026-07-07T17:42:00.148686+00:00","updated_at":"2026-08-20T12:01:38.33968+00:00","categories":[{"slug":"inference-serving","name":"Inference & Serving","url":"https://www.graphcanon.com/categories/inference-serving","markdown_url":"https://www.graphcanon.com/categories/inference-serving.md","api_url":"https://www.graphcanon.com/api/graphcanon/categories/inference-serving"}],"tags":[{"slug":"ai-gateway","name":"ai-gateway"},{"slug":"gateway-services","name":"gateway-services"},{"slug":"generative-ai","name":"generative-ai"},{"slug":"guardrails","name":"guardrails"},{"slug":"llm-cost","name":"llm-cost"},{"slug":"load-balancing","name":"load-balancing"},{"slug":"model-router","name":"model-router"}],"trust":{"provenance":{"is_fork":false,"github_id":951115072,"owner_type":"Organization","methodology":"github_public_v1","parent_repo":null,"near_duplicate_slugs":[]},"computed_at":"2026-08-20T12:01:37.606Z","maintenance":{"label":"Very active","score":96,"methodology":"github_public_v1","releases_90d":30,"days_since_push":0,"last_release_at":"2026-08-19T05:38:08Z","stars_delta_30d":812,"open_issues_delta_30d":268},"security_summary":{"status":"no_lockfile","scanner":null,"low_count":0,"high_count":0,"last_scan_at":"2026-07-12T04:02:02.843Z","medium_count":0,"scan_profile":"none","critical_count":0}},"capability_facts":{"scan":{"source":"repo_scan","observed_at":"2026-08-20T12:01:38.052Z"},"languages":{"value":["go"],"source":"github.language","observed_at":"2026-08-20T12:01:38.052Z"},"license_spdx":{"value":"Apache-2.0","source":"github.license","observed_at":"2026-08-20T12:01:38.052Z"}},"decision_facts":{"hosting":null,"pricing":null,"requirements":null,"constraints":null,"when_to_use":["Use when you require a highly scalable solution that can manage more than 1000 different models with minimal overhead.","Ideal if your deployment demands ultra-low latencies of under 100 microseconds at 5,000 requests per second.","Choose this tool for seamless multi-provider support and automatic fallbacks to ensure service continuity."],"when_not_to_use":["Avoid Bifrost if you are working within a restrictive environment where Go is not the preferred language or Apache-2.0 licensing terms cannot be accepted.","Do not use if you prioritize tools that offer more customization options beyond its unified open-source compatible API and default configurations.","Consider alternatives when your requirements surpass existing load handling capabilities or you need features not covered by Bifrost's focus on speed and diversity."],"source":"enrich:decision_facts","observed_at":"2026-07-14T18:59:49.681Z"},"constraint_facets":null,"decision_summary":[{"label":"Adopt for","value":"Bifrost is an ultra-fast AI gateway with adaptive load balancing and support for over 1000 models, suitable for enterprises requiring low latency and high model diversity in their API serving needs."}]}}