{"data":{"slug":"ggml-org-llama-cpp","name":"llama.cpp","tagline":"LLM inference in C/C++","github_url":"https://github.com/ggml-org/llama.cpp","owner":"ggml-org","repo":"llama.cpp","owner_avatar_url":"https://avatars.githubusercontent.com/u/134263123?v=4","primary_language":"C++","stars":122941,"forks":21406,"topics":["ggml"],"archived":false,"github_pushed_at":"2026-08-07T05:28:54+00:00","maintenance_label":"Very active","stars_delta_30d":3353,"url":"https://www.graphcanon.com/tools/ggml-org-llama-cpp","markdown_url":"https://www.graphcanon.com/tools/ggml-org-llama-cpp.md","api_url":"https://www.graphcanon.com/api/graphcanon/tools/ggml-org-llama-cpp","graph_url":"https://www.graphcanon.com/api/graphcanon/graph?tool=ggml-org-llama-cpp","description":"LLM inference in C/C++","homepage_url":"https://llama.app","license":"MIT","open_issues":1969,"watchers":808,"ai_summary":"llama.cpp provides a framework for LLM inference using C++. It supports installation via package managers, Docker, pre-built binaries, and source builds.","readme_excerpt":"## Quick start\n\nA few options to get `llama.cpp` installed on your machine:\n\n- Visit https://llama.app and follow the instructions\n- Run with Docker - see our [Docker documentation](docs/docker.md)\n- Download pre-built binaries from the [releases page](https://github.com/ggml-org/llama.cpp/releases)\n- Build from source by cloning this repository - check out [our build guide](docs/build.md)\n\nOnce installed:\n\n```sh","github_created_at":"2023-03-10T18:58:00+00:00","created_at":"2026-07-07T22:37:23.741655+00:00","updated_at":"2026-08-07T06:01:07.93018+00:00","categories":[{"slug":"inference-serving","name":"Inference & Serving","url":"https://www.graphcanon.com/categories/inference-serving","markdown_url":"https://www.graphcanon.com/categories/inference-serving.md","api_url":"https://www.graphcanon.com/api/graphcanon/categories/inference-serving"}],"tags":[{"slug":"c","name":"c++"},{"slug":"ggml","name":"ggml"}],"trust":{"provenance":{"is_fork":false,"github_id":612354784,"owner_type":"Organization","methodology":"github_public_v1","parent_repo":null,"near_duplicate_slugs":[]},"computed_at":"2026-08-07T06:01:07.225Z","maintenance":{"label":"Very active","score":96,"methodology":"github_public_v1","releases_90d":30,"days_since_push":0,"last_release_at":"2026-08-06T22:10:39Z","stars_delta_30d":3353,"open_issues_delta_30d":143},"security_summary":{"status":"ok","scanner":"osv@v1","low_count":0,"high_count":0,"last_scan_at":"2026-07-11T10:36:54.049Z","medium_count":0,"scan_profile":"deps","critical_count":0}},"capability_facts":{"scan":{"source":"repo_scan","observed_at":"2026-08-07T06:01:07.672Z"},"has_cli":{"value":true,"source":"pyproject.toml:[project.scripts]","observed_at":"2026-08-07T06:01:07.672Z"},"languages":{"value":["c++","python"],"source":"github.language+pyproject.toml","observed_at":"2026-08-07T06:01:07.672Z"},"license_spdx":{"value":"MIT","source":"github.license","observed_at":"2026-08-07T06:01:07.672Z"}},"decision_facts":{"hosting":{"model":"unknown","summary":"llama.cpp supports various installation methods including package managers (like brew), Docker containers for isolation, pre-built binaries for ease of deployment, and source builds for flexibility."},"pricing":null,"requirements":{"notes":["Installation can be done via multiple channels including package managers, Docker, and direct downloads."],"min_ram_gb":null,"requires_docker":false},"constraints":{"min_ram_gb":null,"hosting_model":"unknown","requires_docker":false},"when_to_use":["- You need high-performance inference capabilities in a lightweight environment where C++ performance benefits are critical.","- Your deployment requires direct model quantization which can be optimized with the provided tools and framework."],"when_not_to_use":["- If you prefer a language other than C++, as this tool lacks support for Python or JavaScript bindings that provide higher-level abstractions.","- When your project demands extensive runtime customization and flexibility that is more easily achieved in languages like Python with libraries such as PyTorch."],"source":"enrich:decision_facts","observed_at":"2026-07-11T10:38:29.330Z"},"constraint_facets":{"min_ram_gb":null,"hosting_model":"unknown","requires_docker":false},"decision_summary":[{"label":"Hosting","value":"unknown - llama.cpp supports various installation methods including package managers (like brew), Docker containers for isolation, pre-built binaries for ease of deployment, and source builds for flexibility."},{"label":"Requirements","value":"Installation can be done via multiple channels including package managers, Docker, and direct downloads."},{"label":"Adopt for","value":"llama.cpp is a C++ framework for LLM inference, offering versatile installation options including package managers, Docker, and binary downloads."},{"label":"License detail","value":"MIT licensed, allowing free use and modification under certain conditions."}]}}