{"data":{"node":{"slug":"databendlabs-databend","name":"databend","tagline":"All-in-One Data Warehouse: Analytics, Search, AI, and Python Sandboxing Reimagined From Scratch.","github_url":"https://github.com/databendlabs/databend","owner":"databendlabs","repo":"databend","owner_avatar_url":"https://avatars.githubusercontent.com/u/80994548?v=4","primary_language":"Rust","stars":9420,"forks":891,"topics":["ai","bigdata","cloud-native","database","elasticsearch","geospatial","lakehouse","olap","rust","serverless","snowflake","sql","vector-database","vector-search"],"archived":false,"github_pushed_at":"2026-08-21T05:00:06+00:00","maintenance_label":"Very active","stars_delta_30d":31,"url":"https://www.graphcanon.com/tools/databendlabs-databend","markdown_url":"https://www.graphcanon.com/tools/databendlabs-databend.md","api_url":"https://www.graphcanon.com/api/graphcanon/tools/databendlabs-databend","graph_url":"https://www.graphcanon.com/api/graphcanon/graph?tool=databendlabs-databend"},"categories":[{"slug":"data-retrieval","name":"Data & Retrieval","url":"https://www.graphcanon.com/categories/data-retrieval","markdown_url":"https://www.graphcanon.com/categories/data-retrieval.md","api_url":"https://www.graphcanon.com/api/graphcanon/categories/data-retrieval"},{"slug":"vector-databases","name":"Vector Databases","url":"https://www.graphcanon.com/categories/vector-databases","markdown_url":"https://www.graphcanon.com/categories/vector-databases.md","api_url":"https://www.graphcanon.com/api/graphcanon/categories/vector-databases"}],"tags":[{"slug":"ai","name":"ai"},{"slug":"bigdata","name":"bigdata"},{"slug":"cloud-native","name":"cloud-native"},{"slug":"database","name":"database"},{"slug":"elasticsearch","name":"elasticsearch"},{"slug":"geospatial","name":"geospatial"},{"slug":"lakehouse","name":"lakehouse"},{"slug":"olap","name":"olap"}],"edges":[{"type":"integrates_with","direction":"out","explanation":"PageIndex aims to provide a document index for vectorless, reasoning-based RAG solutions, complementing Databend's analytical and search capabilities by enabling more detailed document interaction.","successor_context":null,"tool":{"slug":"vectifyai-pageindex","name":"PageIndex","tagline":"Document Index for Vectorless, Reasoning-based RAG","github_url":"https://github.com/VectifyAI/PageIndex","owner":"VectifyAI","repo":"PageIndex","owner_avatar_url":"https://avatars.githubusercontent.com/u/133959746?v=4","primary_language":"Python","stars":35204,"forks":3097,"topics":["agentic-ai","agents","ai","ai-agents","context-engineering","information-retrieval","llm","rag","reasoning","retrieval","retrieval-augmented-generation","vector-database"],"archived":false,"github_pushed_at":"2026-08-14T23:11:17+00:00","maintenance_label":"Very active","stars_delta_30d":1134,"url":"https://www.graphcanon.com/tools/vectifyai-pageindex","markdown_url":"https://www.graphcanon.com/tools/vectifyai-pageindex.md","api_url":"https://www.graphcanon.com/api/graphcanon/tools/vectifyai-pageindex","graph_url":"https://www.graphcanon.com/api/graphcanon/graph?tool=vectifyai-pageindex"}},{"type":"alternative","direction":"out","explanation":"Databend offers comprehensive capabilities including vector search among others, positioning it as an alternative to Qdrant, which specifically focuses on vector similarity search and management. Both tools cater to vector search needs but differ in scope and primary use cases.","successor_context":null,"tool":{"slug":"qdrant-qdrant","name":"qdrant","tagline":"High-performance, massive-scale Vector Database and Vector Search Engine","github_url":"https://github.com/qdrant/qdrant","owner":"qdrant","repo":"qdrant","owner_avatar_url":"https://avatars.githubusercontent.com/u/73504361?v=4","primary_language":"Rust","stars":33629,"forks":2529,"topics":["ai-search","ai-search-engine","embeddings-similarity","hnsw","hybrid-search","image-search","knn-algorithm","machine-learning","mlops","nearest-neighbor-search","neural-network","neural-search","recommender-system","search","search-engine","search-engines","similarity-search","vector-database","vector-search","vector-search-engine"],"archived":false,"github_pushed_at":"2026-07-28T17:03:13+00:00","maintenance_label":"Very active","url":"https://www.graphcanon.com/tools/qdrant-qdrant","markdown_url":"https://www.graphcanon.com/tools/qdrant-qdrant.md","api_url":"https://www.graphcanon.com/api/graphcanon/tools/qdrant-qdrant","graph_url":"https://www.graphcanon.com/api/graphcanon/graph?tool=qdrant-qdrant"}},{"type":"alternative","direction":"out","explanation":"MatrixOne and Databend are both HTAP databases built with capabilities for AI applications such as vector search. They serve similar enterprise data analytics needs.","successor_context":null,"tool":{"slug":"matrixorigin-matrixone","name":"matrixone","tagline":"AI-native HTAP database with Git-for-Data and built-in vector search","github_url":"https://github.com/matrixorigin/matrixone","owner":"matrixorigin","repo":"matrixone","owner_avatar_url":"https://avatars.githubusercontent.com/u/76932962?v=4","primary_language":"Go","stars":1879,"forks":308,"topics":["agents","ai-native","cloud-native","database","distributed-database","distributed-systems","fulltext-support","git-for-data","go","htap","hyperconverged","memory","mysql-compatible","olap","one-size-fits-all","sql","vector-database"],"archived":false,"github_pushed_at":"2026-08-21T11:31:20+00:00","maintenance_label":"Very active","stars_delta_30d":18,"url":"https://www.graphcanon.com/tools/matrixorigin-matrixone","markdown_url":"https://www.graphcanon.com/tools/matrixorigin-matrixone.md","api_url":"https://www.graphcanon.com/api/graphcanon/tools/matrixorigin-matrixone","graph_url":"https://www.graphcanon.com/api/graphcanon/graph?tool=matrixorigin-matrixone"}},{"type":"alternative","direction":"out","explanation":"Databend and infinity both provide data warehousing capabilities for AI workloads, including hybrid search of dense vector, sparse vector, full-text, and more.","successor_context":null,"tool":{"slug":"infiniflow-infinity","name":"infinity","tagline":"AI-native database for LLM applications offering fast hybrid search capabilities.","github_url":"https://github.com/infiniflow/infinity","owner":"infiniflow","repo":"infinity","owner_avatar_url":"https://avatars.githubusercontent.com/u/69962740?v=4","primary_language":"C++","stars":4675,"forks":437,"topics":["ai-native","approximate-nearest-neighbor-search","bm25","cpp20","cpp20-modules","embedding","full-text-search","hnsw","hybrid-search","information-retrival","multi-vector","nearest-neighbor-search","rag","search-engine","tensor-database","vector","vector-database","vector-search","vectordatabase"],"archived":false,"github_pushed_at":"2026-08-17T13:43:09+00:00","maintenance_label":"Very active","stars_delta_30d":51,"url":"https://www.graphcanon.com/tools/infiniflow-infinity","markdown_url":"https://www.graphcanon.com/tools/infiniflow-infinity.md","api_url":"https://www.graphcanon.com/api/graphcanon/tools/infiniflow-infinity","graph_url":"https://www.graphcanon.com/api/graphcanon/graph?tool=infiniflow-infinity"}},{"type":"alternative","direction":"out","explanation":"Databend and Milvus both offer vector search capabilities, but Databend integrates this feature within a broader data analysis and warehousing platform, whereas Milvus is primarily focused on being a high-performance cloud-native vector database.","successor_context":null,"tool":{"slug":"milvus-io-milvus","name":"milvus","tagline":"High-performance cloud-native vector database","github_url":"https://github.com/milvus-io/milvus","owner":"milvus-io","repo":"milvus","owner_avatar_url":"https://avatars.githubusercontent.com/u/51735404?v=4","primary_language":"Go","stars":45402,"forks":4147,"topics":["anns","cloud-native","diskann","distributed","embedding-database","embedding-similarity","embedding-store","faiss","golang","hnsw","image-search","llm","nearest-neighbor-search","rag","vector-database","vector-search","vector-similarity","vector-store"],"archived":false,"github_pushed_at":"2026-07-28T17:42:13+00:00","maintenance_label":"Very active","url":"https://www.graphcanon.com/tools/milvus-io-milvus","markdown_url":"https://www.graphcanon.com/tools/milvus-io-milvus.md","api_url":"https://www.graphcanon.com/api/graphcanon/tools/milvus-io-milvus","graph_url":"https://www.graphcanon.com/api/graphcanon/graph?tool=milvus-io-milvus"}},{"type":"integrates_with","direction":"out","explanation":"Databend, as an enterprise data warehouse designed for large-scale analytics and vector search, integrates with Graphiti, a framework for constructing real-time knowledge graphs, to enable AI agents to perform complex queries on dynamically evolving context graphs that Graphiti maintains. This integration allows Databend to leverage Graphiti's temporal context graph capabilities to enhance its own","successor_context":null,"tool":{"slug":"getzep-graphiti","name":"graphiti","tagline":"Build Real-Time Knowledge Graphs for AI Agents","github_url":"https://github.com/getzep/graphiti","owner":"getzep","repo":"graphiti","owner_avatar_url":"https://avatars.githubusercontent.com/u/132832125?v=4","primary_language":"Python","stars":30018,"forks":3042,"topics":["agents","graph","llms","rag"],"archived":false,"github_pushed_at":"2026-08-17T23:49:16+00:00","maintenance_label":"Very active","stars_delta_30d":1143,"url":"https://www.graphcanon.com/tools/getzep-graphiti","markdown_url":"https://www.graphcanon.com/tools/getzep-graphiti.md","api_url":"https://www.graphcanon.com/api/graphcanon/tools/getzep-graphiti","graph_url":"https://www.graphcanon.com/api/graphcanon/graph?tool=getzep-graphiti"}},{"type":"integrates_with","direction":"out","explanation":"Chroma can integrate with Databend as it provides an infrastructure for searching AI-generated content that could be stored or processed in Databend’s data warehouse.","successor_context":null,"tool":{"slug":"chroma-core-chroma","name":"chroma","tagline":"Search infrastructure for AI","github_url":"https://github.com/chroma-core/chroma","owner":"chroma-core","repo":"chroma","owner_avatar_url":"https://avatars.githubusercontent.com/u/105881770?v=4","primary_language":"Rust","stars":28898,"forks":2409,"topics":["agents","ai","ai-agents","database","rust","rust-lang"],"archived":false,"github_pushed_at":"2026-07-27T22:15:20+00:00","maintenance_label":"Very active","url":"https://www.graphcanon.com/tools/chroma-core-chroma","markdown_url":"https://www.graphcanon.com/tools/chroma-core-chroma.md","api_url":"https://www.graphcanon.com/api/graphcanon/tools/chroma-core-chroma","graph_url":"https://www.graphcanon.com/api/graphcanon/graph?tool=chroma-core-chroma"}},{"type":"alternative","direction":"out","explanation":"Databend includes native support for vector similarity searches while pgvector extends PostgreSQL to offer similar functionalities; both are aimed at providing efficient vector search capabilities, but within different database contexts.","successor_context":null,"tool":{"slug":"pgvector-pgvector","name":"pgvector","tagline":"Open-source vector similarity search for Postgres","github_url":"https://github.com/pgvector/pgvector","owner":"pgvector","repo":"pgvector","owner_avatar_url":"https://avatars.githubusercontent.com/u/98363230?v=4","primary_language":"C","stars":22375,"forks":1257,"topics":["approximate-nearest-neighbor-search","nearest-neighbor-search"],"archived":false,"github_pushed_at":"2026-07-28T09:49:19+00:00","maintenance_label":"Very active","url":"https://www.graphcanon.com/tools/pgvector-pgvector","markdown_url":"https://www.graphcanon.com/tools/pgvector-pgvector.md","api_url":"https://www.graphcanon.com/api/graphcanon/tools/pgvector-pgvector","graph_url":"https://www.graphcanon.com/api/graphcanon/graph?tool=pgvector-pgvector"}},{"type":"integrates_with","direction":"out","explanation":"Databend can integrate with pgvector to provide extended vector similarity search capabilities for Postgres users, enhancing its use cases in hybrid database environments.","successor_context":null,"tool":{"slug":"pgvector-pgvector","name":"pgvector","tagline":"Open-source vector similarity search for Postgres","github_url":"https://github.com/pgvector/pgvector","owner":"pgvector","repo":"pgvector","owner_avatar_url":"https://avatars.githubusercontent.com/u/98363230?v=4","primary_language":"C","stars":22375,"forks":1257,"topics":["approximate-nearest-neighbor-search","nearest-neighbor-search"],"archived":false,"github_pushed_at":"2026-07-28T09:49:19+00:00","maintenance_label":"Very active","url":"https://www.graphcanon.com/tools/pgvector-pgvector","markdown_url":"https://www.graphcanon.com/tools/pgvector-pgvector.md","api_url":"https://www.graphcanon.com/api/graphcanon/tools/pgvector-pgvector","graph_url":"https://www.graphcanon.com/api/graphcanon/graph?tool=pgvector-pgvector"}},{"type":"integrates_with","direction":"out","explanation":"Databend can integrate with MeiliSearch to provide real-time, full-text search capabilities for both structured and unstructured data.","successor_context":null,"tool":{"slug":"meilisearch-meilisearch","name":"meilisearch","tagline":"A lightning-fast search engine API bringing AI-powered hybrid search to your sites and applications.","github_url":"https://github.com/meilisearch/meilisearch","owner":"meilisearch","repo":"meilisearch","owner_avatar_url":"https://avatars.githubusercontent.com/u/43250847?v=4","primary_language":"Rust","stars":59034,"forks":2672,"topics":["ai","api","app-search","database","enterprise-search","faceting","full-text-search","fuzzy-search","geosearch","hybrid-search","instantsearch","search","search-as-you-type","search-engine","semantic-search","site-search","typo-tolerance","vector-database","vector-search","vectors"],"archived":false,"github_pushed_at":"2026-08-14T09:38:01+00:00","maintenance_label":"Very active","stars_delta_30d":352,"url":"https://www.graphcanon.com/tools/meilisearch-meilisearch","markdown_url":"https://www.graphcanon.com/tools/meilisearch-meilisearch.md","api_url":"https://www.graphcanon.com/api/graphcanon/tools/meilisearch-meilisearch","graph_url":"https://www.graphcanon.com/api/graphcanon/graph?tool=meilisearch-meilisearch"}},{"type":"related","direction":"in","explanation":"Both are databases with support for AI applications, but Databend focuses on data warehousing while Milvus is a vector database.","successor_context":null,"tool":{"slug":"milvus-io-milvus","name":"milvus","tagline":"High-performance cloud-native vector database","github_url":"https://github.com/milvus-io/milvus","owner":"milvus-io","repo":"milvus","owner_avatar_url":"https://avatars.githubusercontent.com/u/51735404?v=4","primary_language":"Go","stars":45402,"forks":4147,"topics":["anns","cloud-native","diskann","distributed","embedding-database","embedding-similarity","embedding-store","faiss","golang","hnsw","image-search","llm","nearest-neighbor-search","rag","vector-database","vector-search","vector-similarity","vector-store"],"archived":false,"github_pushed_at":"2026-07-28T17:42:13+00:00","maintenance_label":"Very active","url":"https://www.graphcanon.com/tools/milvus-io-milvus","markdown_url":"https://www.graphcanon.com/tools/milvus-io-milvus.md","api_url":"https://www.graphcanon.com/api/graphcanon/tools/milvus-io-milvus","graph_url":"https://www.graphcanon.com/api/graphcanon/graph?tool=milvus-io-milvus"}},{"type":"related","direction":"in","explanation":"Both are AI-focused databases (HTAP for Databend, vector search extension in Postgres for pgvecto.rs) offering scalable solutions.","successor_context":null,"tool":{"slug":"tensorchord-pgvecto-rs","name":"pgvecto.rs","tagline":"Scalable, Low-latency and Hybrid-enabled Vector Search in Postgres","github_url":"https://github.com/tensorchord/pgvecto.rs","owner":"tensorchord","repo":"pgvecto.rs","owner_avatar_url":"https://avatars.githubusercontent.com/u/100543303?v=4","primary_language":"Rust","stars":2183,"forks":86,"topics":["chatgpt","faiss","gpt","hacktoberfest","llm","nearest-neighbor-search","postgres","rust","vector","vector-database"],"archived":false,"github_pushed_at":"2025-02-26T14:11:43+00:00","maintenance_label":"Dormant","stars_delta_30d":4,"url":"https://www.graphcanon.com/tools/tensorchord-pgvecto-rs","markdown_url":"https://www.graphcanon.com/tools/tensorchord-pgvecto-rs.md","api_url":"https://www.graphcanon.com/api/graphcanon/tools/tensorchord-pgvecto-rs","graph_url":"https://www.graphcanon.com/api/graphcanon/graph?tool=tensorchord-pgvecto-rs"}},{"type":"alternative","direction":"in","explanation":"Deeplake and Databend both aim to provide scalable data management solutions, though Deeplake focuses on a multimodal datalake with support for AI agents, while Databend is more oriented towards enterprise data warehousing.","successor_context":null,"tool":{"slug":"activeloopai-deeplake","name":"deeplake","tagline":"AI Data Runtime for Agents with scalable retrieval and training features","github_url":"https://github.com/activeloopai/deeplake","owner":"activeloopai","repo":"deeplake","owner_avatar_url":"https://avatars.githubusercontent.com/u/34816118?v=4","primary_language":"C++","stars":9224,"forks":721,"topics":["agent","agentic-rag","ai","clawbot","computer-vision","datalake","deep-learning","filesystem","large-language-models","llm","memory","mlops","multimodal","openclaw","postgres","pytorch","rag","skill","vector-database"],"archived":false,"github_pushed_at":"2026-05-21T15:28:00+00:00","maintenance_label":"Steady","stars_delta_30d":16,"url":"https://www.graphcanon.com/tools/activeloopai-deeplake","markdown_url":"https://www.graphcanon.com/tools/activeloopai-deeplake.md","api_url":"https://www.graphcanon.com/api/graphcanon/tools/activeloopai-deeplake","graph_url":"https://www.graphcanon.com/api/graphcanon/graph?tool=activeloopai-deeplake"}},{"type":"alternative","direction":"in","explanation":"Databend is also an enterprise data warehouse that can serve AI-related applications. Infinity and Databend both target high-performance hybrid search capabilities needed for modern AI applications.","successor_context":null,"tool":{"slug":"infiniflow-infinity","name":"infinity","tagline":"AI-native database for LLM applications offering fast hybrid search capabilities.","github_url":"https://github.com/infiniflow/infinity","owner":"infiniflow","repo":"infinity","owner_avatar_url":"https://avatars.githubusercontent.com/u/69962740?v=4","primary_language":"C++","stars":4675,"forks":437,"topics":["ai-native","approximate-nearest-neighbor-search","bm25","cpp20","cpp20-modules","embedding","full-text-search","hnsw","hybrid-search","information-retrival","multi-vector","nearest-neighbor-search","rag","search-engine","tensor-database","vector","vector-database","vector-search","vectordatabase"],"archived":false,"github_pushed_at":"2026-08-17T13:43:09+00:00","maintenance_label":"Very active","stars_delta_30d":51,"url":"https://www.graphcanon.com/tools/infiniflow-infinity","markdown_url":"https://www.graphcanon.com/tools/infiniflow-infinity.md","api_url":"https://www.graphcanon.com/api/graphcanon/tools/infiniflow-infinity","graph_url":"https://www.graphcanon.com/api/graphcanon/graph?tool=infiniflow-infinity"}},{"type":"related","direction":"in","explanation":"Databend is a data warehouse for AI agents, whereas DingoDB focuses on being a multi-modal vector database offering high-level SQL access and real-time strong consistency among other features. They serve related but distinct purposes within the same ecosystem.","successor_context":null,"tool":{"slug":"dingodb-dingo","name":"dingo","tagline":"A multi-modal vector database that supports upserts and vector queries using unified SQL (MySQL-Compatible) on structured and unstructured data","github_url":"https://github.com/dingodb/dingo","owner":"dingodb","repo":"dingo","owner_avatar_url":"https://avatars.githubusercontent.com/u/91237812?v=4","primary_language":"Java","stars":1701,"forks":265,"topics":["embedding-search","embedding-store","hybrid-search","key-value-distributed-store","mysql-compatibility","real-time-semantic-search","serving","structured-data","unified-sql","unstructured-data","vector-database","vector-ocean"],"archived":false,"github_pushed_at":"2026-07-10T11:13:29+00:00","maintenance_label":"Steady","stars_delta_30d":2,"url":"https://www.graphcanon.com/tools/dingodb-dingo","markdown_url":"https://www.graphcanon.com/tools/dingodb-dingo.md","api_url":"https://www.graphcanon.com/api/graphcanon/tools/dingodb-dingo","graph_url":"https://www.graphcanon.com/api/graphcanon/graph?tool=dingodb-dingo"}}],"neighbours":[{"slug":"pathwaycom-llm-app","name":"llm-app","tagline":"Ready-to-run cloud templates for RAG, AI pipelines, and enterprise search with live data.","github_url":"https://github.com/pathwaycom/llm-app","owner":"pathwaycom","repo":"llm-app","owner_avatar_url":"https://avatars.githubusercontent.com/u/25750857?v=4","primary_language":"Jupyter Notebook","stars":59037,"forks":1466,"topics":["chatbot","hugging-face","llm","llm-local","llm-prompting","llm-security","llmops","machine-learning","open-ai","pathway","rag","real-time","retrieval-augmented-generation","vector-database","vector-index"],"archived":false,"github_pushed_at":"2026-07-05T17:59:07+00:00","maintenance_label":"Steady","url":"https://www.graphcanon.com/tools/pathwaycom-llm-app","markdown_url":"https://www.graphcanon.com/tools/pathwaycom-llm-app.md","api_url":"https://www.graphcanon.com/api/graphcanon/tools/pathwaycom-llm-app","graph_url":"https://www.graphcanon.com/api/graphcanon/graph?tool=pathwaycom-llm-app","shared_categories":["vector-databases","data-retrieval"]},{"slug":"zylon-ai-private-gpt","name":"private-gpt","tagline":"Complete API layer for private AI applications on local models","github_url":"https://github.com/zylon-ai/private-gpt","owner":"zylon-ai","repo":"private-gpt","owner_avatar_url":"https://avatars.githubusercontent.com/u/143802295?v=4","primary_language":"Python","stars":57415,"forks":7607,"topics":["ai","ai-tools","on-premise"],"archived":false,"github_pushed_at":"2026-08-06T13:41:08+00:00","maintenance_label":"Active","url":"https://www.graphcanon.com/tools/zylon-ai-private-gpt","markdown_url":"https://www.graphcanon.com/tools/zylon-ai-private-gpt.md","api_url":"https://www.graphcanon.com/api/graphcanon/tools/zylon-ai-private-gpt","graph_url":"https://www.graphcanon.com/api/graphcanon/graph?tool=zylon-ai-private-gpt","shared_categories":[]},{"slug":"sinaptik-ai-pandas-ai","name":"pandas-ai","tagline":"Chat with your database or your datalake using LLMs and RAG.","github_url":"https://github.com/sinaptik-ai/pandas-ai","owner":"sinaptik-ai","repo":"pandas-ai","owner_avatar_url":"https://avatars.githubusercontent.com/u/154438448?v=4","primary_language":"Python","stars":23746,"forks":2342,"topics":["ai","csv","data","data-analysis","data-science","data-visualization","database","datalake","gpt-4","llm","pandas","sql","text-to-sql"],"archived":false,"github_pushed_at":"2025-10-28T10:02:13+00:00","maintenance_label":"Slowing","url":"https://www.graphcanon.com/tools/sinaptik-ai-pandas-ai","markdown_url":"https://www.graphcanon.com/tools/sinaptik-ai-pandas-ai.md","api_url":"https://www.graphcanon.com/api/graphcanon/tools/sinaptik-ai-pandas-ai","graph_url":"https://www.graphcanon.com/api/graphcanon/graph?tool=sinaptik-ai-pandas-ai","shared_categories":["data-retrieval"]},{"slug":"canner-wrenai","name":"WrenAI","tagline":"GenBI for AI agents, turns natural-language questions into trusted dashboards and SQL","github_url":"https://github.com/Canner/WrenAI","owner":"Canner","repo":"WrenAI","owner_avatar_url":"https://avatars.githubusercontent.com/u/7250217?v=4","primary_language":"Python","stars":17295,"forks":1955,"topics":["ai-agents","bigquery","business-intelligence","charts","clickhouse","context-engineering","dashboard","databricks","duckdb","genbi","generative-ai","llm","mcp","postgresql","rag","semantic-layer","snowflake","sql","text-to-sql","text2sql"],"archived":false,"github_pushed_at":"2026-08-18T05:19:19+00:00","maintenance_label":"Very active","url":"https://www.graphcanon.com/tools/canner-wrenai","markdown_url":"https://www.graphcanon.com/tools/canner-wrenai.md","api_url":"https://www.graphcanon.com/api/graphcanon/tools/canner-wrenai","graph_url":"https://www.graphcanon.com/api/graphcanon/graph?tool=canner-wrenai","shared_categories":["data-retrieval"]},{"slug":"neuml-txtai","name":"txtai","tagline":"All-in-one AI framework for semantic search, LLM orchestration and language model workflows","github_url":"https://github.com/neuml/txtai","owner":"neuml","repo":"txtai","owner_avatar_url":"https://avatars.githubusercontent.com/u/59890304?v=4","primary_language":"Python","stars":12890,"forks":873,"topics":["agents","ai","ai-agents","embeddings","information-retrieval","language-model","large-language-models","llm","nlp","python","rag","retrieval-augmented-generation","search","search-engine","semantic-search","sentence-embeddings","transformers","txtai","vector-database","vector-search"],"archived":false,"github_pushed_at":"2026-08-12T13:42:39+00:00","maintenance_label":"Active","url":"https://www.graphcanon.com/tools/neuml-txtai","markdown_url":"https://www.graphcanon.com/tools/neuml-txtai.md","api_url":"https://www.graphcanon.com/api/graphcanon/tools/neuml-txtai","graph_url":"https://www.graphcanon.com/api/graphcanon/graph?tool=neuml-txtai","shared_categories":["data-retrieval"]},{"slug":"mage-ai-mage-ai","name":"mage-ai","tagline":"Build, run and manage data pipelines for integrating and transforming data","github_url":"https://github.com/mage-ai/mage-ai","owner":"mage-ai","repo":"mage-ai","owner_avatar_url":"https://avatars.githubusercontent.com/u/69371472?v=4","primary_language":"Python","stars":8790,"forks":982,"topics":["artificial-intelligence","data","data-engineering","data-integration","data-pipelines","data-science","dbt","elt","etl","machine-learning","orchestration","pipeline","pipelines","python","reverse-etl","spark","sql","transformation"],"archived":false,"github_pushed_at":"2026-08-10T23:12:25+00:00","maintenance_label":"Active","url":"https://www.graphcanon.com/tools/mage-ai-mage-ai","markdown_url":"https://www.graphcanon.com/tools/mage-ai-mage-ai.md","api_url":"https://www.graphcanon.com/api/graphcanon/tools/mage-ai-mage-ai","graph_url":"https://www.graphcanon.com/api/graphcanon/graph?tool=mage-ai-mage-ai","shared_categories":["data-retrieval"]},{"slug":"zilliztech-deep-searcher","name":"deep-searcher","tagline":"Open Source Deep Research Alternative to Reason and Search on Private Data.","github_url":"https://github.com/zilliztech/deep-searcher","owner":"zilliztech","repo":"deep-searcher","owner_avatar_url":"https://avatars.githubusercontent.com/u/18416694?v=4","primary_language":"Python","stars":8060,"forks":775,"topics":["agent","agentic-rag","claude","deep-research","deepseek","deepseek-r1","grok","grok3","llama4","llm","milvus","openai","qwen3","rag","reasoning-models","vector-database","zilliz"],"archived":false,"github_pushed_at":"2025-11-19T06:04:16+00:00","maintenance_label":"Slowing","url":"https://www.graphcanon.com/tools/zilliztech-deep-searcher","markdown_url":"https://www.graphcanon.com/tools/zilliztech-deep-searcher.md","api_url":"https://www.graphcanon.com/api/graphcanon/tools/zilliztech-deep-searcher","graph_url":"https://www.graphcanon.com/api/graphcanon/graph?tool=zilliztech-deep-searcher","shared_categories":["vector-databases"]},{"slug":"ucbepic-docetl","name":"docetl","tagline":"A system for agentic LLM-powered data processing and ETL","github_url":"https://github.com/ucbepic/docetl","owner":"ucbepic","repo":"docetl","owner_avatar_url":"https://avatars.githubusercontent.com/u/88680502?v=4","primary_language":"Python","stars":3961,"forks":421,"topics":["agents","data","data-pipelines","document-analysis","document-processing","elt","etl","llm","python","semantic-data","unstructured-data","unstructured-data-analysis","workflow"],"archived":false,"github_pushed_at":"2026-08-09T23:31:04+00:00","maintenance_label":"Active","url":"https://www.graphcanon.com/tools/ucbepic-docetl","markdown_url":"https://www.graphcanon.com/tools/ucbepic-docetl.md","api_url":"https://www.graphcanon.com/api/graphcanon/tools/ucbepic-docetl","graph_url":"https://www.graphcanon.com/api/graphcanon/graph?tool=ucbepic-docetl","shared_categories":["data-retrieval"]},{"slug":"spiceai-spiceai","name":"spiceai","tagline":"A real-time analytics node for data-grounded AI applications","github_url":"https://github.com/spiceai/spiceai","owner":"spiceai","repo":"spiceai","owner_avatar_url":"https://avatars.githubusercontent.com/u/73862742?v=4","primary_language":"Rust","stars":3047,"forks":212,"topics":["artificial-intelligence","data","data-federation","developers","full-text-search","infrastructure","llm-inference","machine-learning","sql"],"archived":false,"github_pushed_at":"2026-07-25T05:47:50+00:00","maintenance_label":"Active","url":"https://www.graphcanon.com/tools/spiceai-spiceai","markdown_url":"https://www.graphcanon.com/tools/spiceai-spiceai.md","api_url":"https://www.graphcanon.com/api/graphcanon/tools/spiceai-spiceai","graph_url":"https://www.graphcanon.com/api/graphcanon/graph?tool=spiceai-spiceai","shared_categories":["data-retrieval"]},{"slug":"gmpetrov-databerry","name":"databerry","tagline":"The no-code platform for building custom LLM Agents","github_url":"https://github.com/gmpetrov/databerry","owner":"gmpetrov","repo":"databerry","owner_avatar_url":"https://avatars.githubusercontent.com/u/4693180?v=4","primary_language":null,"stars":2965,"forks":420,"topics":["ai","aichatbot","chatbot","chatbots","chatgpt","llm","no-code","openai","qdrant","semantic-search","typescript"],"archived":false,"github_pushed_at":"2024-06-17T12:36:06+00:00","maintenance_label":"Dormant","url":"https://www.graphcanon.com/tools/gmpetrov-databerry","markdown_url":"https://www.graphcanon.com/tools/gmpetrov-databerry.md","api_url":"https://www.graphcanon.com/api/graphcanon/tools/gmpetrov-databerry","graph_url":"https://www.graphcanon.com/api/graphcanon/graph?tool=gmpetrov-databerry","shared_categories":[]},{"slug":"openlit-openlit","name":"openlit","tagline":"A comprehensive open-source platform for AI Engineering with LLM Observability, Monitoring, and Management","github_url":"https://github.com/openlit/openlit","owner":"openlit","repo":"openlit","owner_avatar_url":"https://avatars.githubusercontent.com/u/149867240?v=4","primary_language":"TypeScript","stars":2664,"forks":342,"topics":["ai-observability","amd-gpu","clickhouse","distributed-tracing","genai","gpu-monitoring","grafana","langchain","llmops","llms","metrics","monitoring-tool","nvidia-smi","observability","open-source","openai","opentelemetry","otlp","python","tracing"],"archived":false,"github_pushed_at":"2026-07-31T18:39:37+00:00","maintenance_label":"Active","url":"https://www.graphcanon.com/tools/openlit-openlit","markdown_url":"https://www.graphcanon.com/tools/openlit-openlit.md","api_url":"https://www.graphcanon.com/api/graphcanon/tools/openlit-openlit","graph_url":"https://www.graphcanon.com/api/graphcanon/graph?tool=openlit-openlit","shared_categories":[]},{"slug":"huggingface-aisheets","name":"aisheets","tagline":"Build, enrich, and transform datasets using AI models with no code","github_url":"https://github.com/huggingface/aisheets","owner":"huggingface","repo":"aisheets","owner_avatar_url":"https://avatars.githubusercontent.com/u/25720743?v=4","primary_language":"TypeScript","stars":1638,"forks":140,"topics":["ai","llm-evaluation","llms","nocode","oss","synthetic-data"],"archived":false,"github_pushed_at":"2026-05-26T10:33:23+00:00","maintenance_label":"Steady","url":"https://www.graphcanon.com/tools/huggingface-aisheets","markdown_url":"https://www.graphcanon.com/tools/huggingface-aisheets.md","api_url":"https://www.graphcanon.com/api/graphcanon/tools/huggingface-aisheets","graph_url":"https://www.graphcanon.com/api/graphcanon/graph?tool=huggingface-aisheets","shared_categories":["data-retrieval"]}]}}