{"data":{"node":{"slug":"openrlhf-openrlhf","name":"OpenRLHF","tagline":"Easy-to-use, Scalable and High-performance Agentic RL Framework based on Ray","github_url":"https://github.com/OpenRLHF/OpenRLHF","owner":"OpenRLHF","repo":"OpenRLHF","owner_avatar_url":"https://avatars.githubusercontent.com/u/175771028?v=4","primary_language":"Python","stars":9891,"forks":996,"topics":["large-language-models","proximal-policy-optimization","raylib","reinforcement-learning","reinforcement-learning-from-human-feedback","transformers","visual-language-models","vllm"],"archived":false,"github_pushed_at":"2026-07-14T01:57:21+00:00","maintenance_label":"Active","stars_delta_30d":132,"url":"https://www.graphcanon.com/tools/openrlhf-openrlhf","markdown_url":"https://www.graphcanon.com/tools/openrlhf-openrlhf.md","api_url":"https://www.graphcanon.com/api/graphcanon/tools/openrlhf-openrlhf","graph_url":"https://www.graphcanon.com/api/graphcanon/graph?tool=openrlhf-openrlhf"},"categories":[{"slug":"inference-serving","name":"Inference & Serving","url":"https://www.graphcanon.com/categories/inference-serving","markdown_url":"https://www.graphcanon.com/categories/inference-serving.md","api_url":"https://www.graphcanon.com/api/graphcanon/categories/inference-serving"},{"slug":"model-training","name":"Model Training","url":"https://www.graphcanon.com/categories/model-training","markdown_url":"https://www.graphcanon.com/categories/model-training.md","api_url":"https://www.graphcanon.com/api/graphcanon/categories/model-training"}],"tags":[{"slug":"large-language-models","name":"large language models"},{"slug":"proximal-policy-optimization","name":"proximal-policy-optimization"},{"slug":"raylib","name":"raylib"},{"slug":"reinforcement-learning","name":"reinforcement-learning"},{"slug":"transformers","name":"transformers"},{"slug":"vllm","name":"vllm"}],"edges":[{"type":"related","direction":"out","explanation":"Both platforms aim to work with LLMs and can be considered adjacent projects, but Onyx focuses more on a user-friendly AI chat interface.","successor_context":null,"tool":{"slug":"onyx-dot-app-onyx","name":"onyx","tagline":"Open Source AI Platform - AI Chat with advanced features that works with every LLM","github_url":"https://github.com/onyx-dot-app/onyx","owner":"onyx-dot-app","repo":"onyx","owner_avatar_url":"https://avatars.githubusercontent.com/u/131946000?v=4","primary_language":"Python","stars":31617,"forks":4351,"topics":["ai","ai-chat","chatgpt","chatui","enterprise-search","gen-ai","information-retrieval","llm","llm-ui","nextjs","python","rag","self-hosted","vector-search"],"archived":false,"github_pushed_at":"2026-08-16T10:06:57+00:00","maintenance_label":"Very active","stars_delta_30d":685,"url":"https://www.graphcanon.com/tools/onyx-dot-app-onyx","markdown_url":"https://www.graphcanon.com/tools/onyx-dot-app-onyx.md","api_url":"https://www.graphcanon.com/api/graphcanon/tools/onyx-dot-app-onyx","graph_url":"https://www.graphcanon.com/api/graphcanon/graph?tool=onyx-dot-app-onyx"}},{"type":"depends_on","direction":"out","explanation":"OpenRLHF is built on HuggingFace Transformers for seamless model loading and fine-tuning of pretrained models.","successor_context":null,"tool":{"slug":"huggingface-transformers","name":"transformers","tagline":"Transformers: the model-definition framework for state-of-the-art machine learning models in text, vision, audio, and multimodal models","github_url":"https://github.com/huggingface/transformers","owner":"huggingface","repo":"transformers","owner_avatar_url":"https://avatars.githubusercontent.com/u/25720743?v=4","primary_language":"Python","stars":164121,"forks":34249,"topics":["audio","deep-learning","deepseek","gemma","glm","hacktoberfest","llm","machine-learning","model-hub","natural-language-processing","nlp","pretrained-models","python","pytorch","pytorch-transformers","qwen","speech-recognition","transformer","vlm"],"archived":false,"github_pushed_at":"2026-08-15T22:28:12+00:00","maintenance_label":"Very active","stars_delta_30d":1457,"url":"https://www.graphcanon.com/tools/huggingface-transformers","markdown_url":"https://www.graphcanon.com/tools/huggingface-transformers.md","api_url":"https://www.graphcanon.com/api/graphcanon/tools/huggingface-transformers","graph_url":"https://www.graphcanon.com/api/graphcanon/graph?tool=huggingface-transformers"}},{"type":"related","direction":"out","explanation":"Both serve large language and multimodal models, but sglang is a serving framework while OpenRLHF focuses on RL optimization.","successor_context":null,"tool":{"slug":"sgl-project-sglang","name":"sglang","tagline":"High-performance serving framework for large language and multimodal models","github_url":"https://github.com/sgl-project/sglang","owner":"sgl-project","repo":"sglang","owner_avatar_url":"https://avatars.githubusercontent.com/u/147780389?v=4","primary_language":"Python","stars":31454,"forks":7720,"topics":["attention","blackwell","cuda","deepseek","diffusion","glm","gpt-oss","inference","llama","llm","minimax","moe","qwen","qwen-image","reinforcement-learning","transformer","vlm","wan"],"archived":false,"github_pushed_at":"2026-08-07T06:00:20+00:00","maintenance_label":"Very active","stars_delta_30d":1409,"url":"https://www.graphcanon.com/tools/sgl-project-sglang","markdown_url":"https://www.graphcanon.com/tools/sgl-project-sglang.md","api_url":"https://www.graphcanon.com/api/graphcanon/tools/sgl-project-sglang","graph_url":"https://www.graphcanon.com/api/graphcanon/graph?tool=sgl-project-sglang"}},{"type":"depends_on","direction":"out","explanation":"OpenRLHF uses vLLM to power high-throughput and memory-efficient sample generation, optimizing the RLHF training process.","successor_context":null,"tool":{"slug":"vllm-project-vllm","name":"vllm","tagline":"A high-throughput and memory-efficient inference and serving engine for LLMs","github_url":"https://github.com/vllm-project/vllm","owner":"vllm-project","repo":"vllm","owner_avatar_url":"https://avatars.githubusercontent.com/u/136984999?v=4","primary_language":"Python","stars":87847,"forks":20135,"topics":["amd","blackwell","cuda","deepseek","deepseek-v3","gpt","gpt-oss","inference","kimi","llama","llm","llm-serving","model-serving","moe","openai","pytorch","qwen","qwen3","tpu","transformer"],"archived":false,"github_pushed_at":"2026-08-01T11:55:36+00:00","maintenance_label":"Very active","url":"https://www.graphcanon.com/tools/vllm-project-vllm","markdown_url":"https://www.graphcanon.com/tools/vllm-project-vllm.md","api_url":"https://www.graphcanon.com/api/graphcanon/tools/vllm-project-vllm","graph_url":"https://www.graphcanon.com/api/graphcanon/graph?tool=vllm-project-vllm"}},{"type":"depends_on","direction":"out","explanation":"OpenRLHF leverages Ray for efficient distributed scheduling, separating Actor, Reward, Reference, and Critic models across different GPUs.","successor_context":null,"tool":{"slug":"ray-project-ray","name":"ray","tagline":"Ray is an AI compute engine with a core distributed runtime and AI Libraries for accelerating ML workloads.","github_url":"https://github.com/ray-project/ray","owner":"ray-project","repo":"ray","owner_avatar_url":"https://avatars.githubusercontent.com/u/22125274?v=4","primary_language":"Python","stars":43526,"forks":7929,"topics":["data-science","deep-learning","deployment","distributed","hyperparameter-optimization","hyperparameter-search","large-language-models","llm","llm-inference","llm-serving","machine-learning","optimization","parallel","python","pytorch","ray","reinforcement-learning","rllib","serving","tensorflow"],"archived":false,"github_pushed_at":"2026-08-16T00:26:16+00:00","maintenance_label":"Very active","stars_delta_30d":270,"url":"https://www.graphcanon.com/tools/ray-project-ray","markdown_url":"https://www.graphcanon.com/tools/ray-project-ray.md","api_url":"https://www.graphcanon.com/api/graphcanon/tools/ray-project-ray","graph_url":"https://www.graphcanon.com/api/graphcanon/graph?tool=ray-project-ray"}}],"neighbours":[{"slug":"openai-openai-agents-python","name":"openai-agents-python","tagline":"A lightweight, powerful framework for multi-agent workflows","github_url":"https://github.com/openai/openai-agents-python","owner":"openai","repo":"openai-agents-python","owner_avatar_url":"https://avatars.githubusercontent.com/u/14957082?v=4","primary_language":"Python","stars":28676,"forks":4516,"topics":["agents","ai","framework","harness","llm","openai","python"],"archived":false,"github_pushed_at":"2026-08-16T09:21:37+00:00","maintenance_label":"Very active","url":"https://www.graphcanon.com/tools/openai-openai-agents-python","markdown_url":"https://www.graphcanon.com/tools/openai-openai-agents-python.md","api_url":"https://www.graphcanon.com/api/graphcanon/tools/openai-openai-agents-python","graph_url":"https://www.graphcanon.com/api/graphcanon/graph?tool=openai-openai-agents-python","shared_categories":[]},{"slug":"huggingface-open-r1","name":"open-r1","tagline":"Fully open reproduction of DeepSeek-R1","github_url":"https://github.com/huggingface/open-r1","owner":"huggingface","repo":"open-r1","owner_avatar_url":"https://avatars.githubusercontent.com/u/25720743?v=4","primary_language":"Python","stars":26423,"forks":2447,"topics":[],"archived":false,"github_pushed_at":"2026-04-02T14:03:15+00:00","maintenance_label":"Slowing","url":"https://www.graphcanon.com/tools/huggingface-open-r1","markdown_url":"https://www.graphcanon.com/tools/huggingface-open-r1.md","api_url":"https://www.graphcanon.com/api/graphcanon/tools/huggingface-open-r1","graph_url":"https://www.graphcanon.com/api/graphcanon/graph?tool=huggingface-open-r1","shared_categories":["model-training","inference-serving"]},{"slug":"verl-project-verl","name":"verl","tagline":"A Flexible and Efficient RL Post-Training Framework","github_url":"https://github.com/verl-project/verl","owner":"verl-project","repo":"verl","owner_avatar_url":"https://avatars.githubusercontent.com/u/212961691?v=4","primary_language":"Python","stars":22854,"forks":4353,"topics":[],"archived":false,"github_pushed_at":"2026-08-07T04:17:39+00:00","maintenance_label":"Active","url":"https://www.graphcanon.com/tools/verl-project-verl","markdown_url":"https://www.graphcanon.com/tools/verl-project-verl.md","api_url":"https://www.graphcanon.com/api/graphcanon/tools/verl-project-verl","graph_url":"https://www.graphcanon.com/api/graphcanon/graph?tool=verl-project-verl","shared_categories":["model-training"]},{"slug":"huggingface-trl","name":"trl","tagline":"Train transformer language models with reinforcement learning.","github_url":"https://github.com/huggingface/trl","owner":"huggingface","repo":"trl","owner_avatar_url":"https://avatars.githubusercontent.com/u/25720743?v=4","primary_language":"Python","stars":19016,"forks":2891,"topics":[],"archived":false,"github_pushed_at":"2026-08-06T10:02:43+00:00","maintenance_label":"Active","url":"https://www.graphcanon.com/tools/huggingface-trl","markdown_url":"https://www.graphcanon.com/tools/huggingface-trl.md","api_url":"https://www.graphcanon.com/api/graphcanon/tools/huggingface-trl","graph_url":"https://www.graphcanon.com/api/graphcanon/graph?tool=huggingface-trl","shared_categories":["model-training"]},{"slug":"opendilab-awesome-rlhf","name":"awesome-RLHF","tagline":"A curated list of reinforcement learning with human feedback resources (continually updated)","github_url":"https://github.com/opendilab/awesome-RLHF","owner":"opendilab","repo":"awesome-RLHF","owner_avatar_url":"https://avatars.githubusercontent.com/u/86840398?v=4","primary_language":null,"stars":4422,"forks":258,"topics":["deep-learning","deep-reinforcement-learning","human-feedback","large-language-models","reinforcement-learning","rlhf"],"archived":false,"github_pushed_at":"2026-05-20T12:56:15+00:00","maintenance_label":"Slowing","url":"https://www.graphcanon.com/tools/opendilab-awesome-rlhf","markdown_url":"https://www.graphcanon.com/tools/opendilab-awesome-rlhf.md","api_url":"https://www.graphcanon.com/api/graphcanon/tools/opendilab-awesome-rlhf","graph_url":"https://www.graphcanon.com/api/graphcanon/graph?tool=opendilab-awesome-rlhf","shared_categories":["model-training"]},{"slug":"alibaba-roll","name":"ROLL","tagline":"Scaling Library for Reinforcement Learning with Large Language Models","github_url":"https://github.com/alibaba/ROLL","owner":"alibaba","repo":"ROLL","owner_avatar_url":"https://avatars.githubusercontent.com/u/1961952?v=4","primary_language":"Python","stars":3354,"forks":304,"topics":["agentic","rlhf","rlvr"],"archived":false,"github_pushed_at":"2026-08-07T01:51:58+00:00","maintenance_label":"Active","url":"https://www.graphcanon.com/tools/alibaba-roll","markdown_url":"https://www.graphcanon.com/tools/alibaba-roll.md","api_url":"https://www.graphcanon.com/api/graphcanon/tools/alibaba-roll","graph_url":"https://www.graphcanon.com/api/graphcanon/graph?tool=alibaba-roll","shared_categories":["model-training"]},{"slug":"stanford-crfm-helm","name":"helm","tagline":"Holistic, reproducible and transparent evaluation of foundation models","github_url":"https://github.com/stanford-crfm/helm","owner":"stanford-crfm","repo":"helm","owner_avatar_url":"https://avatars.githubusercontent.com/u/75054807?v=4","primary_language":"Python","stars":2873,"forks":406,"topics":[],"archived":false,"github_pushed_at":"2026-08-01T01:23:17+00:00","maintenance_label":"Active","url":"https://www.graphcanon.com/tools/stanford-crfm-helm","markdown_url":"https://www.graphcanon.com/tools/stanford-crfm-helm.md","api_url":"https://www.graphcanon.com/api/graphcanon/tools/stanford-crfm-helm","graph_url":"https://www.graphcanon.com/api/graphcanon/graph?tool=stanford-crfm-helm","shared_categories":[]},{"slug":"ray-project-ray-llm","name":"ray-llm","tagline":"Archived repository; LLM serving APIs integrated into the Ray project","github_url":"https://github.com/ray-project/ray-llm","owner":"ray-project","repo":"ray-llm","owner_avatar_url":"https://avatars.githubusercontent.com/u/22125274?v=4","primary_language":null,"stars":1261,"forks":90,"topics":["llm","llm-serving","ray"],"archived":true,"github_pushed_at":"2025-03-13T01:13:38+00:00","maintenance_label":"Archived","url":"https://www.graphcanon.com/tools/ray-project-ray-llm","markdown_url":"https://www.graphcanon.com/tools/ray-project-ray-llm.md","api_url":"https://www.graphcanon.com/api/graphcanon/tools/ray-project-ray-llm","graph_url":"https://www.graphcanon.com/api/graphcanon/graph?tool=ray-project-ray-llm","shared_categories":["model-training","inference-serving"]},{"slug":"rath-team-openrath","name":"OpenRath","tagline":"An open-source runtime for dynamic multi-agent workflows","github_url":"https://github.com/Rath-Team/OpenRath","owner":"Rath-Team","repo":"OpenRath","owner_avatar_url":"https://avatars.githubusercontent.com/u/279513473?v=4","primary_language":"Python","stars":1100,"forks":52,"topics":["agent-framework","agentic-ai","ai-agents","anthropic","lllm-agent","llm","memory","model-context-protocol","multi-agent","multi-agent-systems","open-source","openai","provenance","python","runtime-state","sandbox","session-graph","session-state","workflow-orchestration"],"archived":false,"github_pushed_at":"2026-07-22T18:29:53+00:00","maintenance_label":"Steady","url":"https://www.graphcanon.com/tools/rath-team-openrath","markdown_url":"https://www.graphcanon.com/tools/rath-team-openrath.md","api_url":"https://www.graphcanon.com/api/graphcanon/tools/rath-team-openrath","graph_url":"https://www.graphcanon.com/api/graphcanon/graph?tool=rath-team-openrath","shared_categories":[]},{"slug":"joyce94-llm-rlhf-tuning","name":"LLM-RLHF-Tuning","tagline":"LLM Tuning with PEFT (SFT+RM+PPO+DPO with LoRA)","github_url":"https://github.com/Joyce94/LLM-RLHF-Tuning","owner":"Joyce94","repo":"LLM-RLHF-Tuning","owner_avatar_url":"https://avatars.githubusercontent.com/u/28557140?v=4","primary_language":"Python","stars":453,"forks":24,"topics":["fine-tuning","language-model","llama","llm","lora","peft","ppo","reinforcement-learning","rlhf"],"archived":false,"github_pushed_at":"2023-10-11T08:41:20+00:00","maintenance_label":"Dormant","url":"https://www.graphcanon.com/tools/joyce94-llm-rlhf-tuning","markdown_url":"https://www.graphcanon.com/tools/joyce94-llm-rlhf-tuning.md","api_url":"https://www.graphcanon.com/api/graphcanon/tools/joyce94-llm-rlhf-tuning","graph_url":"https://www.graphcanon.com/api/graphcanon/graph?tool=joyce94-llm-rlhf-tuning","shared_categories":["model-training"]}]}}