{"data":{"node":{"slug":"pku-alignment-align-anything","name":"align-anything","tagline":"Training All-modality Model with Feedback","github_url":"https://github.com/PKU-Alignment/align-anything","owner":"PKU-Alignment","repo":"align-anything","owner_avatar_url":"https://avatars.githubusercontent.com/u/129283536?v=4","primary_language":"Python","stars":4666,"forks":505,"topics":["chameleon","dpo","large-language-models","multimodal","rlhf","vision-language-model"],"archived":false,"github_pushed_at":"2025-11-27T06:50:36+00:00","maintenance_label":"Slowing","stars_delta_30d":4,"url":"https://www.graphcanon.com/tools/pku-alignment-align-anything","markdown_url":"https://www.graphcanon.com/tools/pku-alignment-align-anything.md","api_url":"https://www.graphcanon.com/api/graphcanon/tools/pku-alignment-align-anything","graph_url":"https://www.graphcanon.com/api/graphcanon/graph?tool=pku-alignment-align-anything"},"categories":[{"slug":"llm-frameworks","name":"LLM Frameworks","url":"https://www.graphcanon.com/categories/llm-frameworks","markdown_url":"https://www.graphcanon.com/categories/llm-frameworks.md","api_url":"https://www.graphcanon.com/api/graphcanon/categories/llm-frameworks"},{"slug":"model-training","name":"Model Training","url":"https://www.graphcanon.com/categories/model-training","markdown_url":"https://www.graphcanon.com/categories/model-training.md","api_url":"https://www.graphcanon.com/api/graphcanon/categories/model-training"}],"tags":[{"slug":"chameleon","name":"chameleon"},{"slug":"dpo","name":"dpo"},{"slug":"large-language-models","name":"large language models"},{"slug":"multimodal","name":"multimodal"},{"slug":"rlhf","name":"rlhf"},{"slug":"vision-language-model","name":"vision-language-model"}],"edges":[{"type":"related","direction":"out","explanation":"Both repositories focus on multimodal large language models, but 'Align Anything' focuses specifically on training techniques with feedback, while the former is a collection of resources and advances in the field.","successor_context":null,"tool":{"slug":"bradyfu-awesome-multimodal-large-language-models","name":"Awesome-Multimodal-Large-Language-Models","tagline":"Latest Advances on Multimodal Large Language Models","github_url":"https://github.com/BradyFU/Awesome-Multimodal-Large-Language-Models","owner":"BradyFU","repo":"Awesome-Multimodal-Large-Language-Models","owner_avatar_url":"https://avatars.githubusercontent.com/u/54254631?v=4","primary_language":null,"stars":17978,"forks":1133,"topics":["chain-of-thought","in-context-learning","instruction-following","instruction-tuning","large-language-models","large-vision-language-model","large-vision-language-models","multi-modality","multimodal-chain-of-thought","multimodal-in-context-learning","multimodal-instruction-tuning","multimodal-large-language-models","visual-instruction-tuning"],"archived":false,"github_pushed_at":"2026-08-14T17:17:50+00:00","maintenance_label":"Very active","stars_delta_30d":29,"url":"https://www.graphcanon.com/tools/bradyfu-awesome-multimodal-large-language-models","markdown_url":"https://www.graphcanon.com/tools/bradyfu-awesome-multimodal-large-language-models.md","api_url":"https://www.graphcanon.com/api/graphcanon/tools/bradyfu-awesome-multimodal-large-language-models","graph_url":"https://www.graphcanon.com/api/graphcanon/graph?tool=bradyfu-awesome-multimodal-large-language-models"}},{"type":"related","direction":"out","explanation":"Both 'Align Anything' and MGM deal with multimodal vision-language models, but they approach the problem differently with 'Align Anything' focusing on training methodologies and MGM on a framework for both understanding and generating images.","successor_context":null,"tool":{"slug":"jia-lab-research-mgm","name":"MGM","tagline":"Mini-Gemini: Mining the Potential of Multi-modality Vision Language Models","github_url":"https://github.com/JIA-Lab-research/MGM","owner":"JIA-Lab-research","repo":"MGM","owner_avatar_url":"https://avatars.githubusercontent.com/u/64006090?v=4","primary_language":"Python","stars":3331,"forks":276,"topics":["generation","large-language-models","vision-language-model"],"archived":false,"github_pushed_at":"2024-05-04T14:36:51+00:00","maintenance_label":"Dormant","stars_delta_30d":1,"url":"https://www.graphcanon.com/tools/jia-lab-research-mgm","markdown_url":"https://www.graphcanon.com/tools/jia-lab-research-mgm.md","api_url":"https://www.graphcanon.com/api/graphcanon/tools/jia-lab-research-mgm","graph_url":"https://www.graphcanon.com/api/graphcanon/graph?tool=jia-lab-research-mgm"}},{"type":"integrates_with","direction":"out","explanation":"Align-anything provides alignment methods such as Supervised Fine-Tuning (SFT), Direct Preference Optimization (DPO), Proximal Policy Optimization (PPO), and rule-based reinforcement learning to align multimodal large models with human intentions. LlamaFactory offers an efficient fine-tuning process for over 100 language models and vision-language models supported by major tech companies. The 'int","successor_context":null,"tool":{"slug":"hiyouga-llamafactory","name":"LlamaFactory","tagline":"Unified Efficient Fine-Tuning of 100+ LLMs & VLMs","github_url":"https://github.com/hiyouga/LlamaFactory","owner":"hiyouga","repo":"LlamaFactory","owner_avatar_url":"https://avatars.githubusercontent.com/u/16256802?v=4","primary_language":"Python","stars":74132,"forks":9071,"topics":["agent","ai","deepseek","fine-tuning","gemma","gpt","instruction-tuning","large-language-models","llama","llama3","llm","lora","moe","nlp","peft","qlora","quantization","qwen","rlhf","transformers"],"archived":false,"github_pushed_at":"2026-08-13T12:45:56+00:00","maintenance_label":"Very active","stars_delta_30d":803,"url":"https://www.graphcanon.com/tools/hiyouga-llamafactory","markdown_url":"https://www.graphcanon.com/tools/hiyouga-llamafactory.md","api_url":"https://www.graphcanon.com/api/graphcanon/tools/hiyouga-llamafactory","graph_url":"https://www.graphcanon.com/api/graphcanon/graph?tool=hiyouga-llamafactory"}}],"neighbours":[{"slug":"bradyfu-awesome-multimodal-large-language-models","name":"Awesome-Multimodal-Large-Language-Models","tagline":"Latest Advances on Multimodal Large Language Models","github_url":"https://github.com/BradyFU/Awesome-Multimodal-Large-Language-Models","owner":"BradyFU","repo":"Awesome-Multimodal-Large-Language-Models","owner_avatar_url":"https://avatars.githubusercontent.com/u/54254631?v=4","primary_language":null,"stars":17978,"forks":1133,"topics":["chain-of-thought","in-context-learning","instruction-following","instruction-tuning","large-language-models","large-vision-language-model","large-vision-language-models","multi-modality","multimodal-chain-of-thought","multimodal-in-context-learning","multimodal-instruction-tuning","multimodal-large-language-models","visual-instruction-tuning"],"archived":false,"github_pushed_at":"2026-08-14T17:17:50+00:00","maintenance_label":"Active","url":"https://www.graphcanon.com/tools/bradyfu-awesome-multimodal-large-language-models","markdown_url":"https://www.graphcanon.com/tools/bradyfu-awesome-multimodal-large-language-models.md","api_url":"https://www.graphcanon.com/api/graphcanon/tools/bradyfu-awesome-multimodal-large-language-models","graph_url":"https://www.graphcanon.com/api/graphcanon/graph?tool=bradyfu-awesome-multimodal-large-language-models","shared_categories":["llm-frameworks"]},{"slug":"lightning-ai-litgpt","name":"litgpt","tagline":"High-performance LLMs with recipes for pretraining, finetuning and deployment","github_url":"https://github.com/Lightning-AI/litgpt","owner":"Lightning-AI","repo":"litgpt","owner_avatar_url":"https://avatars.githubusercontent.com/u/58386951?v=4","primary_language":"Python","stars":13605,"forks":1483,"topics":["ai","artificial-intelligence","deep-learning","large-language-models","llm","llm-inference","llms"],"archived":false,"github_pushed_at":"2026-07-20T10:24:12+00:00","maintenance_label":"Steady","url":"https://www.graphcanon.com/tools/lightning-ai-litgpt","markdown_url":"https://www.graphcanon.com/tools/lightning-ai-litgpt.md","api_url":"https://www.graphcanon.com/api/graphcanon/tools/lightning-ai-litgpt","graph_url":"https://www.graphcanon.com/api/graphcanon/graph?tool=lightning-ai-litgpt","shared_categories":["model-training","llm-frameworks"]},{"slug":"shishirpatil-gorilla","name":"gorilla","tagline":"Training and Evaluating LLMs for Function Calls (Tool Calls)","github_url":"https://github.com/ShishirPatil/gorilla","owner":"ShishirPatil","repo":"gorilla","owner_avatar_url":"https://avatars.githubusercontent.com/u/30296397?v=4","primary_language":"Python","stars":12988,"forks":1397,"topics":["api","api-documentation","chatgpt","claude-api","gpt-4-api","llm","openai-api","openai-functions"],"archived":false,"github_pushed_at":"2026-04-13T03:19:45+00:00","maintenance_label":"Slowing","url":"https://www.graphcanon.com/tools/shishirpatil-gorilla","markdown_url":"https://www.graphcanon.com/tools/shishirpatil-gorilla.md","api_url":"https://www.graphcanon.com/api/graphcanon/tools/shishirpatil-gorilla","graph_url":"https://www.graphcanon.com/api/graphcanon/graph?tool=shishirpatil-gorilla","shared_categories":["model-training"]},{"slug":"llm-attacks-llm-attacks","name":"llm-attacks","tagline":"Universal and Transferable Attacks on Aligned Language Models","github_url":"https://github.com/llm-attacks/llm-attacks","owner":"llm-attacks","repo":"llm-attacks","owner_avatar_url":"https://avatars.githubusercontent.com/u/140664770?v=4","primary_language":"Python","stars":4756,"forks":633,"topics":[],"archived":false,"github_pushed_at":"2024-08-02T06:02:18+00:00","maintenance_label":"Dormant","url":"https://www.graphcanon.com/tools/llm-attacks-llm-attacks","markdown_url":"https://www.graphcanon.com/tools/llm-attacks-llm-attacks.md","api_url":"https://www.graphcanon.com/api/graphcanon/tools/llm-attacks-llm-attacks","graph_url":"https://www.graphcanon.com/api/graphcanon/graph?tool=llm-attacks-llm-attacks","shared_categories":["llm-frameworks"]},{"slug":"evolvinglmms-lab-lmms-eval","name":"lmms-eval","tagline":"One-for-All Multimodal Evaluation Toolkit Across Text, Image, Video, and Audio Tasks","github_url":"https://github.com/EvolvingLMMs-Lab/lmms-eval","owner":"EvolvingLMMs-Lab","repo":"lmms-eval","owner_avatar_url":"https://avatars.githubusercontent.com/u/154951679?v=4","primary_language":"Python","stars":4368,"forks":639,"topics":["agi","audio-evaluation","benchmark","evaluation","large-language-models","llm-evaluation","multimodal","multimodal-evaluation","video-understanding","vision-language-model","vlm"],"archived":false,"github_pushed_at":"2026-08-06T02:22:23+00:00","maintenance_label":"Active","url":"https://www.graphcanon.com/tools/evolvinglmms-lab-lmms-eval","markdown_url":"https://www.graphcanon.com/tools/evolvinglmms-lab-lmms-eval.md","api_url":"https://www.graphcanon.com/api/graphcanon/tools/evolvinglmms-lab-lmms-eval","graph_url":"https://www.graphcanon.com/api/graphcanon/graph?tool=evolvinglmms-lab-lmms-eval","shared_categories":[]},{"slug":"jia-lab-research-mgm","name":"MGM","tagline":"Mini-Gemini: Mining the Potential of Multi-modality Vision Language Models","github_url":"https://github.com/JIA-Lab-research/MGM","owner":"JIA-Lab-research","repo":"MGM","owner_avatar_url":"https://avatars.githubusercontent.com/u/64006090?v=4","primary_language":"Python","stars":3331,"forks":276,"topics":["generation","large-language-models","vision-language-model"],"archived":false,"github_pushed_at":"2024-05-04T14:36:51+00:00","maintenance_label":"Dormant","url":"https://www.graphcanon.com/tools/jia-lab-research-mgm","markdown_url":"https://www.graphcanon.com/tools/jia-lab-research-mgm.md","api_url":"https://www.graphcanon.com/api/graphcanon/tools/jia-lab-research-mgm","graph_url":"https://www.graphcanon.com/api/graphcanon/graph?tool=jia-lab-research-mgm","shared_categories":["model-training","llm-frameworks"]},{"slug":"eladlev-autoprompt","name":"AutoPrompt","tagline":"Framework for prompt tuning using Intent-based Prompt Calibration","github_url":"https://github.com/Eladlev/AutoPrompt","owner":"Eladlev","repo":"AutoPrompt","owner_avatar_url":"https://avatars.githubusercontent.com/u/28984104?v=4","primary_language":"Python","stars":2993,"forks":264,"topics":["prompt-engineering","prompt-tuning","synthetic-dataset-generation"],"archived":false,"github_pushed_at":"2025-12-02T17:23:20+00:00","maintenance_label":"Slowing","url":"https://www.graphcanon.com/tools/eladlev-autoprompt","markdown_url":"https://www.graphcanon.com/tools/eladlev-autoprompt.md","api_url":"https://www.graphcanon.com/api/graphcanon/tools/eladlev-autoprompt","graph_url":"https://www.graphcanon.com/api/graphcanon/graph?tool=eladlev-autoprompt","shared_categories":["llm-frameworks"]},{"slug":"roboflow-maestro","name":"maestro","tagline":"Streamlines fine-tuning for multimodal models PaliGemma 2, Florence-2, Qwen2.5-VL","github_url":"https://github.com/roboflow/maestro","owner":"roboflow","repo":"maestro","owner_avatar_url":"https://avatars.githubusercontent.com/u/53104118?v=4","primary_language":"Python","stars":2687,"forks":222,"topics":["captioning","fine-tuning","florence-2","multimodal","objectdetection","paligemma","phi-3-vision","qwen2-vl","transformers","vision-and-language","vqa"],"archived":false,"github_pushed_at":"2026-07-20T17:45:05+00:00","maintenance_label":"Steady","url":"https://www.graphcanon.com/tools/roboflow-maestro","markdown_url":"https://www.graphcanon.com/tools/roboflow-maestro.md","api_url":"https://www.graphcanon.com/api/graphcanon/tools/roboflow-maestro","graph_url":"https://www.graphcanon.com/api/graphcanon/graph?tool=roboflow-maestro","shared_categories":["model-training"]},{"slug":"om-ai-lab-omagent","name":"OmAgent","tagline":"Build multimodal language agents for fast prototype and production","github_url":"https://github.com/om-ai-lab/OmAgent","owner":"om-ai-lab","repo":"OmAgent","owner_avatar_url":"https://avatars.githubusercontent.com/u/96569904?v=4","primary_language":"Python","stars":2665,"forks":292,"topics":["agent","chatbot","gemini","gpt","gpt4","gradio","language-agent","large-language-models","llama","llava","llm","multimodal","multimodal-agent","openai","python","rag","smart-hardware","vision-and-language","vlm","workflow"],"archived":false,"github_pushed_at":"2025-03-19T11:36:13+00:00","maintenance_label":"Dormant","url":"https://www.graphcanon.com/tools/om-ai-lab-omagent","markdown_url":"https://www.graphcanon.com/tools/om-ai-lab-omagent.md","api_url":"https://www.graphcanon.com/api/graphcanon/tools/om-ai-lab-omagent","graph_url":"https://www.graphcanon.com/api/graphcanon/graph?tool=om-ai-lab-omagent","shared_categories":["llm-frameworks"]},{"slug":"tianrun-chen-sam-adapter-pytorch","name":"SAM-Adapter-PyTorch","tagline":"Adapting Meta AI's Segment Anything to Downstream Tasks with Adapters and Prompts","github_url":"https://github.com/tianrun-chen/SAM-Adapter-PyTorch","owner":"tianrun-chen","repo":"SAM-Adapter-PyTorch","owner_avatar_url":"https://avatars.githubusercontent.com/u/126600557?v=4","primary_language":"Python","stars":1544,"forks":124,"topics":["2d-segmentation","adapter","camouflage-images","camouflaged-object-detection","camouflaged-target-detection","fine-tune","fine-tuning","image-segmentation","image-segmentation-pytorch","segment-anything","segment-anything-model"],"archived":false,"github_pushed_at":"2026-05-17T04:56:00+00:00","maintenance_label":"Slowing","url":"https://www.graphcanon.com/tools/tianrun-chen-sam-adapter-pytorch","markdown_url":"https://www.graphcanon.com/tools/tianrun-chen-sam-adapter-pytorch.md","api_url":"https://www.graphcanon.com/api/graphcanon/tools/tianrun-chen-sam-adapter-pytorch","graph_url":"https://www.graphcanon.com/api/graphcanon/graph?tool=tianrun-chen-sam-adapter-pytorch","shared_categories":["model-training"]},{"slug":"arahim3-mlx-tune","name":"mlx-tune","tagline":"Fine-tune LLMs on your Mac with Apple Silicon for various tasks including SFT, DPO, GRPO, Vision, TTS, STT, Embedding, and OCR.","github_url":"https://github.com/ARahim3/mlx-tune","owner":"ARahim3","repo":"mlx-tune","owner_avatar_url":"https://avatars.githubusercontent.com/u/41390319?v=4","primary_language":"Python","stars":1372,"forks":88,"topics":["apple-silicon","deep-learning","huggingface","large-language-models","llm","llm-finetuning","local-llm","lora","machine-learning","macos","mlx","on-device-ai","peft","speech-recognition","speech-to-text","text-to-speech","transformers","unsloth","vision-language-model","whisper"],"archived":false,"github_pushed_at":"2026-06-23T12:24:30+00:00","maintenance_label":"Steady","url":"https://www.graphcanon.com/tools/arahim3-mlx-tune","markdown_url":"https://www.graphcanon.com/tools/arahim3-mlx-tune.md","api_url":"https://www.graphcanon.com/api/graphcanon/tools/arahim3-mlx-tune","graph_url":"https://www.graphcanon.com/api/graphcanon/graph?tool=arahim3-mlx-tune","shared_categories":["model-training","llm-frameworks"]},{"slug":"sakanaai-text-to-lora","name":"text-to-lora","tagline":"Hypernetworks for adapting LLMs to specific tasks via textual descriptions","github_url":"https://github.com/SakanaAI/text-to-lora","owner":"SakanaAI","repo":"text-to-lora","owner_avatar_url":"https://avatars.githubusercontent.com/u/140988036?v=4","primary_language":"Python","stars":1294,"forks":88,"topics":["fine-tuning","hypernetworks","llm","lora","machine-learning"],"archived":false,"github_pushed_at":"2025-06-08T14:42:10+00:00","maintenance_label":"Dormant","url":"https://www.graphcanon.com/tools/sakanaai-text-to-lora","markdown_url":"https://www.graphcanon.com/tools/sakanaai-text-to-lora.md","api_url":"https://www.graphcanon.com/api/graphcanon/tools/sakanaai-text-to-lora","graph_url":"https://www.graphcanon.com/api/graphcanon/graph?tool=sakanaai-text-to-lora","shared_categories":["model-training"]}]}}