{"data":{"node":{"slug":"ictnlp-llama-omni","name":"LLaMA-Omni","tagline":"End-to-end speech interaction model based on Llama-3.1-8B-Instruct","github_url":"https://github.com/ictnlp/LLaMA-Omni","owner":"ictnlp","repo":"LLaMA-Omni","owner_avatar_url":"https://avatars.githubusercontent.com/u/45630465?v=4","primary_language":"Python","stars":3146,"forks":224,"topics":["large-language-models","multimodal-large-language-models","speech-interaction","speech-language-model","speech-to-speech","speech-to-text"],"archived":false,"github_pushed_at":"2025-05-19T02:24:42+00:00","maintenance_label":"Dormant","url":"https://www.graphcanon.com/tools/ictnlp-llama-omni","markdown_url":"https://www.graphcanon.com/tools/ictnlp-llama-omni.md","api_url":"https://www.graphcanon.com/api/graphcanon/tools/ictnlp-llama-omni","graph_url":"https://www.graphcanon.com/api/graphcanon/graph?tool=ictnlp-llama-omni"},"categories":[{"slug":"speech-audio","name":"Speech & Audio","url":"https://www.graphcanon.com/categories/speech-audio","markdown_url":"https://www.graphcanon.com/categories/speech-audio.md","api_url":"https://www.graphcanon.com/api/graphcanon/categories/speech-audio"}],"tags":[{"slug":"large-language-models","name":"large language models"},{"slug":"multimodal-large-language-models","name":"multimodal-large-language-models"},{"slug":"speech-interaction","name":"speech-interaction"},{"slug":"speech-language-model","name":"speech-language-model"},{"slug":"speech-to-speech","name":"speech-to-speech"},{"slug":"speech-to-text","name":"speech-to-text"}],"edges":[{"type":"related","direction":"out","explanation":"Both LLaMA-Omni and Whisper deal with speech processing, but they serve different purposes. Whisper focuses on robust speech recognition, while LLaMA-Omni is about generating both text and speech responses from speech inputs.","successor_context":null,"tool":{"slug":"openai-whisper","name":"whisper","tagline":"Robust Speech Recognition via Large-Scale Weak Supervision","github_url":"https://github.com/openai/whisper","owner":"openai","repo":"whisper","owner_avatar_url":"https://avatars.githubusercontent.com/u/14957082?v=4","primary_language":"Python","stars":106740,"forks":12971,"topics":[],"archived":false,"github_pushed_at":"2026-07-28T20:18:29+00:00","maintenance_label":"Active","url":"https://www.graphcanon.com/tools/openai-whisper","markdown_url":"https://www.graphcanon.com/tools/openai-whisper.md","api_url":"https://www.graphcanon.com/api/graphcanon/tools/openai-whisper","graph_url":"https://www.graphcanon.com/api/graphcanon/graph?tool=openai-whisper"}},{"type":"related","direction":"out","explanation":"Although LLaMA-Omni focuses on speech interaction with large language models, the concept of finding free resources (in this case, music) and accessibility might be tangentially related to its overall utility in a resource-constrained environment.","successor_context":null,"tool":{"slug":"nukeop-nuclear","name":"nuclear","tagline":"Streaming music player that finds free music for you","github_url":"https://github.com/nukeop/nuclear","owner":"nukeop","repo":"nuclear","owner_avatar_url":"https://avatars.githubusercontent.com/u/12746779?v=4","primary_language":"TypeScript","stars":18297,"forks":1321,"topics":["agent","ai","desktop-app","linux","mac","mcp","mcp-server","music","music-player","react","rust","spotify","streaming","tauri","typescript","windows"],"archived":false,"github_pushed_at":"2026-08-16T00:50:31+00:00","maintenance_label":"Very active","stars_delta_30d":235,"url":"https://www.graphcanon.com/tools/nukeop-nuclear","markdown_url":"https://www.graphcanon.com/tools/nukeop-nuclear.md","api_url":"https://www.graphcanon.com/api/graphcanon/tools/nukeop-nuclear","graph_url":"https://www.graphcanon.com/api/graphcanon/graph?tool=nukeop-nuclear"}},{"type":"related","direction":"out","explanation":"LLaMA-Omni can be listed under the curated generative AI projects in awesome-generative-ai, given its focus on generating speech and text responses.","successor_context":null,"tool":{"slug":"steven2358-awesome-generative-ai","name":"awesome-generative-ai","tagline":"A curated list of modern Generative Artificial Intelligence projects and services","github_url":"https://github.com/steven2358/awesome-generative-ai","owner":"steven2358","repo":"awesome-generative-ai","owner_avatar_url":"https://avatars.githubusercontent.com/u/164072?v=4","primary_language":null,"stars":12501,"forks":1990,"topics":["ai","artificial-intelligence","awesome","awesome-list","generative-ai","generative-art","large-language-models","llm"],"archived":false,"github_pushed_at":"2026-08-03T10:58:05+00:00","maintenance_label":"Active","stars_delta_30d":160,"url":"https://www.graphcanon.com/tools/steven2358-awesome-generative-ai","markdown_url":"https://www.graphcanon.com/tools/steven2358-awesome-generative-ai.md","api_url":"https://www.graphcanon.com/api/graphcanon/tools/steven2358-awesome-generative-ai","graph_url":"https://www.graphcanon.com/api/graphcanon/graph?tool=steven2358-awesome-generative-ai"}},{"type":"alternative","direction":"out","explanation":"Both LLaMA-Omni and ChatTTS are generative speech models used for daily dialogue, but they have different model bases and focus areas.","successor_context":null,"tool":{"slug":"2noise-chattts","name":"ChatTTS","tagline":"A generative speech model for daily dialogue","github_url":"https://github.com/2noise/ChatTTS","owner":"2noise","repo":"ChatTTS","owner_avatar_url":"https://avatars.githubusercontent.com/u/164844019?v=4","primary_language":"Python","stars":39768,"forks":4257,"topics":["agent","chat","chatgpt","chattts","chinese","chinese-language","english","english-language","gpt","llm","llm-agent","natural-language-inference","python","text-to-speech","torch","torchaudio","tts"],"archived":false,"github_pushed_at":"2026-04-10T16:33:48+00:00","maintenance_label":"Slowing","stars_delta_30d":140,"url":"https://www.graphcanon.com/tools/2noise-chattts","markdown_url":"https://www.graphcanon.com/tools/2noise-chattts.md","api_url":"https://www.graphcanon.com/api/graphcanon/tools/2noise-chattts","graph_url":"https://www.graphcanon.com/api/graphcanon/graph?tool=2noise-chattts"}},{"type":"related","direction":"out","explanation":null,"successor_context":null,"tool":{"slug":"2noise-chattts","name":"ChatTTS","tagline":"A generative speech model for daily dialogue","github_url":"https://github.com/2noise/ChatTTS","owner":"2noise","repo":"ChatTTS","owner_avatar_url":"https://avatars.githubusercontent.com/u/164844019?v=4","primary_language":"Python","stars":39768,"forks":4257,"topics":["agent","chat","chatgpt","chattts","chinese","chinese-language","english","english-language","gpt","llm","llm-agent","natural-language-inference","python","text-to-speech","torch","torchaudio","tts"],"archived":false,"github_pushed_at":"2026-04-10T16:33:48+00:00","maintenance_label":"Slowing","stars_delta_30d":140,"url":"https://www.graphcanon.com/tools/2noise-chattts","markdown_url":"https://www.graphcanon.com/tools/2noise-chattts.md","api_url":"https://www.graphcanon.com/api/graphcanon/tools/2noise-chattts","graph_url":"https://www.graphcanon.com/api/graphcanon/graph?tool=2noise-chattts"}},{"type":"related","direction":"in","explanation":null,"successor_context":null,"tool":{"slug":"openai-whisper","name":"whisper","tagline":"Robust Speech Recognition via Large-Scale Weak Supervision","github_url":"https://github.com/openai/whisper","owner":"openai","repo":"whisper","owner_avatar_url":"https://avatars.githubusercontent.com/u/14957082?v=4","primary_language":"Python","stars":106740,"forks":12971,"topics":[],"archived":false,"github_pushed_at":"2026-07-28T20:18:29+00:00","maintenance_label":"Active","url":"https://www.graphcanon.com/tools/openai-whisper","markdown_url":"https://www.graphcanon.com/tools/openai-whisper.md","api_url":"https://www.graphcanon.com/api/graphcanon/tools/openai-whisper","graph_url":"https://www.graphcanon.com/api/graphcanon/graph?tool=openai-whisper"}},{"type":"alternative","direction":"in","explanation":"LLaMA-Omni and ChatTTS both serve speech interaction tasks but with different underlying models and setups.","successor_context":null,"tool":{"slug":"2noise-chattts","name":"ChatTTS","tagline":"A generative speech model for daily dialogue","github_url":"https://github.com/2noise/ChatTTS","owner":"2noise","repo":"ChatTTS","owner_avatar_url":"https://avatars.githubusercontent.com/u/164844019?v=4","primary_language":"Python","stars":39768,"forks":4257,"topics":["agent","chat","chatgpt","chattts","chinese","chinese-language","english","english-language","gpt","llm","llm-agent","natural-language-inference","python","text-to-speech","torch","torchaudio","tts"],"archived":false,"github_pushed_at":"2026-04-10T16:33:48+00:00","maintenance_label":"Slowing","stars_delta_30d":140,"url":"https://www.graphcanon.com/tools/2noise-chattts","markdown_url":"https://www.graphcanon.com/tools/2noise-chattts.md","api_url":"https://www.graphcanon.com/api/graphcanon/tools/2noise-chattts","graph_url":"https://www.graphcanon.com/api/graphcanon/graph?tool=2noise-chattts"}},{"type":"related","direction":"in","explanation":"Both Nuclear and LLaMA-Omni relate to AI-driven audio applications. While Nuclear is about playing music, LLaMA-Omni involves interacting with large language models for speech processing.","successor_context":null,"tool":{"slug":"nukeop-nuclear","name":"nuclear","tagline":"Streaming music player that finds free music for you","github_url":"https://github.com/nukeop/nuclear","owner":"nukeop","repo":"nuclear","owner_avatar_url":"https://avatars.githubusercontent.com/u/12746779?v=4","primary_language":"TypeScript","stars":18297,"forks":1321,"topics":["agent","ai","desktop-app","linux","mac","mcp","mcp-server","music","music-player","react","rust","spotify","streaming","tauri","typescript","windows"],"archived":false,"github_pushed_at":"2026-08-16T00:50:31+00:00","maintenance_label":"Very active","stars_delta_30d":235,"url":"https://www.graphcanon.com/tools/nukeop-nuclear","markdown_url":"https://www.graphcanon.com/tools/nukeop-nuclear.md","api_url":"https://www.graphcanon.com/api/graphcanon/tools/nukeop-nuclear","graph_url":"https://www.graphcanon.com/api/graphcanon/graph?tool=nukeop-nuclear"}},{"type":"alternative","direction":"in","explanation":"Both Whisper and LLaMA-Omni focus on seamless interaction with models for speech, but LLaMA-Omni integrates speech interaction capabilities into large language models while Whisper specifically handles robust speech recognition.","successor_context":null,"tool":{"slug":"openai-whisper","name":"whisper","tagline":"Robust Speech Recognition via Large-Scale Weak Supervision","github_url":"https://github.com/openai/whisper","owner":"openai","repo":"whisper","owner_avatar_url":"https://avatars.githubusercontent.com/u/14957082?v=4","primary_language":"Python","stars":106740,"forks":12971,"topics":[],"archived":false,"github_pushed_at":"2026-07-28T20:18:29+00:00","maintenance_label":"Active","url":"https://www.graphcanon.com/tools/openai-whisper","markdown_url":"https://www.graphcanon.com/tools/openai-whisper.md","api_url":"https://www.graphcanon.com/api/graphcanon/tools/openai-whisper","graph_url":"https://www.graphcanon.com/api/graphcanon/graph?tool=openai-whisper"}}],"neighbours":[{"slug":"vllm-project-vllm","name":"vllm","tagline":"A high-throughput and memory-efficient inference and serving engine for LLMs","github_url":"https://github.com/vllm-project/vllm","owner":"vllm-project","repo":"vllm","owner_avatar_url":"https://avatars.githubusercontent.com/u/136984999?v=4","primary_language":"Python","stars":87847,"forks":20135,"topics":["amd","blackwell","cuda","deepseek","deepseek-v3","gpt","gpt-oss","inference","kimi","llama","llm","llm-serving","model-serving","moe","openai","pytorch","qwen","qwen3","tpu","transformer"],"archived":false,"github_pushed_at":"2026-08-01T11:55:36+00:00","maintenance_label":"Active","url":"https://www.graphcanon.com/tools/vllm-project-vllm","markdown_url":"https://www.graphcanon.com/tools/vllm-project-vllm.md","api_url":"https://www.graphcanon.com/api/graphcanon/tools/vllm-project-vllm","graph_url":"https://www.graphcanon.com/api/graphcanon/graph?tool=vllm-project-vllm","shared_categories":[]},{"slug":"bradyfu-awesome-multimodal-large-language-models","name":"Awesome-Multimodal-Large-Language-Models","tagline":"Latest Advances on Multimodal Large Language Models","github_url":"https://github.com/BradyFU/Awesome-Multimodal-Large-Language-Models","owner":"BradyFU","repo":"Awesome-Multimodal-Large-Language-Models","owner_avatar_url":"https://avatars.githubusercontent.com/u/54254631?v=4","primary_language":null,"stars":17978,"forks":1133,"topics":["chain-of-thought","in-context-learning","instruction-following","instruction-tuning","large-language-models","large-vision-language-model","large-vision-language-models","multi-modality","multimodal-chain-of-thought","multimodal-in-context-learning","multimodal-instruction-tuning","multimodal-large-language-models","visual-instruction-tuning"],"archived":false,"github_pushed_at":"2026-08-14T17:17:50+00:00","maintenance_label":"Active","url":"https://www.graphcanon.com/tools/bradyfu-awesome-multimodal-large-language-models","markdown_url":"https://www.graphcanon.com/tools/bradyfu-awesome-multimodal-large-language-models.md","api_url":"https://www.graphcanon.com/api/graphcanon/tools/bradyfu-awesome-multimodal-large-language-models","graph_url":"https://www.graphcanon.com/api/graphcanon/graph?tool=bradyfu-awesome-multimodal-large-language-models","shared_categories":[]},{"slug":"lightning-ai-litgpt","name":"litgpt","tagline":"High-performance LLMs with recipes for pretraining, finetuning and deployment","github_url":"https://github.com/Lightning-AI/litgpt","owner":"Lightning-AI","repo":"litgpt","owner_avatar_url":"https://avatars.githubusercontent.com/u/58386951?v=4","primary_language":"Python","stars":13605,"forks":1483,"topics":["ai","artificial-intelligence","deep-learning","large-language-models","llm","llm-inference","llms"],"archived":false,"github_pushed_at":"2026-07-20T10:24:12+00:00","maintenance_label":"Steady","url":"https://www.graphcanon.com/tools/lightning-ai-litgpt","markdown_url":"https://www.graphcanon.com/tools/lightning-ai-litgpt.md","api_url":"https://www.graphcanon.com/api/graphcanon/tools/lightning-ai-litgpt","graph_url":"https://www.graphcanon.com/api/graphcanon/graph?tool=lightning-ai-litgpt","shared_categories":[]},{"slug":"om-ai-lab-omagent","name":"OmAgent","tagline":"Build multimodal language agents for fast prototype and production","github_url":"https://github.com/om-ai-lab/OmAgent","owner":"om-ai-lab","repo":"OmAgent","owner_avatar_url":"https://avatars.githubusercontent.com/u/96569904?v=4","primary_language":"Python","stars":2665,"forks":292,"topics":["agent","chatbot","gemini","gpt","gpt4","gradio","language-agent","large-language-models","llama","llava","llm","multimodal","multimodal-agent","openai","python","rag","smart-hardware","vision-and-language","vlm","workflow"],"archived":false,"github_pushed_at":"2025-03-19T11:36:13+00:00","maintenance_label":"Dormant","url":"https://www.graphcanon.com/tools/om-ai-lab-omagent","markdown_url":"https://www.graphcanon.com/tools/om-ai-lab-omagent.md","api_url":"https://www.graphcanon.com/api/graphcanon/tools/om-ai-lab-omagent","graph_url":"https://www.graphcanon.com/api/graphcanon/graph?tool=om-ai-lab-omagent","shared_categories":[]}]}}