{"data":{"slug":"espnet-espnet","name":"espnet","tagline":"End-to-End Speech Processing Toolkit","github_url":"https://github.com/espnet/espnet","owner":"espnet","repo":"espnet","owner_avatar_url":"https://avatars.githubusercontent.com/u/34493687?v=4","primary_language":"Python","stars":9903,"forks":2421,"topics":["chainer","deep-learning","end-to-end","kaldi","machine-translation","pytorch","singing-voice-synthesis","speaker-diarization","speech-enhancement","speech-recognition","speech-separation","speech-synthesis","speech-translation","spoken-language-understanding","text-to-speech","voice-conversion"],"archived":false,"github_pushed_at":"2026-07-28T14:36:55+00:00","maintenance_label":"Very active","url":"https://www.graphcanon.com/tools/espnet-espnet","markdown_url":"https://www.graphcanon.com/tools/espnet-espnet.md","api_url":"https://www.graphcanon.com/api/graphcanon/tools/espnet-espnet","graph_url":"https://www.graphcanon.com/api/graphcanon/graph?tool=espnet-espnet","description":"End-to-End Speech Processing Toolkit","homepage_url":"https://espnet.github.io/espnet/","license":"Apache-2.0","open_issues":49,"watchers":168,"ai_summary":"ESPNet offers tools for speech recognition, synthesis, translation, and more tasks employing deep learning models via Python.","readme_excerpt":"## Installation\n- If you intend to do full experiments, including DNN training, then see [Installation](https://espnet.github.io/espnet/installation.html).\n- If you just need the Python module only:\n    ```sh\n    # We recommend you install PyTorch before installing espnet following https://pytorch.org/get-started/locally/\n    pip install espnet\n    # To install the latest\n    # pip install git+https://github.com/espnet/espnet\n    # To install additional packages\n    # pip install \"espnet[all]\"\n    ```\n\n    If you use ESPnet1, please install chainer and cupy.\n\n    ```sh\n    pip install chainer==6.0.0 cupy==6.0.0    # [Option]\n    ```\n\n    You might need to install some packages depending on each task. We prepared various installation scripts at [tools/installers](tools/installers).\n\n- (ESPnet2) Once installed, run `wandb login` and set `--use_wandb true` to enable tracking runs using W&B.\n\n---\n\n## Docker Container\n\ngo to [docker/](docker/) and follow [instructions](https://espnet.github.io/espnet/docker.html).","github_created_at":"2017-12-13T00:45:11+00:00","created_at":"2026-07-11T12:05:26.447004+00:00","updated_at":"2026-07-29T06:00:43.111099+00:00","categories":[{"slug":"model-training","name":"Model Training","url":"https://www.graphcanon.com/categories/model-training","markdown_url":"https://www.graphcanon.com/categories/model-training.md","api_url":"https://www.graphcanon.com/api/graphcanon/categories/model-training"},{"slug":"speech-audio","name":"Speech & Audio","url":"https://www.graphcanon.com/categories/speech-audio","markdown_url":"https://www.graphcanon.com/categories/speech-audio.md","api_url":"https://www.graphcanon.com/api/graphcanon/categories/speech-audio"}],"tags":[{"slug":"chainer","name":"chainer"},{"slug":"deep-learning","name":"deep-learning"},{"slug":"kaldi","name":"kaldi"},{"slug":"pytorch","name":"pytorch"},{"slug":"speech-recognition","name":"speech-recognition"},{"slug":"speech-synthesis","name":"speech-synthesis"},{"slug":"speech-translation","name":"speech-translation"},{"slug":"spoken-language-understanding","name":"spoken-language-understanding"}],"trust":{"provenance":{"is_fork":false,"github_id":114054873,"owner_type":"Organization","methodology":"github_public_v1","parent_repo":null,"near_duplicate_slugs":[]},"computed_at":"2026-07-29T06:00:42.347Z","maintenance":{"label":"Very active","score":96,"methodology":"github_public_v1","releases_90d":0,"days_since_push":0,"last_release_at":"2026-04-22T14:11:58Z"},"security_summary":{"status":"no_lockfile","scanner":null,"low_count":0,"high_count":0,"last_scan_at":"2026-07-11T12:05:28.195Z","medium_count":0,"scan_profile":"none","critical_count":0}},"capability_facts":{"scan":{"source":"repo_scan","observed_at":"2026-07-29T06:00:42.778Z"},"languages":{"value":["python"],"source":"github.language+pyproject.toml","observed_at":"2026-07-29T06:00:42.778Z"},"license_spdx":{"value":"Apache-2.0","source":"github.license","observed_at":"2026-07-29T06:00:42.778Z"}},"decision_facts":{"hosting":null,"pricing":null,"requirements":null,"constraints":null,"when_to_use":["When you require comprehensive tools for end-to-end speech processing tasks such as speech recognition, synthesis, translation, and speaker diarization.","If your project involves working with Chainer or PyTorch frameworks for experimenting with DNN training in the domain of speech technology.","You need to use W&B (Weights & Biases) for tracking runs and prefer a toolkit that supports integration directly through configuration."],"when_not_to_use":["If you are working on tasks unrelated to speech or audio processing, such as computer vision, NLP, or any other deep learning areas outside of ESPNet's focus.","Your development environment is limited to languages other than Python or frameworks that do not support Chainer or PyTorch, which are foundational to espnet."],"source":"enrich:decision_facts","observed_at":"2026-07-17T13:29:08.555Z"},"constraint_facets":null,"decision_summary":[{"label":"Adopt for","value":"ESPNet is an End-to-End Speech Processing Toolkit that employs deep learning models for tasks including speech recognition and synthesis."}]}}