dia2 logo

dia2

nari-labs/dia2

TTS model capable of streaming conversational audio in real-time.

GraphCanon updated 3w · GitHub synced 3w

1.2k stars99 forksLast push 9mo Python Apache-2.0

Decision brief

Good fit when

  • When you need real-time streaming capabilities for text-to-speech applications.
  • If your project requires the integration of third-party codecs such as Kyutai Mimi which dia2 offers natively.

Avoid when

  • For projects aimed at batch processing or non-realtime TTS applications where instantaneous audio output isn't required.
  • If you are looking for an open-source model that does not include specific proprietary components like the Kyutai Mimi codec.

Observed Jul 12, 2026 · Source: enrich:decision_facts

Verify the decision

Maintenance and security

Full trust report
Maintenance
Slowing (243d since push)
As of 3w
Provenance
Not a fork · Organization account
As of 3w
Security (OSV)
No lockfile
As of 1mo

Public GitHub metadata and optional OSV scans. Signals, not a guarantee. Trust methodology.

Install

pip install dia2
PyPI

Similar tools

Same-category neighbours. No typed graph edges are catalogued for this tool yet.

Evidence and technical details

Sourced facts, taxonomy, compatibility claims, README excerpt, and machine-readable endpoints.

Overview

This is a text-to-speech model repository that supports real-time streaming of conversational audio using Python. It includes third-party assets like the Kyutai Mimi codec under their respective licenses, with the primary project licensed under Apache 2.0.

Capability facts

Languages
python

Source: github.language+pyproject.toml · Jul 30, 2026

Categories

Tags

README

License & Attribution

Licensed under Apache 2.0. All third-party assets (Kyutai Mimi codec, etc.) retain their original licenses.

For agents

This page has a .md twin and JSON over the API.

Was this helpful?

Anonymous feedback helps us improve pages and translations.