x-stable-diffusion
Real-time inference for Stable Diffusion - 0.88s latency
GraphCanon updated 3w · GitHub synced 3w
Decision brief
x-stable-diffusion offers real-time inference for the Stable Diffusion model with a latency of 0.88s, leveraging AITemplate, nvFuser, TensorRT, and FlashAttention.
Good fit when
- When you require low-latency real-time inference performance at less than 1 second
- If your project involves optimizing Stable Diffusion specifically
Avoid when
- For projects that do not require real-time performance or have higher latency tolerance
- If the specific optimizations for Stable Diffusion are not aligned with your model needs
Observed Jul 12, 2026 · Source: enrich:decision_facts
Verify the decision
Maintenance and security
Full trust report- Maintenance
- Archived (971d since push)
- As of 3w
- Provenance
- Not a fork · Organization account
- As of 3w
- Security (OSV)
- No lockfile
- As of 1mo
Public GitHub metadata and optional OSV scans. Signals, not a guarantee. Trust methodology.
Install
git clone https://github.com/stochasticai/x-stable-diffusionHow it fits your stack(1)
Typed graph edges - alternatives, integrations, successors, and dependencies. Ranked by relationship type, not raw GitHub stars.
Relationship graph
Optional deeper exploration of typed edges and category neighbours.
Similar tools
Same-category neighbours not already linked as typed edges.
Evidence and technical details
Sourced facts, taxonomy, compatibility claims, README excerpt, and machine-readable endpoints.
Overview
A repository focused on optimizing Stable Diffusion model inference using various tools including AITemplate, nvFuser, TensorRT, and FlashAttention.
Capability facts
- Languages
- jupyter notebook
Source: github.language · Aug 2, 2026
Categories
Tags
README
Manual deployment
Check the README.md of the following directories:
- AITemplate
- FlashAttention
- nvFuser
- PyTorch
- TensorRT
For agents
This page has a .md twin and JSON over the API.