posted an update

I’ve released the Windows version! With CUDA + FlashAttention 2 support, systems with a compatible GPU can generate audio in about one second after the initial model load. A Linux version is also available, but it hasn’t been tested yet. Feedback is very welcome!

Log in or sign up for Devpost to join the conversation.