Published: January 3, 2025
6
94
910
new extremely fast text-to-audio model
this is TangoFlux, a new text-to-audio model that can generate 30 seconds of 44.1kHz audio in just 3.7 seconds on a single A40 GPU project page: https://tangoflux.github.io/ code: https://github.com/declare-lab... demo: https://huggingface.co/spaces/...
@michieldoteth donβt think this is gonna hold up for it but report back!
@NimEshed not sure how it has evolved, but check out tortoise and t5-tts
@pureshimon enjoy
