Text-to-audio and video-to-audio using Sony AI's Woosh foundation model.
-
Updated
May 7, 2026 - Python
Text-to-audio and video-to-audio using Sony AI's Woosh foundation model.
Tool-to-Agent Protocol: tools can be smart without embedded LLM calls.
Portable C++17 implementation of mispeech/Dasheng-AudioGen a 2B-parameter flow-matching text-to-audio model using GGML
To associate your repository with the t2a topic, visit your repo's landing page and select "manage topics."