A first approach for general audio generation with high-dimensional LLM + Diffusion.
AI & ML interests
None defined yet.
Recent Activity
Organization Card
models 27
mispeech/midashenglm-spatial
Audio-Text-to-Text • 8B • Updated • 26 • 2
mispeech/midashenglm-gen
Text-to-Audio • 3B • Updated • 173 • 45
mispeech/Dasheng-AudioGen
Text-to-Audio • 2B • Updated • 551 • 18
mispeech/Dasheng-AudioGen-Multilingual
Text-to-Audio • 2B • Updated • 47 • 6
mispeech/dasheng-denoiser
Audio-to-Audio • 0.1B • Updated • 59 • 15
mispeech/dashengtokenizer
Audio-to-Audio • 0.8B • Updated • 762 • 15
mispeech/midashenglm-0.6b-gguf
Audio-Text-to-Text • 0.6B • Updated • 458 • 1
mispeech/midashenglm-7b-1021-gguf
Audio-Text-to-Text • 8B • Updated • 777 • 3
mispeech/midashenglm-0.6b-fp32
Audio-Text-to-Text • 0.7B • Updated • 597 • 4
mispeech/ced-base
Audio Classification • 85.7M • Updated • 20.1k • 17