A first approach for general audio generation with high-dimensional LLM + Diffusion.
AI & ML interests
None defined yet.
Recent Activity
Organization Card
models 27
mispeech/midashenglm-spatial
Audio-Text-to-Text • 8B • Updated • 2
mispeech/midashenglm-gen
Text-to-Audio • 3B • Updated • 184 • 45
mispeech/Dasheng-AudioGen
Text-to-Audio • 2B • Updated • 527 • 18
mispeech/Dasheng-AudioGen-Multilingual
Text-to-Audio • 2B • Updated • 47 • 6
mispeech/dasheng-denoiser
Audio-to-Audio • 0.1B • Updated • 71 • 15
mispeech/dashengtokenizer
Audio-to-Audio • 0.8B • Updated • 758 • 15
mispeech/midashenglm-0.6b-gguf
Audio-Text-to-Text • 0.6B • Updated • 426 • 1
mispeech/midashenglm-7b-1021-gguf
Audio-Text-to-Text • 8B • Updated • 500 • 3
mispeech/midashenglm-0.6b-fp32
Audio-Text-to-Text • 0.7B • Updated • 567 • 4
mispeech/ced-base
Audio Classification • 85.7M • Updated • 20.5k • 17