A first approach for general audio generation with high-dimensional LLM + Diffusion.
AI & ML interests
None defined yet.
Recent Activity
Organization Card
models 26
mispeech/midashenglm-gen
Text-to-Audio • 3B • Updated • 525 • 36
mispeech/Dasheng-AudioGen
Text-to-Audio • 2B • Updated • 519 • 15
mispeech/Dasheng-AudioGen-Multilingual
Text-to-Audio • 2B • Updated • 35 • 6
mispeech/dasheng-denoiser
Audio-to-Audio • 0.1B • Updated • 77 • 15
mispeech/dashengtokenizer
Audio-to-Audio • 0.8B • Updated • 6.71k • 12
mispeech/midashenglm-0.6b-gguf
Audio-Text-to-Text • 0.6B • Updated • 168 • 1
mispeech/midashenglm-7b-1021-gguf
Audio-Text-to-Text • 8B • Updated • 242 • 3
mispeech/midashenglm-0.6b-fp32
Audio-Text-to-Text • 0.7B • Updated • 230 • 4
mispeech/ced-base
Audio Classification • 85.7M • Updated • 9.07k • 14
mispeech/ced-tiny
Audio Classification • 5.5M • Updated • 3.56k • 4