☆ Yσɠƚԋσʂ ☆@lemmy.ml to Technology@lemmy.mlEnglish · 7 days agoQwen-Audio-3.0-TTSfunaudiollm.github.ioexternal-linkmessage-square3linkfedilinkarrow-up114arrow-down12cross-posted to: technology@lemmygrad.mlAii@programming.dev
arrow-up112arrow-down1external-linkQwen-Audio-3.0-TTSfunaudiollm.github.io☆ Yσɠƚԋσʂ ☆@lemmy.ml to Technology@lemmy.mlEnglish · 7 days agomessage-square3linkfedilinkcross-posted to: technology@lemmygrad.mlAii@programming.dev
minus-squareloathsome dongeater@lemmygrad.mllinkfedilinkEnglisharrow-up1·7 days agoSorry but what kind of specs are needed to run something like this? Is it more or less demanding than a 8B moderately quantised LLM?
minus-square☆ Yσɠƚԋσʂ ☆@lemmy.mlOPlinkfedilinkarrow-up3·7 days agoyeah the models are tiny, just 2b https://huggingface.co/collections/Qwen/qwen3-tts
Sorry but what kind of specs are needed to run something like this? Is it more or less demanding than a 8B moderately quantised LLM?
yeah the models are tiny, just 2b https://huggingface.co/collections/Qwen/qwen3-tts
Try out CrispASR.