This directory contains audio.cpp-native GGUF conversions of multiple speech models. These files are intended for use with audio.cpp.
Modellquelle
Quellenbeschreibung
This directory contains audio.cpp-native GGUF conversions of multiple speech models. These files are intended for use with audio.cpp.
For conversion details, supported layouts, direct-file loading, sidecar embedding, and the latest compatibility notes, see the audio.cpp GGUF guide:
Quellen
1 QuelleVerifiziert 3. Aug.
Modellartefakte
64 ArtefakteACE-Step1.5-GGUF/base/ace-step-1.5-base-bf16.gguf
gguf · 9,40 GB · SHA-256 aa4b0f27122b…7cf4 · Hugging Face
HerunterladenACE-Step1.5-GGUF/base/ace-step-1.5-base-q8_0.gguf
gguf · 5,76 GB · SHA-256 761337100558…d5ba · Hugging Face
HerunterladenQuellenauszüge
2 Auszüge!!! Automated audio checks are intentionally strict and may flag length, log-mel, or transcript drift that can still sound acceptable to human listeners. Validate the exact file, backend, and route you plan to use. These converted weights are provided as-is; use them at your own risk.
Tested summarizes the current audio.cpp path-test status. See the GGUF guide above
for the full matrix and drift notes.
| Directory | Files | audio.cpp family | Tested | Original model license |
|---|---|---|---|---|
ACE-Step1.5-GGUF | base/ace-step-1.5-base-bf16.gguf, base/ace-step-1.5-base-q8_0.gguf, turbo/ace-step-1.5-turbo-bf16.gguf, turbo/ace-step-1.5-turbo-q8_0.gguf | ace_step | 16-bit + Q8 drift | See original model package |
BS-RoFormer-ep368-GGUF | bs-roformer-ep368-q8_0.gguf | bs_roformer | Q8 pass | See original model package |
Chatterbox-GGUF | chatterbox-f16.gguf, chatterbox-q8_0.gguf | chatterbox | 16-bit + Q8 ASR-match drift | MIT |
Citrinet-ASR-GGUF | citrinet-asr-q8_0.gguf |
flow_lm.flow_net.time_embed.*.mlp.{0,2}.weight
tensors in Q8 in addition to the default converter selection. conditioner.embed,
cond_embed, and Mimi conv tensors are not forced to Q8 because tested outputs
drifted or the current conv path casts quantized conv weights back to F32.Pass a GGUF file directly as --model:
audiocpp_cli --task tts --family supertonic --model Supertonic-3-GGUF/supertonic-3-orig.gguf --backend cuda --language en --text "Hello." --voice-id M1 --out out.wav
For ASR:
audiocpp_cli --task asr --family qwen3_asr --model Qwen3-ASR-0.6B-GGUF/qwen3-asr-0.6b-f16.gguf --backend cuda --audio speech.wav --text "" --text-out transcript.txt
Each GGUF file is a converted form of its original model. Use and redistribution are governed by the corresponding original model license listed above. Please review the original model card and license terms before using or redistributing any converted weights.
Anforderungen
1 AnforderungVAE · Hugging Face · audio-cpp/audio.cpp-gguf
ACE-Step1.5-GGUF/turbo/ace-step-1.5-turbo-bf16.gguf
gguf · 9,40 GB · SHA-256 93974239a29a…be5f · Hugging Face
HerunterladenACE-Step1.5-GGUF/turbo/ace-step-1.5-turbo-q8_0.gguf
gguf · 5,76 GB · SHA-256 cd7bf272588f…270b · Hugging Face
HerunterladenBS-RoFormer-ep368-GGUF/bs-roformer-ep368-q8_0.gguf
gguf · 165 MB · SHA-256 9a55a8cad369…3ad2 · Hugging Face
HerunterladenChatterbox-GGUF/chatterbox-f16.gguf
gguf · 3,49 GB · SHA-256 b872f908605f…e713 · Hugging Face
HerunterladenChatterbox-GGUF/chatterbox-q8_0.gguf
gguf · 1,94 GB · SHA-256 d586dd1aa596…0656 · Hugging Face
HerunterladenCitrinet-ASR-GGUF/citrinet-asr-q8_0.gguf
gguf · 38,7 MB · SHA-256 e1b9187ecc1c…6368 · Hugging Face
HerunterladenConfucius4-TTS-GGUF/confucius4-tts-orig.gguf
gguf · 7,63 GB · SHA-256 eec4ab3fae3c…6244 · Hugging Face
HerunterladenDotTTS-MF-GGUF/dots-tts-mf-bf16.gguf
gguf · 4,46 GB · SHA-256 384de4f85839…1214 · Hugging Face
HerunterladenDotTTS-SOAR-GGUF/dots-tts-soar-bf16.gguf
gguf · 4,46 GB · SHA-256 5292089b053f…a41c · Hugging Face
HerunterladenDotTTS-SOAR-GGUF/dots-tts-soar-orig.gguf
gguf · 4,81 GB · SHA-256 d4406cc394b9…3e67 · Hugging Face
HerunterladenDramaBox-GGUF/dramabox-q8_0.gguf
gguf · 17,6 GB · SHA-256 75e7e80fc748…451e · Hugging Face
HerunterladenFish-Audio-S2-Pro-GGUF/fish-audio-s2-pro-bf16.gguf
gguf · 9,53 GB · SHA-256 781fdece3ff8…96ff · Hugging Face
HerunterladenFish-Audio-S2-Pro-GGUF/fish-audio-s2-pro-q8_0.gguf
gguf · 5,88 GB · SHA-256 4ffc169447b7…9d0b · Hugging Face
HerunterladenFun-ASR-Nano-2512-GGUF/fun-asr-nano-2512-f16.gguf
gguf · 1,56 GB · SHA-256 3d906c3ccfed…c5be · Hugging Face
HerunterladenFun-ASR-Nano-2512-GGUF/fun-asr-nano-2512-q8_0.gguf
gguf · 997 MB · SHA-256 4d727357574b…5e0a · Hugging Face
Herunterladengemma-3-12b-it-qat-GGUF/gemma-3-12b-it-Q8_0.gguf
gguf · 11,7 GB · SHA-256 a774cd9c6f28…5982 · Hugging Face
Herunterladengemma-3-12b-it-qat-GGUF/gemma-3-12b-it-qat-UD-Q4_K_XL.gguf
gguf · 6,92 GB · SHA-256 da98f81c8691…b53b · Hugging Face
Herunterladengemma-3-12b-it-qat-GGUF/mmproj-BF16.gguf
gguf · 815 MB · SHA-256 e63e63bf219e…e4c5 · Hugging Face
HerunterladenHeartMuLa-GGUF/heartmula-f16.gguf
gguf · 10,4 GB · SHA-256 77d47b5c4ffd…fdf3 · Hugging Face
HerunterladenHeartMuLa-GGUF/heartmula-q8_0.gguf
gguf · 7,13 GB · SHA-256 5c3e42171ec0…cfbd · Hugging Face
HerunterladenHiggs-Audio-v3-STT-GGUF/higgs-audio-v3-stt-f16.gguf
gguf · 5,00 GB · SHA-256 1deab3267267…0c29 · Hugging Face
HerunterladenHiggs-Audio-v3-STT-GGUF/higgs-audio-v3-stt-q8_0.gguf
gguf · 2,94 GB · SHA-256 4149e6820175…d849 · Hugging Face
HerunterladenHiggs-Audio-v3-TTS-4B-GGUF/higgs-audio-v3-tts-4b-bf16.gguf
gguf · 7,92 GB · SHA-256 2c04410d9ae6…69dc · Hugging Face
HerunterladenHiggs-Audio-v3-TTS-4B-GGUF/higgs-audio-v3-tts-4b-q8_0.gguf
gguf · 4,75 GB · SHA-256 79746822045b…a74a · Hugging Face
HerunterladenHTDemucs-GGUF/htdemucs-f16.gguf
gguf · 80,2 MB · SHA-256 f8c54ba35df9…86b6 · Hugging Face
HerunterladenHTDemucs-GGUF/htdemucs-q8_0.gguf
gguf · 59,1 MB · SHA-256 b0f532ac6e5f…4388 · Hugging Face
HerunterladenHviske-v5.3-GGUF/hviske-v5.3-q8_0.gguf
gguf · 2,27 GB · SHA-256 9f910b2610e6…f0ad · Hugging Face
HerunterladenIndexTTS2-GGUF/index-tts2-f16.gguf
gguf · 4,33 GB · SHA-256 0f7b95d3d32e…1523 · Hugging Face
HerunterladenIndexTTS2-GGUF/index-tts2-orig.gguf
gguf · 7,53 GB · SHA-256 c96cb948339b…8e16 · Hugging Face
HerunterladenIndexTTS2-GGUF/index-tts2-q8_0.gguf
gguf · 3,38 GB · SHA-256 f9be73ac5b6c…2e8e · Hugging Face
HerunterladenInflect-Micro-v2-GGUF/inflect-micro-v2-orig.gguf
gguf · 68,7 MB · SHA-256 d4af1cb6a92c…4b03 · Hugging Face
HerunterladenIrodori-TTS-500M-v3-GGUF/irodori-tts-500m-v3-f16.gguf
gguf · 1,17 GB · SHA-256 bc58da653f5d…8318 · Hugging Face
HerunterladenIrodori-TTS-500M-v3-GGUF/irodori-tts-500m-v3-q8_0.gguf
gguf · 1,02 GB · SHA-256 6f1d8d84f922…b0c4 · Hugging Face
HerunterladenIrodori-TTS-600M-v3-VoiceDesign-GGUF/irodori-tts-600m-v3-voicedesign-f16.gguf
gguf · 1,36 GB · SHA-256 b7a1b999f6b3…5568 · Hugging Face
HerunterladenIrodori-TTS-600M-v3-VoiceDesign-GGUF/irodori-tts-600m-v3-voicedesign-q8_0.gguf
gguf · 1,18 GB · SHA-256 eda37824d020…5b33 · Hugging Face
HerunterladenKroko-ASR-GGUF/kroko-en-community-64-l-q8_0.gguf
gguf · 160 MB · SHA-256 596c749fcb6c…baa5 · Hugging Face
HerunterladenLTX-2.3-GGUF-unsloth/ltx-2.3-22b-dev-Q8_0.gguf
gguf · 21,2 GB · SHA-256 c4e78967e6c6…145c · Hugging Face
HerunterladenLTX-2.3-GGUF-unsloth/text_encoders/ltx-2.3-22b-dev_embeddings_connectors.safetensors
safetensors · 2,15 GB · SHA-256 a5c5148788d8…a8ae · Hugging Face
HerunterladenLTX-2.3-GGUF-unsloth/vae/ltx-2.3-22b-dev_audio_vae.safetensors
safetensors · 348 MB · SHA-256 d7711812d938…4dea · Hugging Face
HerunterladenLTX-2.3-GGUF-unsloth/vae/ltx-2.3-22b-dev_video_vae.safetensors
safetensors · 1,35 GB · SHA-256 8732bb70cf43…3f4e · Hugging Face
HerunterladenMagpieTTS-Multilingual-357M-GGUF/magpie-tts-multilingual-357m-orig.gguf
gguf · 1,78 GB · SHA-256 3ed26e41e19e…096f · Hugging Face
HerunterladenMel-Band-RoFormer-GGUF/mel-band-roformer-f16.gguf
gguf · 435 MB · SHA-256 cbdf174dc716…7add · Hugging Face
HerunterladenMel-Band-RoFormer-GGUF/mel-band-roformer-q8_0.gguf
gguf · 240 MB · SHA-256 2dd898ceb0e3…38fd · Hugging Face
HerunterladenMioCodec-25Hz-44.1kHz-v2-GGUF/miocodec-25hz-44khz-v2-f16.gguf
gguf · 432 MB · SHA-256 9f8eb1e69948…d500 · Hugging Face
HerunterladenMioCodec-25Hz-44.1kHz-v2-GGUF/miocodec-25hz-44khz-v2-orig.gguf
gguf · 864 MB · SHA-256 966ff5516355…dabc · Hugging Face
HerunterladenMioCodec-25Hz-44.1kHz-v2-GGUF/miocodec-25hz-44khz-v2-q8_0.gguf
gguf · 285 MB · SHA-256 4c76a3631349…b3bb · Hugging Face
HerunterladenMioTTS-1.7B-GGUF/miotts-1.7b-bf16.gguf
gguf · 3,28 GB · SHA-256 0f035aa84873…2835 · Hugging Face
HerunterladenMioTTS-1.7B-GGUF/miotts-1.7b-orig.gguf
gguf · 3,28 GB · SHA-256 29f5dcfbf5a9…46ed · Hugging Face
HerunterladenMioTTS-1.7B-GGUF/miotts-1.7b-q8_0.gguf
gguf · 2,05 GB · SHA-256 3496847df8bd…3cd8 · Hugging Face
HerunterladenMOSS-TTS-Local-v1.5-GGUF/moss-tts-local-v1.5-bf16.gguf
gguf · 12,4 GB · SHA-256 9ff6c86f3d4c…9415 · Hugging Face
HerunterladenMOSS-TTS-Local-v1.5-GGUF/moss-tts-local-v1.5-q8_0.gguf
gguf · 7,00 GB · SHA-256 ce2d34f73274…08ea · Hugging Face
HerunterladenMOSS-TTS-Nano-100M-GGUF/moss-tts-nano-100m-bf16.gguf
gguf · 317 MB · SHA-256 249a7eb9a8ae…de0b · Hugging Face
HerunterladenMOSS-TTS-Nano-100M-GGUF/moss-tts-nano-100m-q8_0.gguf
gguf · 184 MB · SHA-256 3b9c138ce409…e388 · Hugging Face
HerunterladenNemotron-3.5-ASR-Streaming-0.6B-GGUF/nemotron-3.5-asr-streaming-0.6b-f16.gguf
gguf · 1,19 GB · SHA-256 8bef32306425…137e · Hugging Face
HerunterladenNemotron-3.5-ASR-Streaming-0.6B-GGUF/nemotron-3.5-asr-streaming-0.6b-q8_0.gguf
gguf · 888 MB · SHA-256 7026c80d8b94…2afe · Hugging Face
HerunterladenOmniVoice-GGUF/omnivoice-bf16.gguf
gguf · 1,53 GB · SHA-256 7301a41a0a5a…819c · Hugging Face
HerunterladenOmniVoice-GGUF/omnivoice-f16.gguf
gguf · 1,53 GB · SHA-256 7421b6c38735…ab63 · Hugging Face
HerunterladenOmniVoice-GGUF/omnivoice-q8_0.gguf
gguf · 1,26 GB · SHA-256 2f4be6372780…da8b · Hugging Face
HerunterladenParakeet-TDT-0.6B-v3-GGUF/parakeet-tdt-0.6b-v3-f16.gguf
gguf · 1,17 GB · SHA-256 83622c442756…8649 · Hugging Face
HerunterladenParakeet-TDT-0.6B-v3-GGUF/parakeet-tdt-0.6b-v3-q8_0.gguf
gguf · 873 MB · SHA-256 074e61ac1abd…e64c · Hugging Face
HerunterladenPocketTTS-GGUF/english/embeddings/alba.safetensors
safetensors · 5,91 MB · SHA-256 69c32db63ca5…8845 · Hugging Face
HerunterladenPocketTTS-GGUF/english/embeddings/anna.safetensors
safetensors · 7,45 MB · SHA-256 5ea82f78db00…0dfa · Hugging Face
Herunterladen--- license: other library_name: audio.cpp pipeline_tag: text-to-speech tags: - gguf - audio.cpp - quantized - text-to-speech - automatic-speech-recognition - voice-conversion - text-to-audio - audio-to-audio - source-separation - speaker-diarization - speech base_model_relation: quantized base_model: - ACE-Step/Ace-Step1.5 - ACE-Step/acestep-v15-base - Aratako/Irodori-TTS-500M-v3 - Aratako/Irodori-TTS-600M-v3-VoiceDesign - Aratako/MioCodec-25Hz-44.1kHz-v2 - Aratako/MioTTS-1.7B - Aratako/Semantic-DACVAE-Japanese-32dim - Banafo/Kroko-ASR - dots-studio/dots.tts-soar - fishaudio/s2-pro - FunAudioLLM/Fun-ASR-Nano-2512-hf - HeartMuLa/HeartCodec-oss-20260123 - HeartMuLa/HeartMuLa-oss-3B - HeartMuLa/HeartMuLaGen - OpenBMB/VoxCPM2 - OpenMOSS-Team/MOSS-Audio-Tokenizer-Nano - OpenMOSS-Team/MOSS-Audio-Tokenizer-v2 - OpenMOSS-Team/MOSS-TTS-Local-Transformer-v1.5 - OpenMOSS-Team/MOSS-TTS-Nano-100M - Qwen/Qwen3-ASR-0.6B - Qwen/Qwen3-ASR-1.7B-hf - Qwen/Qwen3-ForcedAligner-0.6B - Qwen/Qwen3-TTS-12Hz-1.7B-Base - Qwen/Qwen3-TTS-12Hz-1.7B-CustomVoice - Qwen/Qwen3-TTS-12Hz-1.7B-VoiceDesign - Qwen/Qwen3-TTS-Tokenizer-12Hz - ResembleAI/chatterbox - RMSnow/Vevo2 - bosonai/higgs-audio-v3-stt - bosonai/higgs-audio-v3-tts-4b - k2-fsa/OmniVoice - kyutai/pocket-tts - llm-jp/llm-jp-3-150m - microsoft/VibeVoice-1.5B - microsoft/VibeVoice-ASR - mistralai/Voxtral-Mini-4B-Realtime-2602 - mirek190/audio.cpp - mlx-community/SeedVC-MLX - mlx-community/index-tts2-mlx - mlx-community/mel-roformer-mlx - mlx-community/supertonic-3-mlx - mlx-community/wavlm-base-plus-mlx - nvidia/diar_sortformer_4spk-v1 - nvidia/nemotron-3.5-asr-streaming-0.6b - nvidia/parakeet-tdt-0.6b-v3 - owensong/Inflect-Micro-v2 - stabilityai/stable-audio-3-medium - stabilityai/stable-audio-3-small-music - stabilityai/stable-audio-3-small-sfx --- # audio.cpp GGUF Model Packages This directory contains audio.cpp-native GGUF conversions of multiple speech models. These files are intended for use with [audio.cpp](https://github.com/0xShug0/audio.cpp). For conversion details, supported layouts, direct-file loading, sidecar embedding, and the latest compatibility notes, see the audio.cpp GGUF guide: - https://github.com/0xShug0/audio.cpp/blob/main/docs/gguf.md !!! Automated audio checks are intentionally strict and may flag length, log-mel, or transcript drift that can still sound acceptable to human listeners. Validate the...
Source context: 61538 downloads · 27 likes · Pipeline text-to-speech · Library audio.cpp · Repo audio-cpp/audio.cpp-gguf
citrinet_asr |
| Q8 pass |
| CC-BY-4.0 |
Confucius4-TTS-GGUF | confucius4-tts-orig.gguf | confucius4_tts | orig pass | See original model package |
DotTTS-SOAR-GGUF | dots-tts-soar-orig.gguf, dots-tts-soar-bf16.gguf | dots_tts | experimental | See original model package |
DramaBox-GGUF | dramabox-q8_0.gguf | dramabox | Q8 pass | See original model package |
Fish-Audio-S2-Pro-GGUF | fish-audio-s2-pro-bf16.gguf, fish-audio-s2-pro-q8_0.gguf | fish_audio | 16-bit + Q8 pass | See original model package |
Fun-ASR-Nano-2512-GGUF | fun-asr-nano-2512-f16.gguf, fun-asr-nano-2512-q8_0.gguf | fun_asr_nano | 16-bit + Q8 pass | FunASR Model Open Source License Agreement v1.1 |
HeartMuLa-GGUF | heartmula-f16.gguf, heartmula-q8_0.gguf | heartmula | 16-bit + Q8 drift | Apache-2.0 |
HTDemucs-GGUF | htdemucs-f16.gguf, htdemucs-q8_0.gguf | htdemucs | 16-bit pass, Q8 drift | See original model package |
Higgs-Audio-v3-STT-GGUF | higgs-audio-v3-stt-f16.gguf, higgs-audio-v3-stt-q8_0.gguf | higgs_audio_stt | 16-bit + Q8 pass | Apache-2.0 |
Higgs-Audio-v3-TTS-4B-GGUF | higgs-audio-v3-tts-4b-bf16.gguf, higgs-audio-v3-tts-4b-q8_0.gguf | higgs_audio_tts | 16-bit + Q8 pass | See original model package |
IndexTTS2-GGUF | index-tts2-orig.gguf, index-tts2-f16.gguf, index-tts2-q8_0.gguf | index_tts2 | orig + 16-bit pass/drift, Q8 ASR-match drift | bilibili Model Use License Agreement |
Inflect-Micro-v2-GGUF | inflect-micro-v2-orig.gguf | inflect_v2 | orig pass | See original model package |
Irodori-TTS-500M-v3-GGUF | irodori-tts-500m-v3-f16.gguf, irodori-tts-500m-v3-q8_0.gguf | irodori_tts | 16-bit pass, Q8 drift | MIT |
Irodori-TTS-600M-v3-VoiceDesign-GGUF | irodori-tts-600m-v3-voicedesign-f16.gguf, irodori-tts-600m-v3-voicedesign-q8_0.gguf | irodori_tts | 16-bit pass, Q8 drift | MIT |
Kroko-ASR-GGUF | kroko-en-community-64-l-q8_0.gguf | kroko_asr | Q8 pass | See original model package |
MOSS-TTS-Local-v1.5-GGUF | moss-tts-local-v1.5-bf16.gguf, moss-tts-local-v1.5-q8_0.gguf | moss_tts_local | 16-bit pass, Q8 ASR-match drift | Apache-2.0 |
MOSS-TTS-Nano-100M-GGUF | moss-tts-nano-100m-bf16.gguf, moss-tts-nano-100m-q8_0.gguf | moss_tts_nano | 16-bit pass, Q8 ASR-match drift | Apache-2.0 |
Mel-Band-RoFormer-GGUF | mel-band-roformer-f16.gguf, mel-band-roformer-q8_0.gguf | mel_band_roformer | 16-bit + Q8 drift | MIT |
MioCodec-25Hz-44.1kHz-v2-GGUF | miocodec-25hz-44khz-v2-orig.gguf, miocodec-25hz-44khz-v2-f16.gguf, miocodec-25hz-44khz-v2-q8_0.gguf | miocodec | orig pass, 16-bit + Q8 drift | MIT |
MioTTS-1.7B-GGUF | miotts-1.7b-orig.gguf, miotts-1.7b-bf16.gguf, miotts-1.7b-q8_0.gguf | miotts | orig pass, 16-bit drift, Q8 ASR-match drift | Apache-2.0 |
Nemotron-3.5-ASR-Streaming-0.6B-GGUF | nemotron-3.5-asr-streaming-0.6b-f16.gguf, nemotron-3.5-asr-streaming-0.6b-q8_0.gguf | nemotron_asr | 16-bit pass, Q8 minor filler drift | OpenMDW-1.1 |
OmniVoice-GGUF | omnivoice-bf16.gguf, omnivoice-f16.gguf, omnivoice-q8_0.gguf | omnivoice | 16-bit + Q8 drift | Apache-2.0 |
Parakeet-TDT-0.6B-v3-GGUF | parakeet-tdt-0.6b-v3-f16.gguf, parakeet-tdt-0.6b-v3-q8_0.gguf | parakeet_tdt | 16-bit + Q8 pass | See original model package |
PocketTTS-GGUF | english/, german/, italian/, portuguese/, spanish/ each contain bf16 and q8_0 GGUFs | pocket_tts | 16-bit pass, Q8 drift | See original model package |
Qwen3-ASR-0.6B-GGUF | qwen3-asr-0.6b-f16.gguf, qwen3-asr-0.6b-q8_0.gguf | qwen3_asr | 16-bit + Q8 pass | Apache-2.0 |
Qwen3-ASR-1.7B-GGUF | qwen3-asr-1.7b-f16.gguf, qwen3-asr-1.7b-q8_0.gguf | qwen3_asr | 16-bit + Q8 pass | Apache-2.0 |
Qwen3-ForcedAligner-0.6B-GGUF | qwen3-forced-aligner-0.6b-f16.gguf, qwen3-forced-aligner-0.6b-q8_0.gguf | qwen3_forced_aligner | 16-bit + Q8 pass | Apache-2.0 |
Qwen3-TTS-12Hz-1.7B-Base-GGUF | qwen3-tts-12hz-1.7b-base-orig.gguf, qwen3-tts-12hz-1.7b-base-bf16.gguf, qwen3-tts-12hz-1.7b-base-q8_0_v2.gguf | qwen3_tts | orig pass, 16-bit + Q8 ASR-match drift | Apache-2.0 |
Qwen3-TTS-12Hz-1.7B-CustomVoice-GGUF | qwen3-tts-12hz-1.7b-customvoice-bf16.gguf, qwen3-tts-12hz-1.7b-customvoice-q8_0.gguf | qwen3_tts | 16-bit + Q8 ASR-match drift | Apache-2.0 |
Qwen3-TTS-12Hz-1.7B-VoiceDesign-GGUF | qwen3-tts-12hz-1.7b-voicedesign-bf16.gguf, qwen3-tts-12hz-1.7b-voicedesign-q8_0.gguf | qwen3_tts | 16-bit + Q8 ASR-match drift | Apache-2.0 |
RVC-GGUF | rvc-f16.gguf | rvc | F16 pass | See original model package |
SeedVC-MLX-GGUF | seed-vc-mlx-orig.gguf, seed-vc-mlx-f16.gguf, seed-vc-mlx-q8_0.gguf | seed_vc | 16-bit + Q8 drift | GPL-3.0 |
Sortformer-Diar-4spk-v1-GGUF | sortformer-diar-4spk-v1-f16.gguf, sortformer-diar-4spk-v1-q8_0.gguf | sortformer_diar | 16-bit + Q8 pass | CC-BY-NC-4.0 |
Stable-Audio-3-Medium-GGUF | stable-audio-3-medium-f16.gguf, stable-audio-3-medium-q8_0.gguf | stable_audio | 16-bit + Q8 drift | Stability AI Community License |
Stable-Audio-3-Small-Music-GGUF | stable-audio-3-small-music-f16.gguf, stable-audio-3-small-music-q8_0.gguf | stable_audio | 16-bit + Q8 drift | Stability AI Community License |
Stable-Audio-3-Small-SFX-GGUF | stable-audio-3-small-sfx-f16.gguf, stable-audio-3-small-sfx-q8_0.gguf | stable_audio | 16-bit + Q8 drift | Stability AI Community License |
Supertonic-3-GGUF | supertonic-3-orig.gguf, supertonic-3-f16.gguf, supertonic-3-q8_0.gguf | supertonic | F32/orig pass; f16 not tested; Q8 unsupported dtype | BigScience Open RAIL-M |
Vevo2-GGUF | vevo2-orig.gguf, vevo2-f16.gguf, vevo2-q8_0.gguf | vevo2 | orig + 16-bit pass/drift; Q8 mixed route drift | See original model package |
VibeVoice-1.5B-GGUF | vibevoice-1.5b-bf16.gguf, vibevoice-1.5b-q8_0.gguf | vibevoice | 16-bit pass, Q8 drift | MIT |
VibeVoice-ASR-GGUF | vibevoice-asr-f16.gguf, vibevoice-asr-q8_0.gguf | vibevoice_asr | 16-bit + Q8 pass | MIT |
VoxCPM2-GGUF | voxcpm2-orig.gguf, voxcpm2-bf16.gguf, voxcpm2-q8_0.gguf | voxcpm2 | orig pass, 16-bit + Q8 ASR-match drift | Apache-2.0 |
Voxtral-Mini-4B-Realtime-2602-GGUF | voxtral-mini-4b-realtime-2602-bf16.gguf, voxtral-mini-4b-realtime-2602-q8_0.gguf, voxtral-mini-4b-realtime-2602-q4_k.gguf | voxtral_realtime | 16-bit + Q8 pass; Q4_K quick check passed | Apache-2.0 |