This is a GGUF conversion of Google's T5 v1.1 XXL encoder model.
Fonte do modelo
Descrição da fonte
This is a GGUF conversion of Google's T5 v1.1 XXL encoder model.
The weights can be used with ./llama-embedding or with the ComfyUI-GGUF custom node together with image generation models.
This is a quant as llama.cpp doesn't support imatrix creation for T5 models at the time of writing. It's therefore recommended to use for the best results, although smaller models may also still provide decent results in resource constrained scenarios.
Recursos relacionados
2 conexõesFontes
1 fonteVerificado 29 de jul.
Artefatos de modelo
11 artefatosTrechos de fonte
2 trechos--- base_model: google/t5-v1_1-xxl library_name: gguf license: apache-2.0 quantized_by: city96 language: en --- This is a GGUF conversion of Google's T5 v1.1 XXL encoder model. The weights can be used with [`./llama-embedding`](https://github.com/ggerganov/llama.cpp/tree/master/examples/embedding) or with the [ComfyUI-GGUF](https://github.com/city96/ComfyUI-GGUF) custom node together with image generation models. This is a **non imatrix** quant as llama.cpp doesn't support imatrix creation for T5 models at the time of writing. It's therefore recommended to use **Q5_K_M or larger** for the best results, although smaller models may also still provide decent results in resource constrained scenarios.
t5-v1_1-xxl-encoder-Q3_K_L.gguf
gguf · 2,29 GB · SHA-256 8f5ab8792343…f64b · Hugging Face
Source context: 50114 downloads · 532 likes · Library gguf · Repo city96/t5-v1_1-xxl-encoder-gguf