This is a GGUF conversion of Google's T5 v1.1 XXL encoder model.
Source du modèle
Description de la source
This is a GGUF conversion of Google's T5 v1.1 XXL encoder model.
The weights can be used with ./llama-embedding or with the ComfyUI-GGUF custom node together with image generation models.
This is a quant as llama.cpp doesn't support imatrix creation for T5 models at the time of writing. It's therefore recommended to use for the best results, although smaller models may also still provide decent results in resource constrained scenarios.
Ressources connexes
2 connexionsSources
1 sourceVérifié 29 juil.
Artefacts du modèle
11 artefactsExtraits de sources
2 extraits--- base_model: google/t5-v1_1-xxl library_name: gguf license: apache-2.0 quantized_by: city96 language: en --- This is a GGUF conversion of Google's T5 v1.1 XXL encoder model. The weights can be used with [`./llama-embedding`](https://github.com/ggerganov/llama.cpp/tree/master/examples/embedding) or with the [ComfyUI-GGUF](https://github.com/city96/ComfyUI-GGUF) custom node together with image generation models. This is a **non imatrix** quant as llama.cpp doesn't support imatrix creation for T5 models at the time of writing. It's therefore recommended to use **Q5_K_M or larger** for the best results, although smaller models may also still provide decent results in resource constrained scenarios.
t5-v1_1-xxl-encoder-Q3_K_L.gguf
gguf · 2,29 GB · SHA-256 8f5ab8792343…f64b · Hugging Face
Source context: 50114 downloads · 532 likes · Library gguf · Repo city96/t5-v1_1-xxl-encoder-gguf