This is a GGUF conversion of Google's T5 v1.1 XXL encoder model. The weights can be used with ./llama embedding or with the ComfyUI GGUF custom node together with image generation models. This is a non imatrix quant as llama.cpp doesn't support imatrix creation for T5 models at the time of writing. It's therefore recommended to use Q5 K M or larger for the best results, although smaller models may also still provide decent results in resource constrained scenarios.
We use cookies for essential functionality and analytics. You can accept or reject analytics cookies.Cookie policy