For a convenient overview and download list, visit our model page for this model.
Fuente del modelo
Descripción de la fuente
weighted/imatrix quants of https://huggingface.co/0xA50C1A1/Llama-3.3-70B-Instruct-SOM-MPOA
For a convenient overview and download list, visit our model page for this model.
static quants are available at
Fuentes
1 fuenteVerificado 6 ago
Artefactos del modelo
24 artefactosExtractos de fuentes
3 extractosIf you are unsure how to use GGUF files, refer to one of TheBloke's READMEs for more details, including on how to concatenate multi-part files.
(sorted by size, not necessarily quality. IQ-quants are often preferable over similar sized non-IQ quants)
| Link | Type | Size/GB | Notes |
|---|---|---|---|
| GGUF | imatrix | 0.1 | imatrix file (for creating your own quants) |
| GGUF | i1-IQ1_S | 15.4 | for the desperate |
| GGUF | i1-IQ1_M | 16.9 | mostly desperate |
| GGUF | i1-IQ2_XXS | 19.2 |
Here is a handy graph by ikawrakow comparing some lower-quality quant types (lower is better):
image.png
And here are Artefact2's thoughts on the matter: https://gist.github.com/Artefact2/b5f810600771265fc1e39442288e8ec9
See https://huggingface.co/mradermacher/model_requests for some answers to questions you might have and/or if you want some other model quantized.
I thank my company, nethype GmbH, for letting me use its servers and providing upgrades to my workstation to enable this work in my free time. Additional thanks to @nicoboss for giving me access to his private supercomputer, enabling me to provide many more imatrix quants, at much higher quality, than I would otherwise be able to.
Llama-3.3-70B-Instruct-SOM-MPOA.i1-IQ2_M.gguf
gguf · 22,5 GB · SHA-256 f21fb8b50651…eb1d · Hugging Face
DescargarLlama-3.3-70B-Instruct-SOM-MPOA.i1-IQ2_S.gguf
gguf · 20,7 GB · SHA-256 f0eb38b0493f…0694 · Hugging Face
DescargarLlama-3.3-70B-Instruct-SOM-MPOA.i1-IQ2_XS.gguf
gguf · 19,7 GB · SHA-256 90ba2dd2d9ac…3d48 · Hugging Face
DescargarLlama-3.3-70B-Instruct-SOM-MPOA.i1-IQ2_XXS.gguf
gguf · 17,8 GB · SHA-256 2fe6dca34582…3db6 · Hugging Face
DescargarLlama-3.3-70B-Instruct-SOM-MPOA.i1-IQ3_M.gguf
gguf · 29,7 GB · SHA-256 49102397f3e8…19c8 · Hugging Face
DescargarLlama-3.3-70B-Instruct-SOM-MPOA.i1-IQ3_S.gguf
gguf · 28,8 GB · SHA-256 639c734ce15d…fa48 · Hugging Face
DescargarLlama-3.3-70B-Instruct-SOM-MPOA.i1-IQ3_XS.gguf
gguf · 27,3 GB · SHA-256 5a3c0406398b…23f6 · Hugging Face
DescargarLlama-3.3-70B-Instruct-SOM-MPOA.i1-IQ3_XXS.gguf
gguf · 25,6 GB · SHA-256 9948e7a8b9a3…73e6 · Hugging Face
DescargarLlama-3.3-70B-Instruct-SOM-MPOA.i1-IQ4_XS.gguf
gguf · 35,3 GB · SHA-256 dbdf313928b8…392e · Hugging Face
DescargarLlama-3.3-70B-Instruct-SOM-MPOA.i1-Q2_K_S.gguf
gguf · 22,8 GB · SHA-256 caab4af93048…13d6 · Hugging Face
DescargarLlama-3.3-70B-Instruct-SOM-MPOA.i1-Q2_K.gguf
gguf · 24,6 GB · SHA-256 2678be5b3198…9804 · Hugging Face
DescargarLlama-3.3-70B-Instruct-SOM-MPOA.i1-Q3_K_L.gguf
gguf · 34,6 GB · SHA-256 5118be57a7a4…a315 · Hugging Face
DescargarLlama-3.3-70B-Instruct-SOM-MPOA.i1-Q3_K_M.gguf
gguf · 31,9 GB · SHA-256 d3de63f4602e…5e5d · Hugging Face
DescargarLlama-3.3-70B-Instruct-SOM-MPOA.i1-Q3_K_S.gguf
gguf · 28,8 GB · SHA-256 d1e977ec6b34…502a · Hugging Face
DescargarLlama-3.3-70B-Instruct-SOM-MPOA.i1-Q4_0.gguf
gguf · 37,4 GB · SHA-256 4a966f31f55d…a12a · Hugging Face
DescargarLlama-3.3-70B-Instruct-SOM-MPOA.i1-Q4_1.gguf
gguf · 41,3 GB · SHA-256 3643579f8a9b…3c53 · Hugging Face
DescargarLlama-3.3-70B-Instruct-SOM-MPOA.i1-Q4_K_M.gguf
gguf · 39,6 GB · SHA-256 1d9ab106dabd…d367 · Hugging Face
DescargarLlama-3.3-70B-Instruct-SOM-MPOA.i1-Q4_K_S.gguf
gguf · 37,6 GB · SHA-256 4c21286a3355…ffeb · Hugging Face
DescargarLlama-3.3-70B-Instruct-SOM-MPOA.i1-Q5_K_M.gguf
gguf · 46,5 GB · SHA-256 d972aa155ebf…2375 · Hugging Face
DescargarLlama-3.3-70B-Instruct-SOM-MPOA.i1-Q5_K_S.gguf
gguf · 45,3 GB · SHA-256 a8464ef021a4…2085 · Hugging Face
DescargarLlama-3.3-70B-Instruct-SOM-MPOA.i1-Q6_K.gguf
gguf · 53,9 GB · SHA-256 0e9093c0c6a8…09a9 · Hugging Face
DescargarLlama-3.3-70B-Instruct-SOM-MPOA.imatrix.gguf
gguf · 23,8 MB · SHA-256 1a7d7a332f6d…6765 · Hugging Face
Descargar--- base_model: 0xA50C1A1/Llama-3.3-70B-Instruct-SOM-MPOA language: - en library_name: transformers license: llama3.3 mradermacher: readme_rev: 1 quantized_by: mradermacher tags: - llama-3 - llama - meta - facebook - unsloth - transformers - pytorch - heretic - uncensored - decensored - abliterated --- ## About <!-- ### quantize_version: 2 --> <!-- ### output_tensor_quantised: 1 --> <!-- ### convert_type: hf --> <!-- ### vocab_type: --> <!-- ### tags: nicoboss --> <!-- ### quants: Q2_K IQ3_M Q4_K_S IQ3_XXS Q3_K_M small-IQ4_NL Q4_K_M IQ2_M Q6_K IQ4_XS Q2_K_S IQ1_M Q3_K_S IQ2_XXS Q3_K_L IQ2_XS Q5_K_S IQ2_S IQ1_S Q5_K_M Q4_0 IQ3_XS Q4_1 IQ3_S --> <!-- ### quants_skip: --> <!-- ### skip_mmproj: --> weighted/imatrix quants of https://huggingface.co/0xA50C1A1/Llama-3.3-70B-Instruct-SOM-MPOA <!-- provided-files --> ***For a convenient overview and download list, visit our [model page for this model](https://hf.tst.eu/model#Llama-3.3-70B-Instruct-SOM-MPOA-i1-GGUF).*** static quants are available at https://huggingface.co/mradermacher/Llama-3.3-70B-Instruct-SOM-MPOA-GGUF ## Usage If you are unsure how to use GGUF files, refer to one of [TheBloke's READMEs](https://huggingface.co/TheBloke/KafkaLM-70B-German-V0.1-GGUF) for more details, including on how to concatenate multi-part files. ## Provided Quants (sorted by size, not necessarily quality. IQ-quants are often preferable over similar sized non-IQ quants) | Link | Type | Size/GB | Notes | |:-----|:-----|--------:|:------| | [GGUF](https://huggingface.co/mradermacher/Llama-3.3-70B-Instruct-SOM-MPOA-i1-GGUF/resolve/main/Llama-3.3-70B-Instruct-SOM-MPOA.imatrix.gguf) | imatrix | 0.1 | imatrix file (for creating your own quants) | | [GGUF](https://huggingface.co/mradermacher/Llama-3.3-70B-Instruct-SOM-MPOA-i1-GGUF/resolve/main/Llama-3.3-70B-Instruct-SOM-MPOA.i1-IQ1_S.gguf) | i1-IQ1_S | 15.4 | for the desperate | | [GGUF](https://huggingface.co/mradermacher/Llama-3.3-70B-Instruct-SOM-MPOA-i1-GGUF/resolve/main/Llama-3.3-70B-Instruct-SOM-MPOA.i1-IQ1_M.gguf) | i1-IQ1_M | 16.9 | mostly desperate | | [GGUF](https://huggingface.co/mradermacher/Llama-3.3-70B-Instruct-SOM-MPOA-i1-GGUF/resolve/main/Llama-3.3-70B-Instruct-SOM-MPOA.i1-IQ2_XXS.gguf) | i1-IQ2_XXS | 19.2 | | | [GGUF](https://huggingface.co/mradermacher/Llama-3.3-70B-Instruct-SOM-MPOA-i1-GGUF/resolve/main/Llama-3.3-70B-Instruct-SOM-MPOA.i1-IQ2_XS.gguf) | i1...
Source context: 83 downloads · 0 likes · Library transformers · Repo mradermacher/Llama-3.3-70B-Instruct-SOM-MPOA-i1-GGUF
Source context: 68 downloads · 0 likes · Library transformers · Repo mradermacher/Llama-3.3-70B-Instruct-SOM-MPOA-i1-GGUF
| GGUF | i1-IQ2_XS | 21.2 |
| GGUF | i1-IQ2_S | 22.3 |
| GGUF | i1-IQ2_M | 24.2 |
| GGUF | i1-Q2_K_S | 24.6 | very low quality |
| GGUF | i1-Q2_K | 26.5 | IQ3_XXS probably better |
| GGUF | i1-IQ3_XXS | 27.6 | lower quality |
| GGUF | i1-IQ3_XS | 29.4 |
| GGUF | i1-IQ3_S | 31.0 | beats Q3_K* |
| GGUF | i1-Q3_K_S | 31.0 | IQ3_XS probably better |
| GGUF | i1-IQ3_M | 32.0 |
| GGUF | i1-Q3_K_M | 34.4 | IQ3_S probably better |
| GGUF | i1-Q3_K_L | 37.2 | IQ3_M probably better |
| GGUF | i1-IQ4_XS | 38.0 |
| GGUF | i1-Q4_0 | 40.2 | fast, low quality |
| GGUF | i1-Q4_K_S | 40.4 | optimal size/speed/quality |
| GGUF | i1-Q4_K_M | 42.6 | fast, recommended |
| GGUF | i1-Q4_1 | 44.4 |
| GGUF | i1-Q5_K_S | 48.8 |
| GGUF | i1-Q5_K_M | 50.0 |
| GGUF | i1-Q6_K | 58.0 | practically like static Q6_K |