inferenceillusionist/l3-70b-euryale-v2.1-imat-gguf IQ3_XXS GGUF - Free GGUF Download is indexed on GraySoft with repository links, GGUF quant files, and Hugging Face metadata. This page helps you pick a local model for guIDE or other runtimes. See related models in the same shard below.
Model Intelligence Sheet
inferenceillusionist/l3-70b-euryale-v2.1-imat-gguf overview
Quantized from fp16. * Weighted quantizations were creating using fp16 GGUF and groupsmerged.txt in 88 chunks and nctx=512
Downloads
247
Likes
2
Pipeline
—
Library
transformers
Visibility
Public
Access
Open
Repository Files & Downloads
22 files detected
Direct downloads for all repository files
| File | Type | Quantization | Size | Link |
|---|---|---|---|---|
| L3-70B-Euryale-v2.1-iMat-IQ1_M.gguf | GGUF | IQ1_M | 15.60 GB | Download |
| L3-70B-Euryale-v2.1-iMat-IQ1_S.gguf | GGUF | IQ1_S | 14.29 GB | Download |
| L3-70B-Euryale-v2.1-iMat-IQ2_M.gguf | GGUF | IQ2_M | 22.46 GB | Download |
| L3-70B-Euryale-v2.1-iMat-IQ2_S.gguf | GGUF | IQ2_S | 20.71 GB | Download |
| L3-70B-Euryale-v2.1-iMat-IQ2_XS.gguf | GGUF | IQ2_XS | 19.69 GB | Download |
| L3-70B-Euryale-v2.1-iMat-IQ2_XXS.gguf | GGUF | IQ2_XXS | 17.79 GB | Download |
| L3-70B-Euryale-v2.1-iMat-IQ3_M.gguf | GGUF | IQ3_M | 29.74 GB | Download |
| L3-70B-Euryale-v2.1-iMat-IQ3_S.gguf | GGUF | IQ3_S | 28.79 GB | Download |
| L3-70B-Euryale-v2.1-iMat-IQ3_XS.gguf | GGUF | IQ3_XS | 27.29 GB | Download |
| L3-70B-Euryale-v2.1-iMat-IQ3_XXS.gguf | GGUF | IQ3_XXS | 25.58 GB | Download |
| L3-70B-Euryale-v2.1-iMat-IQ4_XS.gguf | GGUF | IQ4_XS | 35.30 GB | Download |
| L3-70B-Euryale-v2.1-iMat-Q2_K.gguf | GGUF | Q2_K | 24.56 GB | Download |
| L3-70B-Euryale-v2.1-iMat-Q3_K_M.gguf | GGUF | Q3_K_M | 31.91 GB | Download |
| L3-70B-Euryale-v2.1-iMat-Q3_K_S.gguf | GGUF | Q3_K_S | 28.79 GB | Download |
| L3-70B-Euryale-v2.1-iMat-Q4_K_M.gguf | GGUF | Q4_K_M | 39.60 GB | Download |
| L3-70B-Euryale-v2.1-iMat-Q4_K_S.gguf | GGUF | Q4_K_S | 37.58 GB | Download |
| L3-70B-Euryale-v2.1-iMat-Q5_K_M.gguf | GGUF | Q5_K_M | 46.52 GB | Download |
| L3-70B-Euryale-v2.1-iMat-Q5_K_S.gguf | GGUF | Q5_K_S | 45.32 GB | Download |
| L3-70B-Euryale-v2.1-iMat-Q6_K-00001-of-00002.gguf | GGUF | Q6_K | 44.91 GB | Download |
| L3-70B-Euryale-v2.1-iMat-Q6_K-00002-of-00002.gguf | GGUF | Q6_K | 9.01 GB | Download |
| L3-70B-Euryale-v2.1-iMat-Q8_0-00001-of-00002.gguf | GGUF | — | 43.77 GB | Download |
| L3-70B-Euryale-v2.1-iMat-Q8_0-00002-of-00002.gguf | GGUF | — | 26.06 GB | Download |
Model Details Live
Metadata Inspector
Normalized metadata (stored in metadata_json)
{
"metadata": {},
"card_data": {
"base_model": "Sao10K/L3-70B-Euryale-v2.1",
"language": [
"en"
],
"library_name": "transformers",
"quantized_by": "InferenceIllusionist",
"tags": [
"iMat",
"gguf",
"llama3"
],
"license": "cc-by-nc-4.0",
"frontmatter": {
"base_model": "Sao10K/L3-70B-Euryale-v2.1",
"language": [
"en"
],
"library_name": "transformers",
"quantized_by": "InferenceIllusionist",
"tags": [
"iMat",
"gguf",
"llama3"
],
"license": "cc-by-nc-4.0"
},
"hero_image_url": "https://i.imgur.com/P68dXux.png",
"summary": "Quantized from fp16. * Weighted quantizations were creating using fp16 GGUF and groups_merged.txt in 88 chunks and n_ctx=512",
"quick_links": [],
"benchmark_table_html": "",
"readme_markdown": "---\nbase_model: Sao10K/L3-70B-Euryale-v2.1\nlanguage:\n- en\nlibrary_name: transformers\nquantized_by: InferenceIllusionist\ntags:\n- iMat\n- gguf\n- llama3\nlicense: cc-by-nc-4.0\n---\n<img src=\"https://i.imgur.com/P68dXux.png\" width=\"400\"/>\n\n# L3-70B-Euryale-v2.1-iMat-GGUF\n\nQuantized from fp16.\n* Weighted quantizations were creating using fp16 GGUF and groups_merged.txt in 88 chunks and n_ctx=512\n\n## Recommended Sampler Settings (From Original Model Card)\n```\nTemperature - 1.17\nmin_p - 0.075\nRepetition Penalty - 1.10\n```\n\n**SillyTavern Instruct Settings**:\n<br>Context Template: Llama-3-Instruct-Names\n<br>Instruct Presets: [Euryale-v2.1-Llama-3-Instruct](https://huggingface.co/Sao10K/L3-70B-Euryale-v2.1/blob/main/Euryale-v2.1-Llama-3-Instruct.json)\n\n\nFor a brief rundown of iMatrix quant performance please see this [PR](https://github.com/ggerganov/llama.cpp/pull/5747)\n\n<i>All quants are verified working prior to uploading to repo for your safety and convenience. </i>\n\n\n<b>Tip:</b> Pick a file size under your GPU's VRAM while still allowing some room for context for best speed. You may need to pad this further depending on if you are running image gen or TTS as well.\n\nOriginal model card can be found [here](https://huggingface.co/Sao10K/L3-70B-Euryale-v2.1)",
"related_quantizations": []
},
"tags": [
"transformers",
"gguf",
"iMat",
"llama3",
"en",
"base_model:Sao10K/L3-70B-Euryale-v2.1",
"base_model:quantized:Sao10K/L3-70B-Euryale-v2.1",
"license:cc-by-nc-4.0",
"endpoints_compatible",
"region:us",
"imatrix",
"conversational"
],
"likes": 2,
"downloads": 247,
"gated": false,
"private": false,
"last_modified": "2024-06-15T18:59:01.000Z",
"created_at": "2024-06-13T18:01:24.000Z",
"pipeline_tag": "",
"library_name": "transformers"
}
Source payload excerpt (from Hugging Face API)
{
"_id": "666b33f43214e94388c2b5d5",
"id": "InferenceIllusionist/L3-70B-Euryale-v2.1-iMat-GGUF",
"modelId": "InferenceIllusionist/L3-70B-Euryale-v2.1-iMat-GGUF",
"sha": "2cfd4681ff16a31e4b4375bb1b9286a314f692c4",
"createdAt": "2024-06-13T18:01:24.000Z",
"lastModified": "2024-06-15T18:59:01.000Z",
"author": "InferenceIllusionist",
"downloads": 247,
"likes": 2,
"gated": false,
"private": false,
"pipeline_tag": "",
"library_name": "transformers",
"siblings_count": 24
}