GraySoft
Projects Models About FAQ Contact Download guIDE →

inferenceillusionist/l3-70b-euryale-v2.1-imat-gguf IQ3_XXS GGUF - Free GGUF Download is indexed on GraySoft with repository links, GGUF quant files, and Hugging Face metadata. This page helps you pick a local model for guIDE or other runtimes. See related models in the same shard below.

Model Intelligence Sheet

inferenceillusionist/l3-70b-euryale-v2.1-imat-gguf overview

Quantized from fp16. * Weighted quantizations were creating using fp16 GGUF and groupsmerged.txt in 88 chunks and nctx=512

transformersggufiMatllama3enbase_model:Sao10K/L3-70B-Euryale-v2.1base_model:quantized:Sao10K/L3-70B-Euryale-v2.1license:cc-by-nc-4.0endpoints_compatibleregion:usimatrixconversational
inferenceillusionist/l3-70b-euryale-v2.1-imat-gguf visual
Downloads
247
Likes
2
Pipeline
Library
transformers
Visibility
Public
Access
Open

Repository Files & Downloads

22 files detected
Direct downloads for all repository files
FileTypeQuantizationSizeLink
L3-70B-Euryale-v2.1-iMat-IQ1_M.gguf GGUF IQ1_M 15.60 GB Download
L3-70B-Euryale-v2.1-iMat-IQ1_S.gguf GGUF IQ1_S 14.29 GB Download
L3-70B-Euryale-v2.1-iMat-IQ2_M.gguf GGUF IQ2_M 22.46 GB Download
L3-70B-Euryale-v2.1-iMat-IQ2_S.gguf GGUF IQ2_S 20.71 GB Download
L3-70B-Euryale-v2.1-iMat-IQ2_XS.gguf GGUF IQ2_XS 19.69 GB Download
L3-70B-Euryale-v2.1-iMat-IQ2_XXS.gguf GGUF IQ2_XXS 17.79 GB Download
L3-70B-Euryale-v2.1-iMat-IQ3_M.gguf GGUF IQ3_M 29.74 GB Download
L3-70B-Euryale-v2.1-iMat-IQ3_S.gguf GGUF IQ3_S 28.79 GB Download
L3-70B-Euryale-v2.1-iMat-IQ3_XS.gguf GGUF IQ3_XS 27.29 GB Download
L3-70B-Euryale-v2.1-iMat-IQ3_XXS.gguf GGUF IQ3_XXS 25.58 GB Download
L3-70B-Euryale-v2.1-iMat-IQ4_XS.gguf GGUF IQ4_XS 35.30 GB Download
L3-70B-Euryale-v2.1-iMat-Q2_K.gguf GGUF Q2_K 24.56 GB Download
L3-70B-Euryale-v2.1-iMat-Q3_K_M.gguf GGUF Q3_K_M 31.91 GB Download
L3-70B-Euryale-v2.1-iMat-Q3_K_S.gguf GGUF Q3_K_S 28.79 GB Download
L3-70B-Euryale-v2.1-iMat-Q4_K_M.gguf GGUF Q4_K_M 39.60 GB Download
L3-70B-Euryale-v2.1-iMat-Q4_K_S.gguf GGUF Q4_K_S 37.58 GB Download
L3-70B-Euryale-v2.1-iMat-Q5_K_M.gguf GGUF Q5_K_M 46.52 GB Download
L3-70B-Euryale-v2.1-iMat-Q5_K_S.gguf GGUF Q5_K_S 45.32 GB Download
L3-70B-Euryale-v2.1-iMat-Q6_K-00001-of-00002.gguf GGUF Q6_K 44.91 GB Download
L3-70B-Euryale-v2.1-iMat-Q6_K-00002-of-00002.gguf GGUF Q6_K 9.01 GB Download
L3-70B-Euryale-v2.1-iMat-Q8_0-00001-of-00002.gguf GGUF 43.77 GB Download
L3-70B-Euryale-v2.1-iMat-Q8_0-00002-of-00002.gguf GGUF 26.06 GB Download

Model Details Live

Model Slug
inferenceillusionist/l3-70b-euryale-v2.1-imat-gguf
Author
InferenceIllusionist
Pipeline Task
Library
transformers
Created
2024-06-13
Last Modified
2024-06-15
Gated
No
Private
No
HF SHA
2cfd4681ff16a31e4b4375bb1b9286a314f692c4
License
cc-by-nc-4.0
Language
en
Base Model
Sao10K/L3-70B-Euryale-v2.1

Metadata Inspector

Normalized metadata (stored in metadata_json)
{
  "metadata": {},
  "card_data": {
    "base_model": "Sao10K/L3-70B-Euryale-v2.1",
    "language": [
      "en"
    ],
    "library_name": "transformers",
    "quantized_by": "InferenceIllusionist",
    "tags": [
      "iMat",
      "gguf",
      "llama3"
    ],
    "license": "cc-by-nc-4.0",
    "frontmatter": {
      "base_model": "Sao10K/L3-70B-Euryale-v2.1",
      "language": [
        "en"
      ],
      "library_name": "transformers",
      "quantized_by": "InferenceIllusionist",
      "tags": [
        "iMat",
        "gguf",
        "llama3"
      ],
      "license": "cc-by-nc-4.0"
    },
    "hero_image_url": "https://i.imgur.com/P68dXux.png",
    "summary": "Quantized from fp16. * Weighted quantizations were creating using fp16 GGUF and groups_merged.txt in 88 chunks and n_ctx=512",
    "quick_links": [],
    "benchmark_table_html": "",
    "readme_markdown": "---\nbase_model: Sao10K/L3-70B-Euryale-v2.1\nlanguage:\n- en\nlibrary_name: transformers\nquantized_by: InferenceIllusionist\ntags:\n- iMat\n- gguf\n- llama3\nlicense: cc-by-nc-4.0\n---\n<img src=\"https://i.imgur.com/P68dXux.png\" width=\"400\"/>\n\n# L3-70B-Euryale-v2.1-iMat-GGUF\n\nQuantized from fp16.\n* Weighted quantizations were creating using fp16 GGUF and groups_merged.txt in 88 chunks and n_ctx=512\n\n## Recommended Sampler Settings (From Original Model Card)\n```\nTemperature - 1.17\nmin_p - 0.075\nRepetition Penalty - 1.10\n```\n\n**SillyTavern Instruct Settings**:\n<br>Context Template: Llama-3-Instruct-Names\n<br>Instruct Presets: [Euryale-v2.1-Llama-3-Instruct](https://huggingface.co/Sao10K/L3-70B-Euryale-v2.1/blob/main/Euryale-v2.1-Llama-3-Instruct.json)\n\n\nFor a brief rundown of iMatrix quant performance please see this [PR](https://github.com/ggerganov/llama.cpp/pull/5747)\n\n<i>All quants are verified working prior to uploading to repo for your safety and convenience. </i>\n\n\n<b>Tip:</b> Pick a file size under your GPU's VRAM while still allowing some room for context for best speed. You may need to pad this further depending on if you are running image gen or TTS as well.\n\nOriginal model card can be found [here](https://huggingface.co/Sao10K/L3-70B-Euryale-v2.1)",
    "related_quantizations": []
  },
  "tags": [
    "transformers",
    "gguf",
    "iMat",
    "llama3",
    "en",
    "base_model:Sao10K/L3-70B-Euryale-v2.1",
    "base_model:quantized:Sao10K/L3-70B-Euryale-v2.1",
    "license:cc-by-nc-4.0",
    "endpoints_compatible",
    "region:us",
    "imatrix",
    "conversational"
  ],
  "likes": 2,
  "downloads": 247,
  "gated": false,
  "private": false,
  "last_modified": "2024-06-15T18:59:01.000Z",
  "created_at": "2024-06-13T18:01:24.000Z",
  "pipeline_tag": "",
  "library_name": "transformers"
}
Source payload excerpt (from Hugging Face API)
{
  "_id": "666b33f43214e94388c2b5d5",
  "id": "InferenceIllusionist/L3-70B-Euryale-v2.1-iMat-GGUF",
  "modelId": "InferenceIllusionist/L3-70B-Euryale-v2.1-iMat-GGUF",
  "sha": "2cfd4681ff16a31e4b4375bb1b9286a314f692c4",
  "createdAt": "2024-06-13T18:01:24.000Z",
  "lastModified": "2024-06-15T18:59:01.000Z",
  "author": "InferenceIllusionist",
  "downloads": 247,
  "likes": 2,
  "gated": false,
  "private": false,
  "pipeline_tag": "",
  "library_name": "transformers",
  "siblings_count": 24
}