artefact2/cat-8x7b-gguf Q5_K_S GGUF - Free GGUF Download is indexed on GraySoft with repository links, GGUF quant files, and Hugging Face metadata. This page helps you pick a local model for guIDE or other runtimes. See related models in the same shard below.
Model Intelligence Sheet
artefact2/cat-8x7b-gguf overview
These are GGUF quantized versions of Envoid/Cat-8x7B. The importance matrix was trained for 100K tokens (200 batches of 512 tokens) using wiki.train.raw. The IQ2XXS and IQ2XS versions are compatible with llama.cpp, version 147b17a or later. The IQ3XXS requires version f4d7e54 or later. Some model files above 50GB are split into smaller files. To concatenate them, use the cat command (on Windows, use PowerShell): cat foo-Q6K.gguf.* foo-Q6_K.gguf
Downloads
97
Likes
1
Pipeline
—
Library
—
Visibility
Public
Access
Open
Repository Files & Downloads
12 files detected
Direct downloads for all repository files
| File | Type | Quantization | Size | Link |
|---|---|---|---|---|
| Cat-8x7B-IQ2_XS.gguf | GGUF | IQ2_XS | 12.73 GB | Download |
| Cat-8x7B-IQ2_XXS.gguf | GGUF | IQ2_XXS | 11.44 GB | Download |
| Cat-8x7B-IQ3_XXS.gguf | GGUF | IQ3_XXS | 17.05 GB | Download |
| Cat-8x7B-Q2_K.gguf | GGUF | Q2_K | 16.12 GB | Download |
| Cat-8x7B-Q2_K_S.gguf | GGUF | Q2_K_S | 14.93 GB | Download |
| Cat-8x7B-Q3_K_L.gguf | GGUF | Q3_K_L | 22.51 GB | Download |
| Cat-8x7B-Q3_K_M.gguf | GGUF | Q3_K_M | 21.00 GB | Download |
| Cat-8x7B-Q4_K_M.gguf | GGUF | Q4_K_M | 26.49 GB | Download |
| Cat-8x7B-Q4_K_S.gguf | GGUF | Q4_K_S | 24.91 GB | Download |
| Cat-8x7B-Q5_K_M.gguf | GGUF | Q5_K_M | 30.95 GB | Download |
| Cat-8x7B-Q5_K_S.gguf | GGUF | Q5_K_S | 30.02 GB | Download |
| Cat-8x7B-Q6_K.gguf | GGUF | Q6_K | 35.74 GB | Download |
Model Details Live
Metadata Inspector
Normalized metadata (stored in metadata_json)
{
"metadata": {},
"card_data": {
"language": [
"en"
],
"license": "cc-by-nc-4.0",
"tags": [
"not-for-all-audiences"
],
"frontmatter": {
"language": [
"en"
],
"license": "cc-by-nc-4.0",
"tags": [
"not-for-all-audiences"
]
},
"hero_image_url": "",
"summary": "These are GGUF quantized versions of Envoid/Cat-8x7B. The importance matrix was trained for 100K tokens (200 batches of 512 tokens) using wiki.train.raw. The IQ2_XXS and IQ2_XS versions are compatible with llama.cpp, version 147b17a or later. The IQ3_XXS requires version f4d7e54 or later. Some model files above 50GB are split into smaller files. To concatenate them, use the cat command (on Windows, use PowerShell): cat foo-Q6_K.gguf.* > foo-Q6_K.gguf",
"quick_links": [],
"benchmark_table_html": "",
"readme_markdown": "---\nlanguage:\n- en\nlicense: cc-by-nc-4.0\ntags:\n- not-for-all-audiences\n---\n\nThese are GGUF quantized versions of [Envoid/Cat-8x7B](https://huggingface.co/Envoid/Cat-8x7B).\n\nThe importance matrix was trained for 100K tokens (200 batches of 512 tokens) using `wiki.train.raw`.\n\nThe IQ2_XXS and IQ2_XS versions are compatible with llama.cpp, version `147b17a` or later. The IQ3_XXS requires version `f4d7e54` or later.\n\nSome model files above 50GB are split into smaller files. To concatenate them, use the `cat` command (on Windows, use PowerShell): `cat foo-Q6_K.gguf.* > foo-Q6_K.gguf`",
"related_quantizations": []
},
"tags": [
"gguf",
"not-for-all-audiences",
"en",
"license:cc-by-nc-4.0",
"endpoints_compatible",
"region:us",
"conversational"
],
"likes": 1,
"downloads": 97,
"gated": false,
"private": false,
"last_modified": "2024-03-11T11:47:35.000Z",
"created_at": "2024-02-13T23:40:37.000Z",
"pipeline_tag": "",
"library_name": ""
}
Source payload excerpt (from Hugging Face API)
{
"_id": "65cbfdf53431b27016eafb0b",
"id": "Artefact2/Cat-8x7B-GGUF",
"modelId": "Artefact2/Cat-8x7B-GGUF",
"sha": "8f630185fe113f0955dc912ca99421f0a08849db",
"createdAt": "2024-02-13T23:40:37.000Z",
"lastModified": "2024-03-11T11:47:35.000Z",
"author": "Artefact2",
"downloads": 97,
"likes": 1,
"gated": false,
"private": false,
"pipeline_tag": "",
"library_name": "",
"siblings_count": 15
}