Xenna/stella-embed-0.6b-gguf overview
Stella Embed 0.6b q8 0 GGUF GGUF Q8 0 quantized version of Stella Embed 0.6b , a 600M parameter text embedding model from the StelNet ML V2 suite. Model Detail…
Runs locally from ~609.5 MB disk (4 GB VRAM class GPUs with llama.cpp / guIDE).
Repository Files & Downloads
| File | Type | Quantization | Size | Link |
|---|---|---|---|---|
| Stella-Embed-0.6b-q8_0.gguf | GGUF | Q8_0 | 609.5 MB | Download |
Model Details
Model README
---
license: apache-2.0
language: en
base_model: Xenna/stella-embed-0.6b
tags:
- embedding
- gguf
- stelnet
- stella
- q8_0
---
Stella-Embed-0.6b-q8_0-GGUF
GGUF Q8_0 quantized version of Stella-Embed-0.6b, a 600M parameter text embedding model from the StelNet ML V2 suite.
Model Details
| Key | Value |
|-----|-------|
| Model Name | Stella-Embed-0.6b-q8_0 |
| Base Model | Stella-Embed-0.6b |
| Quantization | GGUF Q8_0 |
| Parameters | ~600M |
| Framework | GGUF / llama.cpp |
| Organization | Xenna (StelNet) |
Use Cases
- Local text embedding inference via llama.cpp
- StelNet ML V2: Stella (DELTA tier) — security & support backbone
- Efficient CPU inference with Q8_0 quantization
Usage
# Pull via HF
hf download Xenna/stella-embed-0.6b-gguf
License
Apache 2.0
Run Xenna/stella-embed-0.6b-gguf with guIDE
Download guIDE — the AI-native code editor with local LLM inference and 69 built-in tools.
Source: Hugging Face · Compare models