Model Intelligence Sheet
filvyb/llama-joycaption-beta-one-hf-llava-GGUF overview
Model Card for Llama JoyCaption Beta One Github https://github.com/fpgaminer/joycaption JoyCaption is an image captioning Visual Language Model VLM built from …
Runs locally from ~565.4 MB disk (4 GB VRAM class GPUs with llama.cpp / guIDE).
Repository Files & Downloads
4 GGUF files detected
Direct downloads for local inference
Model Details
| Model ID | filvyb/llama-joycaption-beta-one-hf-llava-GGUF |
|---|---|
| Author | filvyb |
| Pipeline | image-text-to-text |
| License | llama3.1 |
| Base model | fancyfeast/llama-joycaption-beta-one-hf-llava |
| Last modified | 2026-09-16T17:19:57.000Z |
Model README
---
license: llama3.1
base_model:
- fancyfeast/llama-joycaption-beta-one-hf-llava
tags:
- captioning
pipeline_tag: image-text-to-text
library_name: transformers
---
Model Card for Llama JoyCaption Beta One
JoyCaption is an image captioning Visual Language Model (VLM) built from the ground up as a free, open, and uncensored model for the community to use in training Diffusion models.
Key Features:
- Free and Open: Always released for free, open weights, no restrictions, and just like bigASP, will come with training scripts and lots of juicy details on how it gets built.
- Uncensored: Equal coverage of SFW and NSFW concepts. No "cylindrical shaped object with a white substance coming out on it" here.
- Diversity: All are welcome here. Do you like digital art? Photoreal? Anime? Furry? JoyCaption is for everyone. Pains are being taken to ensure broad coverage of image styles, content, ethnicity, gender, orientation, etc.
- Minimal Filtering: JoyCaption is trained on large swathes of images so that it can understand almost all aspects of our world. almost. Illegal content will never be tolerated in JoyCaption's training.
Run filvyb/llama-joycaption-beta-one-hf-llava-GGUF with guIDE
Download guIDE — the AI-native code editor with local LLM inference and 69 built-in tools.
Source: Hugging Face · Compare models