GraySoft
Projects Models Compare Cloud benchmarks FAQ Download guIDE →
Model Intelligence Sheet

mondk/MiniCPM5-2B-IQ4_XS-GGUF overview

Model Information MiniCPM5 2B has the following features: Type : Causal Language Model Architecture : Standard LlamaForCausalLM Number of Parameters : 2,516,75…

transformersggufminicpmminicpm5llamatext-generationlong-contexttool-callingon-deviceedge-aillama-cppgguf-my-repoenzhdataset:openbmb/Ultra-FineWebdataset:openbmb/UltraX-Previewdataset:openbmb/Ultra-FineWeb-L3dataset:openbmb/UltraData-Mathdataset:openbmb/UltraData-Codedataset:openbmb/UltraData-SFT-2605dataset:openbmb/UltraData-SFT-Agent-2609dataset:openbmb/UltraData-RL-2609base_model:openbmb/MiniCPM5-2Bbase_model:quantized:openbmb/MiniCPM5-2B

Runs locally from ~1.33 GB disk (4 GB VRAM class GPUs with llama.cpp / guIDE).

Downloads
0
Likes
0
Pipeline
text-generation
Author

Repository Files & Downloads

1 GGUF files detected
Direct downloads for local inference
FileTypeQuantizationSizeLink
minicpm5-2b-IQ4_XS.ggufGGUFIQ4_XS1.33 GBDownload

Model Details

Model IDmondk/MiniCPM5-2B-IQ4_XS-GGUF
Authormondk
Pipelinetext-generation
Licenseapache-2.0
Base modelopenbmb/MiniCPM5-2B
Last modified2026-09-08T11:03:10.000Z

Model README

---

license: apache-2.0

language:

  • en
  • zh

library_name: transformers

pipeline_tag: text-generation

tags:

  • minicpm
  • minicpm5
  • llama
  • text-generation
  • long-context
  • tool-calling
  • on-device
  • edge-ai
  • llama-cpp
  • gguf-my-repo

datasets:

  • openbmb/Ultra-FineWeb
  • openbmb/UltraX-Preview
  • openbmb/Ultra-FineWeb-L3
  • openbmb/UltraData-Math
  • openbmb/UltraData-Code
  • openbmb/UltraData-SFT-2605
  • openbmb/UltraData-SFT-Agent-2609
  • openbmb/UltraData-RL-2609

base_model: openbmb/MiniCPM5-2B

---

Model Information

MiniCPM5-2B has the following features:

  • Type: Causal Language Model
  • Architecture: Standard LlamaForCausalLM
  • Number of Parameters: 2,516,756,480
  • Number of Non-Embedding Parameters: 1,981,982,720
  • Number of Layers: 42
  • Number of Attention Heads (GQA): 16 for Q and 2 for KV
  • Context Length: 131,072

Limitations and Disclaimer

This model has no autonomous intent or legal personhood; its outputs are text generated from statistical patterns and may be inaccurate, biased, or offensive, and may be manipulated by carefully crafted prompts ("jailbreaks") into producing unintended content. Its responses on sensitive topics such as politics, health, finance, and law are not reviewed by experts and should not be treated as professional advice.

This model is provided "AS IS", without warranty of any kind, express or implied, and the developers are not liable for any damages arising from its use. Users must employ the model only for lawful, compliant, and ethical purposes, configure their own safeguards, and label AI-generated content where required; deliberate jailbreaking, injection attacks, or inducing harmful output is prohibited, and any such testing is at the user's own risk.

License

This repository and MiniCPM model weights are released under the Apache-2.0 License.

Run mondk/MiniCPM5-2B-IQ4_XS-GGUF with guIDE

Download guIDE — the AI-native code editor with local LLM inference and 69 built-in tools.

Download guIDE → · Browse 524k+ models · Compare models

Source: Hugging Face · Compare models