dahara1 author hub
Gemma4は小さいモデルに先行させる事で推論を高速化するMTP Multi Token Prediction という仕組みがあります。 このモデルはdahara1/gemma 4 E4B it UD japanese imatrixをMTPで動かすための小モデルです。 Gemma4 has a mechanism called MTP Multi Token Prediction that speeds up inference by running a smaller model before the main model. This model …
Models
2
Downloads
0
Run models locally with guIDE
Download guIDE — the AI-native code editor with local LLM inference and 69 built-in tools.