Run MiniCPM5-1B via Ollama on macOS / Linux laptop using the released GGUF. Use when the user wants "ollama run", "ollama pull", a Modelfile-driven setup, or one-line laptop deployment.
73
91%
Does it follow best practices?
Run evals on this skill
Adds up to 20 points to the overall score
View guide
Low
Low-risk findings worth noting
One-binary, no-Python laptop deployment. Consumes the released GGUF.
| Var | Example | Default |
|---|---|---|
GGUF_REPO | openbmb/MiniCPM5-1B-GGUF | required |
QUANT | Q4_K_M (657 MB, recommended) / Q8_0 / F16 | Q4_K_M |
MODEL_NAME | minicpm5-1b | minicpm5-1b |
brew install ollama # macOS
# or:
curl -fsSL https://ollama.com/install.sh | sh # Linux
OLLAMA_FLASH_ATTENTION=1 OLLAMA_KV_CACHE_TYPE=q8_0 ollama serve &mkdir -p ~/${MODEL_NAME} && cd ~/${MODEL_NAME}
huggingface-cli download ${GGUF_REPO} MiniCPM5-1B-${QUANT}.gguf --local-dir .
cat > Modelfile <<EOF
FROM ./MiniCPM5-1B-${QUANT}.gguf
# MiniCPM5 chat template (matches release tokenizer)
TEMPLATE """{{- if .Messages -}}
{{- range .Messages -}}
<|im_start|>{{ .Role }}
{{ .Content }}<|im_end|>
{{ end -}}
<|im_start|>assistant
{{ end -}}"""
PARAMETER stop "<|im_end|>"
PARAMETER stop "</s>"
# Defaults tuned for nothink mode
PARAMETER temperature 0.7
PARAMETER top_p 0.95
PARAMETER num_ctx 8192
EOFollama create ${MODEL_NAME} -f Modelfile
ollama run ${MODEL_NAME}curl http://localhost:11434/v1/chat/completions \
-H "Content-Type: application/json" \
-d '{
"model": "minicpm5-1b",
"messages": [{"role": "user", "content": "1+1=?"}],
"temperature": 0.7, "top_p": 0.95, "max_tokens": 64
}'Expected: "2" in the reply.
Default Modelfile is nothink. For think:
ollama run ${MODEL_NAME} --temperature 0.9 --top-p 0.95Or bake it into a separate model tag by flipping temperature 0.7 to temperature 0.9 (top_p stays 0.95) and ollama create ${MODEL_NAME}-think -f Modelfile.think.
To force the auto-injected <think>\n prefix, use raw mode:
curl http://localhost:11434/api/generate -d '{
"model": "minicpm5-1b",
"raw": true,
"prompt": "<|im_start|>user\n鸡兔同笼…<|im_end|>\n<|im_start|>assistant\n<think>\n",
"options": {"temperature": 0.9, "top_p": 0.95}
}'Error: invalid file magic: corrupted download. Re-run huggingface-cli download.minicpm5-deploy-mlx (Q4 build)minicpm5-deploy-lmstudiominicpm5-deploy-vllm719e4fc
If you maintain this skill, you can claim it as your own. Once claimed, you can manage eval scenarios, bundle related skills, attach documentation or rules, and ensure cross-agent compatibility.