Issue 2026-09-15 · Industry · 开源 · 模型

llama.cpp adds Maple 20B-A1B ternary MoE (CPU)

A PR to ggml-org/llama.cpp adds the Maple 20B-A1B ternary MoE architecture by AlexGabbia, runnable on CPU and visible as deepgrove/maple-preview on Hugging Face. The change extends open-source options for users with low VRAM.

r/LocalLLaMA3 d ago
Read original ↗