Skip to content
RCreddit.com·
Not on the current live radar

llama: add Maple 20B-A1B ternary MoE architecture (CPU) by AlexGabbia · Pull Request #27000 · ggml-org/llama.cpp

AI summary

A new Maple 20B-A1B ternary MoE architecture for CPU is being added to llama.cpp, as seen in Pull Request #27000 by AlexGabbia. This development, which includes a preview available on Hugging Face at deepgrove/maple-preview, suggests a potential benefit for users with low VRAM.

Why this one

This pull request marks the first time a ternary MoE architecture, specifically the Maple 20B-A1B, is being integrated into llama.cpp for CPU, unlike previous additions focused on other architectures.

Time & source

Times shown in UTC

Display time zone: UTC

Local time zone unavailable; showing UTC.

IngestedOffset at this time: UTC+0Sep 14, 2026, 17:01 UTC

Ingested
Sep 14, 2026, 17:01
Source type
Dev community

Full text isn't available here.

Read at source →
Source·reddit.com