Back
Ssignal
16
·6 days ago·1 signals
Archived topic · source no longer tracked

A llama.cpp PR caches “hot” MoE experts on the GPU — 33 → 56 tok/s reported with 8GB VRAM

Heat trend

Collecting trend data

The percentage is based on available heat signal, not comment count or independent people.