Ssignal
16
·6 days ago·1 signals
Archived topic · source no longer tracked
A llama.cpp PR caches “hot” MoE experts on the GPU — 33 → 56 tok/s reported with 8GB VRAM
Heat trend
Collecting trend data
The percentage is based on available heat signal, not comment count or independent people.