Back
Ssignal
17
·3 days ago·1 signals
Archived topic · source no longer tracked

A llama.cpp PR makes Q2_0 3.0–3.6x faster on x86 CPUs, 8B decode goes 2.39 → 8.20 tok/s

Llama

Heat trend

Collecting trend data

The percentage is based on available heat signal, not comment count or independent people.