Back
Ssignal
16
·4 days ago·1 signals
Archived topic · source no longer tracked

I ported vLLM's serving stack to C++20: 66 MiB binary, no Python at inference, output checked token-for-token against vLLM

Heat trend

Collecting trend data

The percentage is based on available heat signal, not comment count or independent people.