Skip to content
RCreddit.com·
Not on the current live radar

What TPS is too slow for you?

AI summary

The discussion revolves around acceptable Tokens Per Second (TPS) for running local models. Some users are content with sub-30 TPS for high-quality models, while others prioritize speed, aiming for 100-150 TPS or more, even if it means compromising quality. The core question is to define the absolute minimum TPS that users would tolerate for their local models.

Why this one

This discussion uniquely highlights the wide disparity in user expectations for local model performance, ranging from sub-30 TPS for quality to over 100-150 TPS for speed.

Time & source

Times shown in UTC

Display time zone: UTC

Local time zone unavailable; showing UTC.

IngestedOffset at this time: UTC+0Sep 24, 2026, 12:02 UTC

Ingested
Sep 24, 2026, 12:02
Source type
Dev community

Full text isn't available here.

Read at source →
Source·reddit.com