Tokens per Second

TPS

A performance metric measuring the rate at which an inference engine generates output tokens per second.

Explore more about Metrics

Signal, not noise.

Focused newsletter for builders and knowledge workers tracking how AI is changing real work. We surface what matters in practice, not every headline. Curated for practitioners, not spectators.