The article discusses the concept of tokens per second in the context of language models and provides a tool to visualize and understand the speed of token generation. It highlights the difference in perceived speed between code and text, as well as the impact of context size on token generation. The tool allows users to experiment with different speeds and modes, including code, text, and reasoning models. This helps to internalize the meaning of tokens per second and understand the limitations of language models.