The article discusses a new approach to training large language models (LLMs) that enables them to reason and think more like humans. This approach involves generating 'think out loud' tokens and using them to improve the model's performance. The model has achieved impressive results, including a high ELO rating and strong performance on various benchmarks. However, the significance and implications of these results are debated in the comments.