The Qwen3.8 Max model has been ranked as the best overall model by the Artificial Analysis Agentic Index, a benchmark that evaluates AI models on their ability to perform real-world tasks. The index assesses models based on their performance in various evaluations, including GDPval-AA v2 and Tau³-Banking. The Qwen3.8 Max model achieved a score of 55.4 on the Agentic Index, outperforming other models. However, some commenters have pointed out that the model's performance may not be significantly better than other models, and that the cost of running the model may be a limiting factor.