Xiaomi has released the MiMo-V2.5-Pro-UltraSpeed model, which achieves a generation speed of 1000 tokens per second on a 1-trillion-parameter model. This is made possible through a combination of model-system codesign, FP4 quantization, and DFlash speculative decoding. The model is available through a limited-time application-based trial. The achievement has implications for real-time decision-making and productivity.