Mistral AI has released Mixtral 8x7B, a high-quality sparse mixture of experts model with open weights, which outperforms Llama 2 70B on most benchmarks with 6x faster inference. Mixtral has 46.7B total parameters but only uses 12.9B parameters per token, making it more efficient. The model is pre-trained on data extracted from the open Web and can be fine-tuned for specific tasks. Mixtral is available for use, with Mistral AI providing an endpoint for testing.