Mozilla's Common Voice initiative has released a new dataset with 16 new languages and 4,622 new hours of speech, aiming to make voice technology more inclusive. The project relies on contributors donating speech data to a public dataset, which can be used to train voice-enabled technology. The new release includes languages such as Basaa, Kazakh, and Guarani, with English being the top language by total hours. The project has been made possible through a partnership with NVIDIA and a $3.4 million investment from the Gates Foundation, Giz, and FCDO.