Google has released Gemini 3 Pro, a multimodal model that delivers state-of-the-art performance in document, spatial, screen, and video understanding. The model excels in complex visual reasoning, document processing, and understanding spatial relationships. Gemini 3 Pro represents a significant improvement over its predecessors and other models in the field. It can be used for tasks such as OCR, visual extraction, and causal logic.