Skip to main content
Gemini 3.5 Flash is Google’s May 2026 Flash reasoning model for agentic workflows, coding, and multimodal understanding. It accepts text, image, audio, and video input with a 1M-token context window and returns up to 64K output tokens. Reasoning effort is tunable with reasoning_effort set to minimal, low, medium, high, or max.

Model

Supported Endpoints

All three API formats are supported:

Example

Volume discounts are available for almost all models. Reach out at contact@unifically.com to discuss custom pricing.