Skip to main content
DeepSeek V4.1 Flash is DeepSeek’s September 2026 Flash model for agentic coding and long-context work, with a 1M-token context window, 384K max output, and thinking on by default. The deepseek/deepseek-flash ID always points at DeepSeek’s latest Flash model.

Model

Supported Endpoints

All three API formats are supported:
DeepSeek models bill at 2x during peak hours: Monday to Friday, 01:00–04:00 and 06:00–10:00 UTC. The rate is fixed when the request starts. See pricing for current rates.

Example

Volume discounts are available for almost all models. Reach out at contact@unifically.com to discuss custom pricing.