Compare models
Put up to 4 models beside each other — token prices, context windows, capabilities and provider, from the same catalogue the model pages read.
| Attribute | DeepSeek V4.1 Flashdeepseek-v4.1-flash | Gemini 3.1 Pro Previewgemini-3.1-pro-preview | Gemini 3.8 Flashgemini-3.8-flash |
|---|---|---|---|
| Pricing | |||
| Input | $0.30 / 1M | $2.00 / 1M | $0.75 / 1M |
| Output | $1.20 / 1M | $12.00 / 1M | $3.75 / 1M |
| Cache Write (5m) | $0.30 / 1M | $2.00 / 1M | $0.75 / 1M |
| Cache Write (1h) | $0.30 / 1M | $2.00 / 1M | $0.75 / 1M |
| Cache Read | $0.30 / 1M | $2.00 / 1M | $0.75 / 1M |
| Web Search | $0 / 1M | $0 / 1M | $0 / 1M |
| Context | |||
| Max context | 1M | 1.0M | 1M |
| Max output | N/A | N/A | N/A |
| Capabilities | |||
| Vision | Yes | Yes | Yes |
| Function Calling | Yes | Yes | Yes |
| JSON Mode | Yes | Yes | Yes |
| Streaming | Yes | Yes | Yes |
| Catalogue | |||
| Provider | DeepSeek | ||
| Category | chat | chat | chat |
| Charge type | Pay As You Go | Pay As You Go | Pay As You Go |
| Released | — | — | — |
| Description | |||
| Summary | DeepSeek V4.1 Flash is a cost-efficient sparse Mixture-of-Experts (MoE) model in DeepSeek's V4.1 family, optimized for coding, reasoning, and agentic workflows. Despite its efficiency-focused positioning, DeepSeek reports that it surpasses the previous V4 Pro in performance, inference speed, and overall task completion time. The model is particularly strong at long-horizon, multi-step execution, making it well suited for coding agents, complex problem solving, and autonomous workflows that must reliably carry tasks through to completion. | Gemini 3.1 Pro Preview is Google's frontier reasoning model, delivering enhanced software engineering performance, improved agentic reliability, and more efficient token usage across complex workflows. Built on the multimodal foundation of the Gemini 3 series, it combines high-precision reasoning across text, image, video, audio, and code with a 1M-token context window for large-scale tasks. The 3.1 update introduces measurable gains on SWE benchmarks and real-world coding environments, along with stronger autonomous execution in structured domains such as finance and spreadsheet-based workflows. Designed for advanced development and agentic systems, it improves long-horizon stability and tool orchestration while adding a new medium thinking level to better balance cost, speed, and performance. Gemini 3.1 Pro Preview is well suited for agentic coding, structured planning, multimodal analysis, financial modeling, spreadsheet automation, and high-context enterprise applications. | Gemini 3.8 Flash is Google's most intelligent Flash-class model, delivering significant improvements over Gemini 3.7 Flash across software engineering, agentic workflows, and complex multi-step reasoning. Designed to combine strong capability with Flash-tier efficiency, it is well suited for coding assistants, autonomous agents, and high-throughput production workflows that require responsive performance without sacrificing reasoning quality. |