LLM Comparison
DeepSeek V4 Flash 0731 vs Maple-Preview
Side-by-side specs, pricing & capabilities · Updated August 2026
Price vs Intelligence
Add to comparison
2/6 modelsSame tier:
| Organization | ||
| OpenTools Score | 46 219 | 50 |
| Family | DeepSeek V4 Flash | Maple |
| Status | Current | Current |
| Release Date | Jul 2026 | Aug 2026 |
| Context Window | 1.0M tokens | 131K tokens |
| Input Price | $0.14/M tokens | Free |
| Output Price | $0.28/M tokens | Free |
| Pricing Notes | DeepSeek API pricing lists $0.0028 cache-hit input, $0.14 cache-miss input, and $0.28 output per 1M tokens for deepseek-v4-flash; peak pricing may change after a future official announcement. | Open-source model card; no hosted API pricing found in the official source. Local inference costs depend on user hardware and runtime. |
| Capabilities | reasoningcodingtool-usestructured-outputlong-contextresponses-api | textreasoningcodeon-device-inference |
| Max Output | 393K tokens | 131K tokens |
| API Identifier | deepseek-v4-flash | deepgrove/maple-preview |
| Benchmarks | ||
| Artificial Analysis Intelligence Index v4.1 | 50artificial-analysis | — |
| Artificial Analysis Intelligence Index v4.1.1 | 52artificial-analysis | — |
| LiveCodeBench v6 | — | 75.1DeepGrove model card |
| AIME 2026 | — | 87.5DeepGrove model card |
| HMMT 2026 | — | 78.8DeepGrove model card |
| GPQA Diamond | — | 73.5DeepGrove model card |
| View DeepSeek V4 Flash 0731 | View Maple-Preview | |
Cost Calculator
Enter your expected monthly token usage to compare costs.
| Model | Input | Output | Total / mo | vs Best |
|---|---|---|---|---|
| Maple-PreviewCheapest | $0.00 | $0.00 | $0.00 | — |
| DeepSeek V4 Flash 0731 | $0.14 | $0.14 | $0.28 | +0% |
DeepSeek
DeepSeek V4 Flash 0731
DeepSeek V4 Flash 0731 is an open-weight, cost-efficient reasoning model for long-context coding, agent tasks, and API workflows through DeepSeek.
DeepGrove
Maple-Preview
Maple-Preview is DeepGrove’s open-source 20B-A1B ternary-weight reasoning LLM for efficient on-device inference. The model card reports 20.2B total parameters, 1.49B active parameters, a 131,072-token context window, and 5.31 GB checkpoint size.
More Comparisons
Looking for more AI models?
Browse All LLMs