LLM Comparison

DiffusionGemma vs MiniCPM

Side-by-side specs, pricing & capabilities · Updated September 2026

Add to comparison

2/6 models
Same tier:
 GoogleDiffusionGemmaOpenBMBMiniCPM
OrganizationGoogleGoogleOpenBMBOpenBMB
OpenTools Score
29
FamilyGemmaMiniCPM
StatusCurrentCurrent
Release DateJun 2026—
Context Window256K tokens128K tokens
Input Price——
Output Price——
Pricing NotesGoogle publishes open weights for local deployment. A hosted API price has not been verified; hosting and infrastructure costs depend on the provider or deployment.Open-weight GitHub and Hugging Face model family. There is no fixed vendor API price; runtime cost depends on the host, hardware, or inference provider.
Capabilities
textvisioncodereasoninglocal-inference
textcodereasoninglocal-inference
Training Cutoff—Not publicly specified in queued source
Max Output—33K tokens
API Identifiergoogle/diffusiongemma-26b-a4b-itOpenBMB/MiniCPM
Benchmarks
MMLU Pro
77.6official-google-model-card
—
GPQA Diamond
73.2official-google-model-card
—
LiveCodeBench v6
69.1official-google-model-card
—
MMMLU
81.5official-google-model-card
—
HLE no tools
11official-google-model-card
—
MiniCPM-SALA standard benchmark average—
76.53official-github-readme
MiniCPM-SALA long-context average—
38.97official-github-readme
MiniCPM-SALA 2048K extrapolation score—
81.6official-github-readme
MiniCPM4.1 reasoning decoding speedup—
3official-github-readme
MiniCPM4 Jetson AGX Orin decoding speedup vs Qwen3-8B—
7official-github-readme
 View DiffusionGemmaView MiniCPM

Looking for more AI models?

Browse All LLMs