LLM Comparison

GLM-5.3-Flash vs MiniCPM

Side-by-side specs, pricing & capabilities · Updated September 2026

Add to comparison

2/6 models
Same tier:
 Z.aiGLM-5.3-FlashOpenBMBMiniCPM
OrganizationZ.aiZ.aiOpenBMBOpenBMB
OpenTools Score
75
463
FamilyGLMMiniCPM
StatusCurrentCurrent
Release DateAug 2026
Context Window1.3M tokens128K tokens
Input Price$0.07/M tokensFree
Output Price$0.25/M tokensFree
Pricing NotesOpenRouter lists a limited-time discounted price of $0.075/M input and $0.25/M output. Standard first-party style pricing is commonly $0.15/M input and $0.50/M output; cache-read pricing may vary by provider.Open-weight GitHub and Hugging Face model family. There is no fixed vendor API price; runtime cost depends on the host, hardware, or inference provider.
Capabilities
textvisionvideo-inputcodetool-usereasoninglong-contextopen-weights
textcodereasoninglocal-inference
Training CutoffNot disclosedNot publicly specified in queued source
Max Output131K tokens33K tokens
API Identifierz-ai/glm-5.3-flashOpenBMB/MiniCPM
Benchmarks
Artificial Analysis Intelligence Index v4.1.1
57artificial-analysis
MMLU
88.1official
MiniCPM-SALA standard benchmark average
76.53official-github-readme
MiniCPM-SALA long-context average
38.97official-github-readme
MiniCPM-SALA 2048K extrapolation score
81.6official-github-readme
MiniCPM4.1 reasoning decoding speedup
3official-github-readme
MiniCPM4 Jetson AGX Orin decoding speedup vs Qwen3-8B
7official-github-readme
 View GLM-5.3-FlashView MiniCPM

Cost Calculator

Enter your expected monthly token usage to compare costs.

ModelInputOutputTotal / movs Best
MiniCPMCheapest$0.00$0.00$0.00
GLM-5.3-Flash$0.08$0.13$0.20+0%

Looking for more AI models?

Browse All LLMs