AI Models & LLMs
Browse and compare AI models from leading organizations. Find the right model for your use case based on pricing, capabilities, and context window.29 models listed
Intelligence vs. Cost
Average benchmark score plotted against cost per million tokens. Dot size reflects context window.
OrganizationAllAI21 LabsAI4Finance FoundationAlibabaAllenAIAmazonAnthropicApodexArcee AIBaiduByteDanceCactus ComputeCohereDeepGroveDeepSeekGoogleIBMIndexTeamInflection AIJingyao GongKyutaiLightricksLiquid AIMetaMicrosoftMiniMaxMistral AIMoonshot AINVIDIANous ResearchOpenAIOpenBMBOpenMOSSPerplexity AIPrior LabsRekaRobloxRoboflowStepFunTencentTsinghua UniversityUpstageWriterXiaomiZ.aixAI
Capability3d-generationagenticagentic-codingagentsanomaly-detectionaudioaudio-generationaudio-inputaudio-understandingcinematic-controlclassificationcodecodingcomputer-usedevice-controledge-inferenceeducationembeddingsemotion-controlextended-thinkingfile-searchfinance-nlpfinancial-modelingfine-tuningformal-verificationfunction-callingimage-segmentationimage-text-retrievalimage-to-videoimage-understandinginstance-segmentationjson-modejson-schemakeypoint-detectionlocal-deploymentlocal-inferencelong-contextlora-trainingmultilingualmultimodalmultimodal-inputobject-detectionofflineon-deviceon-device-inferenceopen-weightspart-controllable-generationpdf-inputpromptable-segmentationpronunciation-controlreasoningregressionresearchresponses-apisearch-groundingself-supervised-learningsentiment-analysisshape-generationspeed-controlstreamingstreaming-inferencestreaming-ttsstructured-outputsynthetic-datatexttext-to-3dtext-to-speechtext-to-videotime-series-forecastingtool-usetrainingvideovideo-editingvideo-extensionvideo-generationvideo-inputvideo-segmentationvideo-understandingvisionvoice-cloningvoice-designvolatility-predictionweb-searchzero-shot-classification
claude-opus-5
95
6.3
95
6.3
gemini-3.6-flash
36
8.0
36
8.0
gpt-5.6-sol
78
6.5
78
6.5
gpt-5.6-terra
57
6.5
57
6.5
gpt-5.6-luna
26
7.5
26
7.5
claude-sonnet-5
47
7.9
47
7.9
claude-fable-5
86
2.9
86
2.9
deepseek-v4-pro
40
61.8
40
61.8
gemini-3.7-flash
57
25.2
57
25.2
nvidia/nemotron-3.5-lightning-30b-a3b
7
13.3
7
13.3
gemini-3.5-flash-lite
18
12.8
18
12.8
muse-spark-1.1
56
20.4
56
20.4
z-ai/glm-5.3-flash
60
372
60
372
Cactus-Compute/needle2
Free / Free256 ctx
Free / Free256 ctx
deepseek-v4-flash
40
192
40
192
labs-leanstral-1-5
Free / Free256K ctx
Free / Free256K ctx
north-mini-code-1-0
7
7
NeedleSmall Language Model
Cactus-Compute/needle
Free / Free
Free / Free
GPT-5.5Multimodal LLM
openai/gpt-5.5
77
4.4
77
4.4
GPT-5.5 ProMultimodal LLM
openai/gpt-5.5-pro
83
0.8
83
0.8
Claude Opus 4.7Multimodal LLM
anthropic/claude-opus-4.7
60
4.0
60
4.0
Mistral Small 4Multimodal LLM
mistralai/mistral-small-2603
26
69.9
26
69.9
Claude Sonnet 4.6Multimodal LLM
anthropic/claude-sonnet-4.6
27
3.0
27
3.0
Claude Opus 4.6Multimodal LLM
anthropic/claude-opus-4.6
48
3.2
48
3.2