Models
AI model ratings
Compare model fit by overall score, coding, reasoning, vision, context, speed and price/power. Values are demo data for MVP evaluation.
Model filters
Sort and narrow the demo model board by task profile.
Filtered table
| Rank | Model | Provider | Intelligence | Coding | Agent Power | Speed | Cost Efficiency | Context | Overall Score |
|---|---|---|---|---|---|---|---|---|---|
| #1 | GPT-5.5 | OpenAI | 96 | 95 | 94 | 78 | 72 | 128K tokens | 89.9 |
| #2 | Claude Opus / Sonnet | Anthropic | 95 | 93 | 91 | 82 | 76 | 200K tokens | 89.6 |
| #3 | DeepSeek | DeepSeek | 88 | 90 | 82 | 84 | 93 | 128K tokens | 87.7 |
| #4 | Gemini | 91 | 86 | 84 | 88 | 80 | 1M tokens | 86.4 | |
| #5 | Qwen | Alibaba | 86 | 88 | 80 | 86 | 89 | 128K tokens | 85.7 |
| #6 | Kimi | Moonshot AI | 87 | 84 | 81 | 83 | 86 | 1M tokens | 84.5 |
| #7 | Mistral | Mistral AI | 84 | 82 | 78 | 89 | 88 | 128K tokens | 83.4 |
| #8 | Llama | Meta / local hosts | 80 | 79 | 72 | 74 | 95 | 32K tokens | 79.8 |