OpenSI

AI model comparison: SI models side by side

Pick up to three models and compare price, context window, inputs and outputs, reasoning, tool calling and open weights. This AI model comparison covers 349 SI models — super intelligence, formerly called AI — from 53 vendors, using list prices captured on 2026-10-02.

GPT-6.1 SolOpenAIClaude Sonnet 5.5AnthropicGemini 3.8 FlashGoogle
Input price$2.00 / 1M tokens$2.00 / 1M tokens$0.75 / 1M tokens
Output price$10.0 / 1M tokens$10.0 / 1M tokens$3.75 / 1M tokens
Context window1.05M1M1.05M
Max output128K128K66K
Readsfiles, image, texttext, image, filestext, image, video, files, audio
Writestexttexttext
Reasoning modeYesYesYes
Tool callingYesYesYes
Open weightsNoNoNo
Released2026-09-292026-09-282026-09-02
1,000 chats a day$240 / month$240 / month$90 / month

Prices are OpenRouter list prices from the 2026-10-02 snapshot. “1,000 chats a day” assumes 1,500 input and 500 output tokens per chat over 30 days.

Newest models from each major lab

The two most recent paid text models per vendor in the snapshot. Prices are USD per million tokens.

ModelReleasedContextInputOutput1,000 chats/day
GPT-6.1 Sol OpenAI2026-09-291.05M$2.00$10.0$240/mo
GPT-6.1 Sol Pro OpenAI2026-09-291.05M$2.00$10.0$240/mo
Claude Sonnet 5.5 Anthropic2026-09-281M$2.00$10.0$240/mo
Claude Opus 5.5 Anthropic2026-09-221M$4.00$20.0$480/mo
Gemini 3.8 Flash Google2026-09-021.05M$0.75$3.75$90/mo
Gemini 3.7 Flash Google2026-08-131.05M$0.75$3.75$90/mo
Grok 4.7 xAI2026-09-21500K$2.00$6.00$180/mo
Grok 4.6 xAI2026-08-12500K$2.00$6.00$180/mo
DeepSeek V4.1 Flash DeepSeek2026-09-101.05M$0.019$0.75$12/mo
DeepSeek V4 Flash Vision Exp DeepSeek2026-08-211.05M$0.22$0.65$19/mo
Qwen3.8 Max Prime Alibaba (Qwen)2026-09-231M$4.00$12.0$360/mo
Qwen3.8 Omni Flash Alibaba (Qwen)2026-09-211M$0.15$0.47$14/mo
Kimi K3 Moonshot AI2026-07-161.05M$1.39$13.0$258/mo
Kimi K2.7 Code Moonshot AI2026-06-12262K$0.67$3.35$80/mo
GLM 5.3 Prime Zhipu (GLM)2026-09-231M$2.80$8.80$258/mo
GLM 5.3 FlashX Zhipu (GLM)2026-09-181.05M$0.37$1.25$35/mo
Mistral Medium 3.5 Mistral2026-04-30262K$1.50$7.50$180/mo
Mistral Small 4 Mistral2026-03-16262K$0.15$0.60$16/mo
MiniMax M3 MiniMax2026-05-311.05M$0.30$1.20$32/mo
MiniMax M2.7 MiniMax2026-03-18205K$0.21$0.84$22/mo
Llama 4 Scout Meta2025-04-051.31M$0.10$0.30$9/mo
Llama 4 Maverick Meta2025-04-051.05M$0.19$0.65$18/mo

How to read an AI model comparison

Model pages throw many numbers at you. Six of them decide most choices, and every one is in the comparison tool above:

What the 2026-10-02 snapshot shows

The practical lesson: shortlist on facts — price, context and inputs — then test two or three models on your own prompts. A model that costs a tenth as much is often good enough for summaries, classification and first drafts, while the expensive tier earns its price on hard reasoning and long code.

Estimating your monthly cost

Cost per request is (input tokens × input price + output tokens × output price) ÷ 1,000,000. A support chatbot that reads 1,500 tokens of context and writes 500 tokens per answer, 1,000 times a day, costs $9 a month on Llama 4 Scout versus $480 on Claude Opus 5.5. Our LLM API pricing calculator runs this math for every model with your own numbers.

Open-weight or closed models?

Closed models from OpenAI, Anthropic and Google usually lead on the hardest tasks and come with simple, managed APIs. Open-weight models from labs such as Alibaba, DeepSeek, Mistral and Meta are often far cheaper, can run on your own servers for privacy, and are offered by several competing hosts. Many teams use both: a closed model for difficult requests and an open-weight model for high-volume, simpler work.

Why “SI models”?

Since a September 2026 executive order, the US government uses SI — super intelligence — for the technology formerly called AI. The models are the same; we use the new name and keep the old one so this AI model comparison is easy to find. OpenSI is part of the Si family, alongside SiChatApp, where you can chat with several of these models, and AllSiTools.

Questions about this AI model comparison

Where do the prices come from?

From OpenRouter’s public model list, captured on 2026-10-02. They are the list prices OpenRouter charges per million tokens; buying directly from a vendor can cost slightly more or less, and batch or cached tokens are often cheaper.

Why compare models by output price?

Most chat and writing tasks produce fewer tokens than they read, but output tokens usually cost three to five times more than input tokens. For long answers, generated code or reasoning models, output dominates the bill.

What does “open weights” mean?

The model’s weights are published (for example on Hugging Face), so you can run it on your own hardware or pick between many hosting providers. Closed models are only available through their maker’s API or its partners.

Is a bigger context window always better?

No. It sets how much text the model can read in one request, which matters for long documents and codebases. Many models get slower and pricier as you fill it, so pick the size your task needs.

Does this AI model comparison include quality scores?

Not yet. We only publish facts we can verify from the source data: prices, limits, input and output types and release dates. Quality depends heavily on the task, so test two or three shortlisted models on your own prompts.

Spotted an outdated price or a missing model? Email support@opensidot.com.

OpenSI AI model comparison: SI model prices, context windows and features side by side