Price list
Every world, priced in the open
Each row shows what the upstream publishes and what Antimatter charges, side by side. The difference is the margin, and it is stated, not hidden in a rate.
+10%
The upstream's published price plus 10%. That is the whole margin, on every world.
The same for every account. No fee to add credit: what you send is what you are credited. It is not cheaper than buying direct and does not pretend to be: it is the upstream's price plus this, for one key, one wire and every world.
- 464
- 458
- 65
- Mistral: Ministral 3 14B 2512mistralai/ministral-14b-2512visiontoolsjson262k contextIn / 1M$0.22up $0.2Out / 1M$0.22up $0.2Margin+10%1st tokenunmeasured
- NVIDIA: Nemotron 3.5 Content Safetynvidia/nemotron-3.5-content-safetyvision131k contextIn / 1M$0.22up $0.2Out / 1M$0.22up $0.2Margin+10%1st tokenunmeasured
- OpenAI: GPT Luna Latest~openai/gpt-luna-latestvisiontoolsjson1.1M contextIn / 1M$0.11up $0.1Out / 1M$0.55up $0.5Margin+10%1st tokenunmeasured
- OpenAI: GPT-6 Lunaopenai/gpt-6-lunavisiontoolsjson1.1M contextIn / 1M$0.11up $0.1Out / 1M$0.55up $0.5Margin+10%1st tokenunmeasured
- OpenAI: GPT-6 Luna Proopenai/gpt-6-luna-provisiontoolsjson1.1M contextIn / 1M$0.11up $0.1Out / 1M$0.55up $0.5Margin+10%1st tokenunmeasured
- Qwen: Qwen3 VL 8B Instructqwen/qwen3-vl-8b-instructvisiontoolsjson262k contextIn / 1M$0.1287up $0.117Out / 1M$0.5005up $0.455Margin+10%1st tokenunmeasured
- OpenAI: GPT-5.6 Luna Pro (batch)openai/gpt-5.6-luna-pro:batchvisiontoolsjson1.1M contextIn / 1M$0.11up $0.1Out / 1M$0.66up $0.6Margin+10%1st tokenunmeasured
- OpenAI: GPT-5.6 Luna (batch)openai/gpt-5.6-luna:batchvisiontoolsjson1.1M contextIn / 1M$0.11up $0.1Out / 1M$0.66up $0.6Margin+10%1st tokenunmeasured
- Qwen: Qwen3.8 Flashqwen/qwen3.8-flashvisiontoolsjson1M contextIn / 1M$0.165up $0.15Out / 1M$0.517up $0.47Margin+10%1st tokenunmeasured
- Qwen: Qwen3.8 Omni Flashqwen/qwen3.8-omni-flashvisiontoolsjson1M contextIn / 1M$0.165up $0.15Out / 1M$0.517up $0.47Margin+10%1st tokenunmeasured
- OpenAI: GPT-5.4 Nano (batch)openai/gpt-5.4-nano:batchvisiontoolsjson400k contextIn / 1M$0.11up $0.1Out / 1M$0.6875up $0.625Margin+10%1st tokenunmeasured
- Z.ai: GLM 5.3 Flashz-ai/glm-5.3-flashvisiontoolsjson1M contextIn / 1M$0.165up $0.15Out / 1M$0.55up $0.5Margin+10%1st tokenunmeasured
- Z.ai: GLM Flash Latest~z-ai/glm-flash-latestvisiontoolsjson1M contextIn / 1M$0.02888up $0.02625Out / 1M$0.99up $0.9Margin+10%1st tokenunmeasured
- Mistral: Mistral Small 4mistralai/mistral-small-2603visiontoolsjson262k contextIn / 1M$0.165up $0.15Out / 1M$0.66up $0.6Margin+10%1st tokenunmeasured
- OpenAI: GPT-4o-miniopenai/gpt-4o-minivisiontoolsjson128k contextIn / 1M$0.165up $0.15Out / 1M$0.66up $0.6Margin+10%1st tokenunmeasured
- OpenAI: GPT-4o-mini (2024-07-18)openai/gpt-4o-mini-2024-07-18visiontoolsjson128k contextIn / 1M$0.165up $0.15Out / 1M$0.66up $0.6Margin+10%1st tokenunmeasured
- Qwen: Qwen3 VL 30B A3B Instructqwen/qwen3-vl-30b-a3b-instructvisiontoolsjson262k contextIn / 1M$0.165up $0.15Out / 1M$0.66up $0.6Margin+10%1st tokenunmeasured
- Google: Gemini 3.1 Flash Lite (batch)google/gemini-3.1-flash-lite:batchvisiontoolsjson1M contextIn / 1M$0.1375up $0.125Out / 1M$0.825up $0.75Margin+10%1st tokenunmeasured
- DeepSeek: DeepSeek Flash Latest~deepseek/deepseek-flash-latestvisiontoolsjson1M contextIn / 1M$0.005032up $0.004575Out / 1M$1.32up $1.20Margin+10%1st tokenunmeasured
- Meta: Llama 4 Maverickmeta-llama/llama-4-maverickvisiontoolsjson1M contextIn / 1M$0.2062up $0.1875Out / 1M$0.7177up $0.6525Margin+10%1st tokenunmeasured
- DeepSeek: DeepSeek V4 Flash Vision Expdeepseek/deepseek-v4-flash-vision-expvisiontoolsjson1M contextIn / 1M$0.2372up $0.2156Out / 1M$0.7115up $0.6468Margin+10%1st tokenunmeasured
- OpenAI: GPT-5 Mini (batch)openai/gpt-5-mini:batchvisiontoolsjson400k contextIn / 1M$0.1375up $0.125Out / 1M$1.10up $1.00Margin+10%1st tokenunmeasured
- OpenAI: GPT-4.1 Mini (batch)openai/gpt-4.1-mini:batchvisiontoolsjson1M contextIn / 1M$0.22up $0.2Out / 1M$0.88up $0.8Margin+10%1st tokenunmeasured
- Qwen: Qwen3.5-35B-A3Bqwen/qwen3.5-35b-a3bvisiontoolsjson262k contextIn / 1M$0.165up $0.15Out / 1M$1.10up $1.00Margin+10%1st tokenunmeasured
- Qwen: Qwen3.6 35B A3Bqwen/qwen3.6-35b-a3bvisiontoolsjson262k contextIn / 1M$0.165up $0.15Out / 1M$1.10up $1.00Margin+10%1st tokenunmeasured
- Mistral: Mistral Large 3 2512 (batch)mistralai/mistral-large-2512:batchvisiontoolsjson262k contextIn / 1M$0.275up $0.25Out / 1M$0.825up $0.75Margin+10%1st tokenunmeasured
- Mistral: Mistral Medium 3.1 (batch)mistralai/mistral-medium-3.1:batchvisiontoolsjson131k contextIn / 1M$0.22up $0.2Out / 1M$1.10up $1.00Margin+10%1st tokenunmeasured
- Mistral: Mistral Small 3.1 24Bmistralai/mistral-small-3.1-24b-instructvisiontools128k contextIn / 1M$0.3861up $0.351Out / 1M$0.6105up $0.555Margin+10%1st tokenunmeasured
- Qwen: Qwen3.6 Flashqwen/qwen3.6-flashvisiontoolsjson1M contextIn / 1M$0.2062up $0.1875Out / 1M$1.238up $1.125Margin+10%1st tokenunmeasured
- Google: Gemini 2.5 Flash (batch)google/gemini-2.5-flash:batchvisiontoolsjson1M contextIn / 1M$0.165up $0.15Out / 1M$1.375up $1.25Margin+10%1st tokenunmeasured
- Google: Gemini 3.5 Flash Lite (batch)google/gemini-3.5-flash-lite:batchvisiontoolsjson1M contextIn / 1M$0.165up $0.15Out / 1M$1.375up $1.25Margin+10%1st tokenunmeasured
- MiniMax: MiniMax-01minimax/minimax-01vision1M contextIn / 1M$0.22up $0.2Out / 1M$1.21up $1.10Margin+10%1st tokenunmeasured
- StepFun: Step 3.7 Flashstepfun/step-3.7-flashvisiontoolsjson262k contextIn / 1M$0.22up $0.2Out / 1M$1.265up $1.15Margin+10%1st tokenunmeasured
- OpenAI: GPT-5.6 Lunaopenai/gpt-5.6-lunavisiontoolsjson1.1M contextIn / 1M$0.22up $0.2Out / 1M$1.32up $1.20Margin+10%1st tokenunmeasured
- OpenAI: GPT-5.6 Luna Proopenai/gpt-5.6-luna-provisiontoolsjson1.1M contextIn / 1M$0.22up $0.2Out / 1M$1.32up $1.20Margin+10%1st tokenunmeasured
- Z.ai: GLM 4.6Vz-ai/glm-4.6vvisiontoolsjson131k contextIn / 1M$0.33up $0.3Out / 1M$0.99up $0.9Margin+10%1st tokenunmeasured
- OpenAI: GPT-5.4 Nanoopenai/gpt-5.4-nanovisiontoolsjson400k contextIn / 1M$0.22up $0.2Out / 1M$1.375up $1.25Margin+10%1st tokenunmeasured
- Perceptron: Perceptron Mk1perceptron/perceptron-mk1visionjson33k contextIn / 1M$0.165up $0.15Out / 1M$1.65up $1.50Margin+10%1st tokenunmeasured
- Perceptron: Perceptron Mk1.5perceptron/perceptron-mk1.5visiontoolsjson37k contextIn / 1M$0.165up $0.15Out / 1M$1.65up $1.50Margin+10%1st tokenunmeasured
- DeepSeek: DeepSeek V4.1 Flashdeepseek/deepseek-v4.1-flashvisiontoolsjson1M contextIn / 1M$0.33up $0.3Out / 1M$1.32up $1.20Margin+10%1st tokenunmeasured
“up” is the upstream's published rate, read from its catalog. A tag with a tick was measured working by a probe on this deployment; one without is the upstream's own word. First token is the median measured here over the last week, from real traffic and probes. Prices follow the upstream: when it changes a rate, this list and the bill change with it.
Calculator
What a request costs
The three cheapest worlds in the list as it is filtered now. Add up to four of your own with Compare. Computed with the same function the gateway bills with. Token counts are yours to set; the tasks are round numbers to start from.
| World | Antimatter | Upstream | Margin |
|---|---|---|---|
| inclusionAI: Ling 3.0 Flash VLinclusionai/ling-3.0-flash-vl | $0.0514 | $0.0468 | $0.004676 |
| Nex AGI: Nex-N2.5-Mininex-agi/nex-n2.5-mini | $0.0715 | $0.0650 | $0.006500 |
| Qwen: Qwen3.7 Flashqwen/qwen3.7-flash | $0.0896 | $0.0815 | $0.008150 |
A thousand requests of 1,200 tokens in and 350 out. A cached prompt token costs less where the upstream publishes a cache rate.
By need
The cheapest world for a need
The worlds that read images, hold a long context, call tools or return JSON, ranked by what a typical message costs at these prices.
Autopilot
Let the router choose
An autopilot id picks a world per request and is billed at that world's price from this list, with the same margin. The answer names the world in x-antimatter-model.
antimatter/autoBalanced: price, measured speed and measured quality, scored together.antimatter/auto:cheapThe lowest blended price in the catalog.antimatter/auto:fastThe lowest measured time to first token.antimatter/auto:codeTool calls measured working and a context sized for code.antimatter/auto:longA long context window, from the catalog.antimatter/auto:visionA world that takes images as input.antimatter/auto:toolsA world whose tool calls were measured working.antimatter/auto:jsonA world whose structured output was measured holding.