Docs/Models
Models
Every model on askr, what it costs, and where it runs.
Checked against the code on
This list is the live catalog, read when the page loads and cached for five minutes. Ids are exact and are what you pass as model. Prices are per million tokens for chat, per picture or per clip for media.
317 chat models on the API, 3 of them make pictures. 21 video models in the workspace. 153 listed but not routable.
All providersQwen 53OpenAI 48Google 26Anthropic 18Mistral 18Z.ai 17DeepSeek 16Meta 12ByteDance 10MoonshotAI 9MiniMax 9xAI 8
| Model | Id | Runs on | In / out, $ per 1M | Turn | Context |
|---|---|---|---|---|---|
| Mistral 18 | |||||
| Mistral Nemo | mistralai/mistral-nemo | API, workspace | 0.02 / 0.032 | 0.036 cr | 131K |
| Mistral Small 3 | mistralai/mistral-small-24b-instruct-2501 | API, workspace | 0.053 / 0.084 | 0.095 cr | 33K |
| Ministral 3 3B 2512 | mistralai/ministral-3b-2512 | API, workspace | 0.105 / 0.105 | 0.16 cr | 131K |
| Mistral Small 3.2 24B | mistralai/mistral-small-3.2-24b-instruct | API, workspace | 0.099 / 0.264 | 0.23 cr | 256K |
| Ministral 3 8B 2512 | mistralai/ministral-8b-2512 | API, workspace | 0.158 / 0.158 | 0.24 cr | 262K |
| Voxtral Small 24B 2507 | mistralai/voxtral-small-24b-2507 | API, workspace | 0.105 / 0.317 | 0.26 cr | 33K |
| Ministral 3 14B 2512 | mistralai/ministral-14b-2512 | API, workspace | 0.211 / 0.211 | 0.32 cr | 262K |
| Mistral Small 4 | mistralai/mistral-small-2603 | API, workspace | 0.158 / 0.633 | 0.47 cr | 262K |
| Saba | mistralai/mistral-saba | API, workspace | 0.211 / 0.633 | 0.53 cr | 33K |
| Mistral Small 3.1 24B | mistralai/mistral-small-3.1-24b-instruct | API, workspace | 0.37 / 0.586 | 0.66 cr | 128K |
| Codestral 2508 | mistralai/codestral-2508 | API, workspace | 0.317 / 0.95 | 0.79 cr | 256K |
| Devstral 2 2512 | mistralai/devstral-2512 | API, workspace | 0.422 / 2.11 | 1.5 cr | 262K |
| Mistral Medium 3 | mistralai/mistral-medium-3 | API, workspace | 0.422 / 2.11 | 1.5 cr | 131K |
| Mistral Medium 3.1 | mistralai/mistral-medium-3.1 | API, workspace | 0.422 / 2.11 | 1.5 cr | 131K |
| Mistral Large | mistralai/mistral-large | API, workspace | 2.11 / 6.33 | 5.3 cr | 128K |
| Mistral Large 2407 | mistralai/mistral-large-2407 | API, workspace | 2.11 / 6.33 | 5.3 cr | 131K |
| Mixtral 8x22B Instruct | mistralai/mixtral-8x22b-instruct | API, workspace | 2.11 / 6.33 | 5.3 cr | 66K |
| Mistral Medium 3.5 | mistralai/mistral-medium-3-5 | API, workspace | 1.58 / 7.91 | 5.5 cr | 262K |
Listed but not routable (153)
Audio and embedding models, image models that are not chat models, models with no published rate, and the ids their provider is not serving. None of these can be called. Why.
amazon/nova-premier-v1refusedanthropic/claude-opus-4refusedaura-srimageautochatautoclawchatbaai/bge-base-en-v1.5embeddingbaai/bge-large-en-v1.5embeddingbaai/bge-m3embeddingbirefnet-v2imagecohere/north-mini-coderefusedcrystal-upscalerimagedeepgram-nova-3audiodeepseek/deepseek-v4-pro-0813refusedeleven_flash_v2_5audioeleven_v3audioelevenlabs-music-v1audiofast-sdxlimageflux-2-fleximageflux-2-proimageflux-2-pro-i2iimageflux-2-pro-outpaintimageflux-3-i2vimageflux-kontext-maximageflux-kontext-proimageflux-proimagegemini-omni-flash-i2vimagegemini-omni-flash-r2vimagegemini-omni-flash-v2vimagegoogle/gemini-3-pro-imagerefusedgoogle/gemini-embedding-001embeddinggoogle/gemini-embedding-2embeddinggoogle/lyria-3-clip-previewchatgoogle/lyria-3-pro-previewchatgpt-image-1imagegpt-image-1.5imagegpt-image-2imagegpt-image-2.5-flareimagegpt-image-2.5-sunburstimagegrok-imagineimagegrok-imagine-editimagegrok-imagine-image-2imagegrok-imagine-image-2-editimagegrok-imagine-video-1.5-i2vimagegrok-imagine-video-i2vimagegrok-imagine-video-v2vimagehappy-horse-1.1-i2vimagehappy-horse-1.1-r2vimageideogram-v3imageideogram-v3-remiximageintfloat/e5-base-v2embeddingintfloat/e5-large-v2embeddingintfloat/multilingual-e5-largeembeddingkling-2.5-turbo-i2vimagekling-o3-pro-i2vimagekling-o3-pro-v2vimagekling-o3-standard-i2vimagekling-o3-standard-v2vimagekling-v3-imageimagekling-v3-image-editimagekling-v3-pro-i2vimagekling-v3-standard-i2vimagekrea-2-turboimageliquid/lfm-2.5-2.6brefusedliquid/lfm-2.5-embedding-350m:freeembeddingmancer/weaverrefusedmeta/muse-spark-1.1refusedmeta/muse-spark-1.2refusedminimax-h3-i2vimageminimax-h3-r2vimagemistralai/codestral-embed-2505embeddingmistralai/mistral-embed-2312embeddingnano-banana-2imagenano-banana-2-editimagenano-banana-proimagenano-banana-pro-directimagenano-banana-pro-direct-editimagenvidia/llama-nemotron-embed-vl-1b-v2:freeembeddingnvidia/nemotron-3-embed-1b:freeembeddingnvidia/nemotron-3-nano-omni-30b-a3b-reasoningrefusednvidia/nemotron-3.5-content-safetyrefusedopenai/gpt-5-imagerefusedopenai/gpt-5-image-minirefusedopenai/gpt-5-prorefusedopenai/gpt-5.2-chatrefusedopenai/gpt-5.2-prorefusedopenai/gpt-5.4-image-2refusedopenai/gpt-audiorefusedopenai/gpt-audio-minirefusedopenai/o1refusedopenai/o1-prorefusedopenai/o3refusedopenai/o3-minirefusedopenai/o3-mini-highrefusedopenai/o3-prorefusedopenai/o4-minirefusedopenai/o4-mini-highrefusedopenai/text-embedding-3-largeembeddingopenai/text-embedding-3-smallembeddingopenai/text-embedding-ada-002embeddingperplexity/pplx-embed-v1-0.6bembeddingperplexity/pplx-embed-v1-4bembeddingperplexity/sonarrefusedperplexity/sonar-deep-researchrefusedperplexity/sonar-prorefusedperplexity/sonar-pro-searchrefusedperplexity/sonar-reasoning-prorefusedpika-i2vimageprivate/gemma4-31brefusedprivate/gpt-oss-120brefusedprivate/kimi-k3refusedprivate/llama3-3-70brefusedqwen-image-2imageqwen-image-2-editimageqwen-image-3imageqwen-image-3-editimageqwen/qwen3-8brefusedqwen/qwen3-embedding-4bembeddingqwen/qwen3-embedding-8bembeddingrecraft-v4.1imagerecraft-v4.1-svgimagerelace/relace-apply-3refusedsakana/fugu-ultrarefusedsakana/sakana-namazurefusedseedance-2-5-i2vimageseedance-2-fast-i2vimageseedance-2-i2vimageseedream-4.5imageseedream-4.5-editimageseedream-v5-liteimageseedream-v5-lite-editimageseedream-v5-proimageseedream-v5-pro-editimagesentence-transformers/all-minilm-l12-v2embeddingsentence-transformers/all-minilm-l6-v2embeddingsentence-transformers/all-mpnet-base-v2embeddingsentence-transformers/multi-qa-mpnet-base-dot-v1embeddingsentence-transformers/paraphrase-minilm-l6-v2embeddingthedrummer/cydonia-24b-v4.1refusedthenlper/gte-baseembeddingthenlper/gte-largeembeddingtopaz-upscaleimageveo3-fast-i2vimageveo3-i2vimagevoyageai/voyage-4embeddingvoyageai/voyage-4-largeembeddingvoyageai/voyage-4-liteembeddingvoyageai/voyage-code-4embeddingvoyageai/voyage-multimodal-3.5embeddingwan-3.0-i2vimagewan-3.0-prime-i2vimagewan-3.0-prime-r2vimagewan-3.0-r2vimagexai-realtime-voiceaudio
Reading the table
- Runs on
- API means
POST /v1/chat/completionsaccepts it. Workspace means the browser can run it. Video is workspace only. - Turn
- Credits for a typical chat turn: 1,000 tokens in, 500 out. Thinking models run above this because reasoning is billed as output.
- Picture
- The measured credits per image where we have measured it; the reservation otherwise.
- Clip
- The cheapest published shape and length. Longer and larger cost more; the workspace shows the exact price before you render.
Ids that the catalog lists but the provider is not currently serving are left out. Listed but refused explains.