Docs/Models
Models
Every model on askr, what it costs, and where it runs.
Checked against the code on
This list is the live catalog, read when the page loads and cached for five minutes. Ids are exact and are what you pass as model. Prices are per million tokens for chat, per picture or per clip for media.
317 chat models on the API, 3 of them make pictures. 21 video models in the workspace. 153 listed but not routable.
All providersQwen 53OpenAI 48Google 26Anthropic 18Mistral 18Z.ai 17DeepSeek 16Meta 12ByteDance 10MoonshotAI 9MiniMax 9xAI 8
| Model | Id | Runs on | In / out, $ per 1M | Turn | Context |
|---|---|---|---|---|---|
| Z.ai 17 | |||||
| GLM 5.3popular | glm-5.3 | API, workspace | 1.48 / 4.64 | 3.8 cr | 1311K |
| GLM 5.2 (Fast)popular | glm-5.2-fast | API, workspace | 2.22 / 6.96 | 5.7 cr | 1049K |
| GLM Flash Latest | ~z-ai/glm-flash-latest | API, workspace | 0.079 / 0.264 | 0.21 cr | 1311K |
| GLM 4.7 Flash | z-ai/glm-4.7-flash | API, workspace | 0.064 / 0.422 | 0.27 cr | 200K |
| GLM 5.3 Flash | z-ai/glm-5.3-flash | API, workspace | 0.158 / 0.527 | 0.42 cr | 1311K |
| GLM 4.5 Air | z-ai/glm-4.5-air | API, workspace | 0.137 / 0.897 | 0.59 cr | 131K |
| GLM 4.6V | z-ai/glm-4.6v | API, workspace | 0.317 / 0.95 | 0.79 cr | 131K |
| GLM 4.7 | z-ai/glm-4.7 | API, workspace | 0.422 / 1.85 | 1.3 cr | 205K |
| GLM 4.6 | z-ai/glm-4.6 | API, workspace | 0.454 / 1.85 | 1.4 cr | 205K |
| GLM 4.5V | z-ai/glm-4.5v | API, workspace | 0.633 / 1.90 | 1.6 cr | 66K |
| GLM 5 | z-ai/glm-5 | API, workspace | 0.633 / 2.03 | 1.6 cr | 205K |
| GLM 4.5 | z-ai/glm-4.5 | API, workspace | 0.633 / 2.32 | 1.8 cr | 131K |
| GLM Latest | ~z-ai/glm-latest | API, workspace | 0.941 / 2.96 | 2.4 cr | 1311K |
| GLM 5.1 | z-ai/glm-5.1 | API, workspace | 1.02 / 3.20 | 2.6 cr | 205K |
| GLM 5 Turbo | z-ai/glm-5-turbo | API, workspace | 1.27 / 4.22 | 3.4 cr | 203K |
| GLM 5V Turbo | z-ai/glm-5v-turbo | API, workspace | 1.27 / 4.22 | 3.4 cr | 203K |
| GLM 5.2 | z-ai/glm-5.2 | API, workspace | 1.48 / 4.64 | 3.8 cr | 1049K |
Listed but not routable (153)
Audio and embedding models, image models that are not chat models, models with no published rate, and the ids their provider is not serving. None of these can be called. Why.
amazon/nova-premier-v1refusedanthropic/claude-opus-4refusedaura-srimageautochatautoclawchatbaai/bge-base-en-v1.5embeddingbaai/bge-large-en-v1.5embeddingbaai/bge-m3embeddingbirefnet-v2imagecohere/north-mini-coderefusedcrystal-upscalerimagedeepgram-nova-3audiodeepseek/deepseek-v4-pro-0813refusedeleven_flash_v2_5audioeleven_v3audioelevenlabs-music-v1audiofast-sdxlimageflux-2-fleximageflux-2-proimageflux-2-pro-i2iimageflux-2-pro-outpaintimageflux-3-i2vimageflux-kontext-maximageflux-kontext-proimageflux-proimagegemini-omni-flash-i2vimagegemini-omni-flash-r2vimagegemini-omni-flash-v2vimagegoogle/gemini-3-pro-imagerefusedgoogle/gemini-embedding-001embeddinggoogle/gemini-embedding-2embeddinggoogle/lyria-3-clip-previewchatgoogle/lyria-3-pro-previewchatgpt-image-1imagegpt-image-1.5imagegpt-image-2imagegpt-image-2.5-flareimagegpt-image-2.5-sunburstimagegrok-imagineimagegrok-imagine-editimagegrok-imagine-image-2imagegrok-imagine-image-2-editimagegrok-imagine-video-1.5-i2vimagegrok-imagine-video-i2vimagegrok-imagine-video-v2vimagehappy-horse-1.1-i2vimagehappy-horse-1.1-r2vimageideogram-v3imageideogram-v3-remiximageintfloat/e5-base-v2embeddingintfloat/e5-large-v2embeddingintfloat/multilingual-e5-largeembeddingkling-2.5-turbo-i2vimagekling-o3-pro-i2vimagekling-o3-pro-v2vimagekling-o3-standard-i2vimagekling-o3-standard-v2vimagekling-v3-imageimagekling-v3-image-editimagekling-v3-pro-i2vimagekling-v3-standard-i2vimagekrea-2-turboimageliquid/lfm-2.5-2.6brefusedliquid/lfm-2.5-embedding-350m:freeembeddingmancer/weaverrefusedmeta/muse-spark-1.1refusedmeta/muse-spark-1.2refusedminimax-h3-i2vimageminimax-h3-r2vimagemistralai/codestral-embed-2505embeddingmistralai/mistral-embed-2312embeddingnano-banana-2imagenano-banana-2-editimagenano-banana-proimagenano-banana-pro-directimagenano-banana-pro-direct-editimagenvidia/llama-nemotron-embed-vl-1b-v2:freeembeddingnvidia/nemotron-3-embed-1b:freeembeddingnvidia/nemotron-3-nano-omni-30b-a3b-reasoningrefusednvidia/nemotron-3.5-content-safetyrefusedopenai/gpt-5-imagerefusedopenai/gpt-5-image-minirefusedopenai/gpt-5-prorefusedopenai/gpt-5.2-chatrefusedopenai/gpt-5.2-prorefusedopenai/gpt-5.4-image-2refusedopenai/gpt-audiorefusedopenai/gpt-audio-minirefusedopenai/o1refusedopenai/o1-prorefusedopenai/o3refusedopenai/o3-minirefusedopenai/o3-mini-highrefusedopenai/o3-prorefusedopenai/o4-minirefusedopenai/o4-mini-highrefusedopenai/text-embedding-3-largeembeddingopenai/text-embedding-3-smallembeddingopenai/text-embedding-ada-002embeddingperplexity/pplx-embed-v1-0.6bembeddingperplexity/pplx-embed-v1-4bembeddingperplexity/sonarrefusedperplexity/sonar-deep-researchrefusedperplexity/sonar-prorefusedperplexity/sonar-pro-searchrefusedperplexity/sonar-reasoning-prorefusedpika-i2vimageprivate/gemma4-31brefusedprivate/gpt-oss-120brefusedprivate/kimi-k3refusedprivate/llama3-3-70brefusedqwen-image-2imageqwen-image-2-editimageqwen-image-3imageqwen-image-3-editimageqwen/qwen3-8brefusedqwen/qwen3-embedding-4bembeddingqwen/qwen3-embedding-8bembeddingrecraft-v4.1imagerecraft-v4.1-svgimagerelace/relace-apply-3refusedsakana/fugu-ultrarefusedsakana/sakana-namazurefusedseedance-2-5-i2vimageseedance-2-fast-i2vimageseedance-2-i2vimageseedream-4.5imageseedream-4.5-editimageseedream-v5-liteimageseedream-v5-lite-editimageseedream-v5-proimageseedream-v5-pro-editimagesentence-transformers/all-minilm-l12-v2embeddingsentence-transformers/all-minilm-l6-v2embeddingsentence-transformers/all-mpnet-base-v2embeddingsentence-transformers/multi-qa-mpnet-base-dot-v1embeddingsentence-transformers/paraphrase-minilm-l6-v2embeddingthedrummer/cydonia-24b-v4.1refusedthenlper/gte-baseembeddingthenlper/gte-largeembeddingtopaz-upscaleimageveo3-fast-i2vimageveo3-i2vimagevoyageai/voyage-4embeddingvoyageai/voyage-4-largeembeddingvoyageai/voyage-4-liteembeddingvoyageai/voyage-code-4embeddingvoyageai/voyage-multimodal-3.5embeddingwan-3.0-i2vimagewan-3.0-prime-i2vimagewan-3.0-prime-r2vimagewan-3.0-r2vimagexai-realtime-voiceaudio
Reading the table
- Runs on
- API means
POST /v1/chat/completionsaccepts it. Workspace means the browser can run it. Video is workspace only. - Turn
- Credits for a typical chat turn: 1,000 tokens in, 500 out. Thinking models run above this because reasoning is billed as output.
- Picture
- The measured credits per image where we have measured it; the reservation otherwise.
- Clip
- The cheapest published shape and length. Longer and larger cost more; the workspace shows the exact price before you render.
Ids that the catalog lists but the provider is not currently serving are left out. Listed but refused explains.