# AI models

[AI Inference](/en/documentation/platform/ai-inference/) runs the models on this page. You select a model by its id, a string the model answers to, rather than by the name it is listed under. Each row of the catalog carries that id next to the capabilities the model page states for itself. For the request body the id travels in, refer to [Model invocation](/en/documentation/platform/ai-inference/model-invocation/).

## Model catalog

Each row is one model, and the model name links the model page, which carries a request example. The Category column reads LLM for a large language model, VLM for a vision language model, Embedding, or Reranker. A dash means the model page states no value for that capability.

| Model                                                                                       | Category  | Input          | Context length | Size            | Tool calling | LoRA | Model id                                           |
| ------------------------------------------------------------------------------------------- | --------- | -------------- | -------------- | --------------- | ------------ | ---- | -------------------------------------------------- |
| [Mistral 3 Small (24B AWQ)](/en/documentation/platform/ai-inference/mistral-3-small/)       | LLM       | Text           | 32k tokens     | 24B parameters  | Yes          | No   | `casperhansen-mistral-small-24b-instruct-2501-awq` |
| [BAAI/bge-reranker-v2-m3](/en/documentation/platform/ai-inference/baai-bge-reranker-v2-m3/) | Reranker  | Text           | 8k tokens      | 568M parameters | —            | No   | `baai-bge-reranker-v2-m3`                          |
| [InternVL3](/en/documentation/platform/ai-inference/internvl3/)                             | VLM       | Text and image | 16k tokens     | 1B parameters   | No           | No   | `opengvlab-internvl3-1b-instruct`                  |
| [Qwen2.5 VL AWQ 3B](/en/documentation/platform/ai-inference/qwen-2-5-vl-3b/)                | VLM       | Text and image | 32k tokens     | 3B parameters   | Yes          | Yes  | `qwen-qwen25-vl-3b-instruct-awq`                   |
| [Qwen2.5 VL AWQ 7B](/en/documentation/platform/ai-inference/qwen-2-5-vl-7b/)                | VLM       | Text and image | 32k tokens     | 7B parameters   | Yes          | Yes  | `qwen-qwen25-vl-7b-instruct-awq`                   |
| [Qwen3 30B A3B Instruct 2507 FP8](/en/documentation/platform/ai-inference/qwen3-30ba3b/)    | LLM       | Text           | 64k tokens     | 30B parameters  | Yes          | Yes  | `Qwen/Qwen3-30B-A3B-Instruct-2507-FP8`             |
| [Qwen3 Embedding 4B](/en/documentation/platform/ai-inference/qwen3-embedding-4b/)           | Embedding | Text           | 32k tokens     | 4B parameters   | —            | —    | `Qwen/Qwen3-Embedding-4B`                          |
| [Nanonets-OCR-s](/en/documentation/platform/ai-inference/nanonets-ocr-s/)                   | —         | Text and image | 32k tokens     | —               | —            | —    | `nanonets/Nanonets-OCR-s`                          |
| [GPT-OSS 20B](/en/documentation/platform/ai-inference/gpt-oss-20b/)                         | LLM       | Text           | 131k tokens    | 20B parameters  | Yes          | No   | `gpt-oss-20b`                                      |

---

## Model ids

A model id is the string that selects the model: the first argument to `Azion.AI.run` inside a [function](/en/documentation/platform/functions/), and the `model` field of an HTTP request body. The ids do not share one form. Some are lowercase and hyphenated with no slash, and others carry a publisher prefix and mixed capitalization. Copy an id from the table character for character rather than deriving it from the model name.

---

## Related resources

- [Model invocation](/en/documentation/platform/ai-inference/model-invocation.md): The request body an id travels in, and the two interfaces that carry it.
- [AI Inference](/en/documentation/platform/ai-inference.md): The product that runs the models in this catalog.
- [LoRA Fine-Tune](/en/documentation/platform/ai-inference/lora-fine-tune.md): The extension of AI Inference behind the LoRA column.
- [Vector search](/en/documentation/platform/sql-database/vector-search.md): Where the vectors an embedding model returns are stored and queried.
- [Functions](/en/documentation/platform/functions.md): The product whose runtime binding takes the ids in this table.
- [Glossary](/en/documentation/platform/ai-inference/glossary.md): Where model id, LoRA Fine-Tune, and the rest of the AI Inference vocabulary are defined.
