AI MODEL DIRECTORY

Model Wiki

Explore major large language models with capability notes, technical parameters, use cases and trade-offs

0models0Provider0Family

domainVolcengine1 models

Volcenginedoubao

Doubao Lite 32K

Doubao Lite 32K model profile for capabilities, pricing and use cases

Doubao Lite 32K is a large language model from Volcengine. This entry summarizes its positioning, typical use cases, strengths, limitations and related pricing signals for quick comparison.

domainXiaomi (MiMo)6 models

Xiaomi (MiMo)mimo

MiMo-V2.5-Pro-UltraSpeed

MiMo-V2.5-Pro-UltraSpeed: a mimo model from Xiaomi, ~1M context, knowledge cutoff 2024-12

MiMo-V2.5-Pro-UltraSpeed is a mimo model from Xiaomi (~1M context, input around $1.305/1M tokens). Suitable for assistants, content generation, knowledge Q&A and business automation

Released June 2026
Xiaomi (MiMo)mimo

MiMo V2.5

MiMo V2.5 for Chinese conversation, content generation and tool-use workflows

MiMo V2.5 is a Xiaomi MiMo model for Chinese conversation, content generation and tool-use workflows, useful for Chinese assistants, tool use and ecosystem-oriented applications.

1000K contextMultimodalInput ¥1/1M tokensOutput ¥2/1M tokens
Xiaomi (MiMo)mimo

MiMo V2.5 Pro

MiMo V2.5 Pro for complex reasoning, coding and long-text tasks

MiMo V2.5 Pro is a Xiaomi MiMo model for complex reasoning, coding and long-text tasks, useful for Chinese assistants, tool use and ecosystem-oriented applications.

1000K contextInput ¥3/1M tokensOutput ¥6/1M tokens
Xiaomi (MiMo)mimo

MiMo V2 Flash

MiMo V2 Flash for low-latency high-frequency chat and quick responses

MiMo V2 Flash is a Xiaomi MiMo model for low-latency high-frequency chat and quick responses, useful for Chinese assistants, tool use and ecosystem-oriented applications.

262K contextInput $0.14/1M tokensOutput $0.28/1M tokens
Xiaomi (MiMo)mimo

MiMo V2 Omni

MiMo V2 Omni for multimodal understanding and integrated interactive experiences

MiMo V2 Omni is a Xiaomi MiMo model for multimodal understanding and integrated interactive experiences, useful for Chinese assistants, tool use and ecosystem-oriented applications.

262K contextMultimodalInput $0.14/1M tokensOutput $0.28/1M tokens
Xiaomi (MiMo)mimo

MiMo V2 Pro

MiMo V2 Pro for complex tasks, coding assistance and business automation

MiMo V2 Pro is a Xiaomi MiMo model for complex tasks, coding assistance and business automation, useful for Chinese assistants, tool use and ecosystem-oriented applications.

1049K contextInput $0.435/1M tokensOutput $0.87/1M tokens

domainBaidu (ERNIE)6 models

Baidu (ERNIE)ernie

ERNIE 4.0 Turbo 8K

ERNIE 4.0 Turbo 8K for Chinese understanding, content generation and enterprise applications

ERNIE 4.0 Turbo 8K is a Baidu ERNIE/Qianfan model for Chinese understanding, content generation and enterprise applications, commonly evaluated for Chinese enterprise workloads.

Baidu (ERNIE)ernie-4.5

ERNIE 4.5 Turbo

ERNIE 4.5 Turbo for multi-scenario Chinese tasks, knowledge Q&A and business assistants

ERNIE 4.5 Turbo is a Baidu ERNIE/Qianfan model for multi-scenario Chinese tasks, knowledge Q&A and business assistants, commonly evaluated for Chinese enterprise workloads.

Input ¥0.8/per_thousand_tokensOutput ¥3.2/per_thousand_tokens
Baidu (ERNIE)ernie-x1

ERNIE X1

ERNIE X1 for complex analysis, logical reasoning and multi-step problem solving

ERNIE X1 is a Baidu ERNIE/Qianfan model for complex analysis, logical reasoning and multi-step problem solving, commonly evaluated for Chinese enterprise workloads.

Baidu (ERNIE)ernie-speed

ERNIE Speed

ERNIE Speed for low-latency conversations and high-frequency basic text tasks

ERNIE Speed is a Baidu ERNIE/Qianfan model for low-latency conversations and high-frequency basic text tasks, commonly evaluated for Chinese enterprise workloads.

Baidu (ERNIE)ernie-lite

ERNIE Lite

ERNIE Lite for cost-sensitive Q&A, summarization and content generation

ERNIE Lite is a Baidu ERNIE/Qianfan model for cost-sensitive Q&A, summarization and content generation, commonly evaluated for Chinese enterprise workloads.

Baidu (ERNIE)ernie

ERNIE 4.0 Turbo 8K

ERNIE 4.0 Turbo 8K model profile for capabilities, pricing and use cases

ERNIE 4.0 Turbo 8K is a large language model from Baidu. This entry summarizes its positioning, typical use cases, strengths, limitations and related pricing signals for quick comparison.

domainMistral31 models

Mistralmistral-medium

Mistral Medium (latest)

Mistral Medium (latest): a mistral-medium model from Mistral

Mistral Medium (latest) is a mistral-medium model from Mistral. Suitable for assistants, content generation, knowledge Q&A and business automation

Released April 2026
Mistralmistral-medium

Mistral Medium 3.5

Mistral Medium 3.5: a mistral-medium model from Mistral

Mistral Medium 3.5 is a mistral-medium model from Mistral. Suitable for assistants, content generation, knowledge Q&A and business automation

Released April 2026
Mistralmistral-small

Mistral Small 4

Mistral Small 4: a mistral-small model from Mistral

Mistral Small 4 is a mistral-small model from Mistral. Suitable for assistants, content generation, knowledge Q&A and business automation

Released March 2026
Mistraldevstral

Devstral 2

Devstral 2: a devstral model from Mistral

Devstral 2 is a devstral model from Mistral. Suitable for assistants, content generation, knowledge Q&A and business automation

Released December 2025
Mistraldevstral

Devstral Small 2

Devstral Small 2: a devstral model from Mistral

Devstral Small 2 is a devstral model from Mistral. Suitable for assistants, content generation, knowledge Q&A and business automation

Released December 2025
Mistraldevstral

Devstral 2

Devstral 2: a devstral model from Mistral

Devstral 2 is a devstral model from Mistral. Suitable for assistants, content generation, knowledge Q&A and business automation

Released December 2025
Mistraldevstral

Devstral 2 (latest)

Devstral 2 (latest): a devstral model from Mistral

Devstral 2 (latest) is a devstral model from Mistral. Suitable for assistants, content generation, knowledge Q&A and business automation

Released December 2025
Mistralmistral-medium

Mistral Medium 3.1

Mistral Medium 3.1: a mistral-medium model from Mistral

Mistral Medium 3.1 is a mistral-medium model from Mistral. Suitable for assistants, content generation, knowledge Q&A and business automation

Released August 2025
Mistraldevstral

Devstral Medium

Devstral Medium: a devstral model from Mistral

Devstral Medium is a devstral model from Mistral. Suitable for assistants, content generation, knowledge Q&A and business automation

Released July 2025
Mistraldevstral

Devstral Small

Devstral Small: a devstral model from Mistral

Devstral Small is a devstral model from Mistral. Suitable for assistants, content generation, knowledge Q&A and business automation

Released July 2025
Mistralmistral-small

Mistral Small 3.2

Mistral Small 3.2: a mistral-small model from Mistral

Mistral Small 3.2 is a mistral-small model from Mistral. Suitable for assistants, content generation, knowledge Q&A and business automation

Released June 2025
Mistraldevstral

Devstral Small 2505

Devstral Small 2505: a devstral model from Mistral

Devstral Small 2505 is a devstral model from Mistral. Suitable for assistants, content generation, knowledge Q&A and business automation

Released May 2025
Mistralmistral-medium

Mistral Medium 3

Mistral Medium 3: a mistral-medium model from Mistral

Mistral Medium 3 is a mistral-medium model from Mistral. Suitable for assistants, content generation, knowledge Q&A and business automation

Released May 2025
Mistralmagistral-medium

Magistral Medium (latest)

Magistral Medium (latest): a magistral-medium model from Mistral

Magistral Medium (latest) is a magistral-medium model from Mistral. Suitable for assistants, content generation, knowledge Q&A and business automation

Released March 2025
Mistralmagistral-small

Magistral Small

Magistral Small: a magistral-small model from Mistral

Magistral Small is a magistral-small model from Mistral. Suitable for assistants, content generation, knowledge Q&A and business automation

Released March 2025
Mistralmistral-large

Mistral Large 2.1

Mistral Large 2.1: a mistral-large model from Mistral

Mistral Large 2.1 is a mistral-large model from Mistral. Suitable for assistants, content generation, knowledge Q&A and business automation

Released November 2024
Mistralmistral-large

Mistral Large 3

Mistral Large 3: a mistral-large model from Mistral

Mistral Large 3 is a mistral-large model from Mistral. Suitable for assistants, content generation, knowledge Q&A and business automation

Released November 2024
Mistralministral

Ministral 3B (latest)

Ministral 3B (latest): a ministral model from Mistral

Ministral 3B (latest) is a ministral model from Mistral. Suitable for assistants, content generation, knowledge Q&A and business automation

Released October 2024
Mistralpixtral

Pixtral 12B

Pixtral 12B: a pixtral model from Mistral

Pixtral 12B is a pixtral model from Mistral. Suitable for assistants, content generation, knowledge Q&A and business automation

Released September 2024
Mistralmistral-nemo

Mistral Nemo

Mistral Nemo: a mistral-nemo model from Mistral

Mistral Nemo is a mistral-nemo model from Mistral. Suitable for assistants, content generation, knowledge Q&A and business automation

Released July 2024
Mistralmistral-nemo

Open Mistral Nemo

Open Mistral Nemo: a mistral-nemo model from Mistral

Open Mistral Nemo is a mistral-nemo model from Mistral. Suitable for assistants, content generation, knowledge Q&A and business automation

Released July 2024
Mistralmixtral

Mixtral 8x22B

Mixtral 8x22B: a mixtral model from Mistral

Mixtral 8x22B is a mixtral model from Mistral. Suitable for assistants, content generation, knowledge Q&A and business automation

Released April 2024
Mistralmixtral

Mixtral 8x7B

Mixtral 8x7B: a mixtral model from Mistral

Mixtral 8x7B is a mixtral model from Mistral. Suitable for assistants, content generation, knowledge Q&A and business automation

Released December 2023
Mistralmistral-embed

Mistral Embed

Mistral Embed: a mistral-embed model from Mistral

Mistral Embed is a mistral-embed model from Mistral. Suitable for assistants, content generation, knowledge Q&A and business automation

Released December 2023
Mistralmistral

Mistral 7B

Mistral 7B: a mistral model from Mistral

Mistral 7B is a mistral model from Mistral. Suitable for assistants, content generation, knowledge Q&A and business automation

Released September 2023
Mistralcodestral

Codestral

Codestral for code completion, code generation and developer assistance

Codestral is a Mistral model for code completion, code generation and developer assistance, with options across enterprise, coding, vision and open-weight workflows.

Mistralmistral-small

Mistral Small

Mistral Small for low-latency high-concurrency and cost-sensitive workloads

Mistral Small is a Mistral model for low-latency high-concurrency and cost-sensitive workloads, with options across enterprise, coding, vision and open-weight workflows.

Mistralmistral-large

Mistral Large

Mistral Large for complex reasoning, enterprise Q&A and multilingual tasks

Mistral Large is a Mistral model for complex reasoning, enterprise Q&A and multilingual tasks, with options across enterprise, coding, vision and open-weight workflows.

Mistralministral

Ministral 8B

Ministral 8B for edge deployment, low-cost usage and basic text tasks

Ministral 8B is a Mistral model for edge deployment, low-cost usage and basic text tasks, with options across enterprise, coding, vision and open-weight workflows.

Mistralpixtral

Pixtral Large

Pixtral Large for image-text understanding, visual Q&A and multimodal analysis

Pixtral Large is a Mistral model for image-text understanding, visual Q&A and multimodal analysis, with options across enterprise, coding, vision and open-weight workflows.

Mistralmixtral

Mixtral 8x7B

Mixtral 8x7B for general language tasks, research and self-hosted deployments

Mixtral 8x7B is a Mistral model for general language tasks, research and self-hosted deployments, with options across enterprise, coding, vision and open-weight workflows.

domainCohere20 models

Coherenorth

North Mini Code

North Mini Code: a north model from Cohere

North Mini Code is a north model from Cohere. Suitable for assistants, content generation, knowledge Q&A and business automation

Released June 2026
Coherecommand-a

Command A Plus

Command A Plus: a command-a model from Cohere

Command A Plus is a command-a model from Cohere. Suitable for assistants, content generation, knowledge Q&A and business automation

Released May 2026
Coherecommand-a

Command A Translate

Command A Translate: a command-a model from Cohere

Command A Translate is a command-a model from Cohere. Suitable for assistants, content generation, knowledge Q&A and business automation

Released August 2025
Coherecommand-a

Command A Reasoning

Command A Reasoning: a command-a model from Cohere

Command A Reasoning is a command-a model from Cohere. Suitable for assistants, content generation, knowledge Q&A and business automation

Released August 2025
Coherecommand-a

Command A Vision

Command A Vision: a command-a model from Cohere

Command A Vision is a command-a model from Cohere. Suitable for assistants, content generation, knowledge Q&A and business automation

Released July 2025
Coherecommand-a

Command A

Command A: a command-a model from Cohere

Command A is a command-a model from Cohere. Suitable for assistants, content generation, knowledge Q&A and business automation

Released March 2025
Cohere

Aya Vision 32B

Aya Vision 32B: a AI model from Cohere

Aya Vision 32B is a AI model from Cohere. Suitable for assistants, content generation, knowledge Q&A and business automation

Released March 2025
Cohere

Aya Vision 8B

Aya Vision 8B: a AI model from Cohere

Aya Vision 8B is a AI model from Cohere. Suitable for assistants, content generation, knowledge Q&A and business automation

Released March 2025
Coherecommand-r

Command R7B Arabic

Command R7B Arabic: a command-r model from Cohere

Command R7B Arabic is a command-r model from Cohere. Suitable for assistants, content generation, knowledge Q&A and business automation

Released February 2025
Coherecommand-r

Command R7B

Command R7B: a command-r model from Cohere

Command R7B is a command-r model from Cohere. Suitable for assistants, content generation, knowledge Q&A and business automation

Released December 2024
Cohere

Aya Expanse 8B

Aya Expanse 8B: a AI model from Cohere

Aya Expanse 8B is a AI model from Cohere. Suitable for assistants, content generation, knowledge Q&A and business automation

Released October 2024
Cohere

Aya Expanse 32B

Aya Expanse 32B: a AI model from Cohere

Aya Expanse 32B is a AI model from Cohere. Suitable for assistants, content generation, knowledge Q&A and business automation

Released October 2024
Coherecommand-r

Command R

Command R: a command-r model from Cohere

Command R is a command-r model from Cohere. Suitable for assistants, content generation, knowledge Q&A and business automation

Released August 2024
Coherecommand-r

Command R+

Command R+: a command-r model from Cohere

Command R+ is a command-r model from Cohere. Suitable for assistants, content generation, knowledge Q&A and business automation

Released August 2024
Coherecommand

Command A

Command A for enterprise agents and complex generation tasks

Command A is a Cohere generation model for enterprise Q&A, RAG, agents and multilingual generation.

Coherecommand-r

Command R+

Command R+ for advanced retrieval-augmented generation, tool use and enterprise Q&A

Command R+ is a Cohere generation model for enterprise Q&A, RAG, agents and multilingual generation.

Coherecommand-r

Command R

Command R for RAG, long-context and multilingual Q&A

Command R is a Cohere generation model for enterprise Q&A, RAG, agents and multilingual generation.

Coherecommand

Command Light

Command Light for low-latency text generation and basic chat

Command Light is a Cohere generation model for enterprise Q&A, RAG, agents and multilingual generation.

Cohereembed

Embed v4.0

Embed v4.0 for semantic search, clustering and RAG knowledge-base indexing

Embed v4.0 is a Cohere embedding model for semantic search, RAG indexing, clustering and similarity workflows.

Coherererank

Rerank v3.5

Rerank v3.5 for reranking search results and improving RAG answer quality

Rerank v3.5 is a Cohere reranking model for search reranking, RAG refinement and answer-quality improvements.

domainMeta Llama13 models

Meta Llamallama

Llama-4-Maverick-17B-128E-Instruct-FP8

Llama-4-Maverick-17B-128E-Instruct-FP8: a llama model from llama, ~128K context, knowledge cutoff 2024-08

Llama-4-Maverick-17B-128E-Instruct-FP8 is a llama model from llama (~128K context, input around $0/1M tokens). Suitable for assistants, content generation, knowledge Q&A and business automation

128K contextMultimodalInput $0/1M tokensOutput $0/1M tokensReleased April 2025
Meta Llamallama

Cerebras-Llama-4-Scout-17B-16E-Instruct

Cerebras-Llama-4-Scout-17B-16E-Instruct: a llama model from llama, ~128K context, knowledge cutoff 2025-01

Cerebras-Llama-4-Scout-17B-16E-Instruct is a llama model from llama (~128K context, input around $0/1M tokens). Suitable for assistants, content generation, knowledge Q&A and business automation

128K contextInput $0/1M tokensOutput $0/1M tokensReleased April 2025
Meta Llamallama

Llama-4-Scout-17B-16E-Instruct-FP8

Llama-4-Scout-17B-16E-Instruct-FP8: a llama model from llama, ~128K context, knowledge cutoff 2024-08

Llama-4-Scout-17B-16E-Instruct-FP8 is a llama model from llama (~128K context, input around $0/1M tokens). Suitable for assistants, content generation, knowledge Q&A and business automation

128K contextMultimodalInput $0/1M tokensOutput $0/1M tokensReleased April 2025
Meta Llamallama

Cerebras-Llama-4-Maverick-17B-128E-Instruct

Cerebras-Llama-4-Maverick-17B-128E-Instruct: a llama model from llama, ~128K context, knowledge cutoff 2025-01

Cerebras-Llama-4-Maverick-17B-128E-Instruct is a llama model from llama (~128K context, input around $0/1M tokens). Suitable for assistants, content generation, knowledge Q&A and business automation

128K contextInput $0/1M tokensOutput $0/1M tokensReleased April 2025
Meta Llamallama

Groq-Llama-4-Maverick-17B-128E-Instruct

Groq-Llama-4-Maverick-17B-128E-Instruct: a llama model from llama, ~128K context, knowledge cutoff 2025-01

Groq-Llama-4-Maverick-17B-128E-Instruct is a llama model from llama (~128K context, input around $0/1M tokens). Suitable for assistants, content generation, knowledge Q&A and business automation

128K contextInput $0/1M tokensOutput $0/1M tokensReleased April 2025
Meta Llamallama

Llama-3.3-8B-Instruct

Llama-3.3-8B-Instruct: a llama model from llama, ~128K context, knowledge cutoff 2023-12

Llama-3.3-8B-Instruct is a llama model from llama (~128K context, input around $0/1M tokens). Suitable for assistants, content generation, knowledge Q&A and business automation

128K contextInput $0/1M tokensOutput $0/1M tokensReleased December 2024
Meta Llamallama

Llama-3.3-70B-Instruct

Llama-3.3-70B-Instruct: a llama model from llama, ~128K context, knowledge cutoff 2023-12

Llama-3.3-70B-Instruct is a llama model from llama (~128K context, input around $0/1M tokens). Suitable for assistants, content generation, knowledge Q&A and business automation

128K contextInput $0/1M tokensOutput $0/1M tokensReleased December 2024
Meta Llamallama-3.1

Llama 3.1 8B

Llama 3.1 8B for local deployment, basic chat and low-cost inference

Llama 3.1 8B is a Meta Llama model for local deployment, basic chat and low-cost inference, commonly evaluated for open ecosystems, self-hosting and customizable AI products.

Meta Llamallama-3.1

Llama 3.1 70B

Llama 3.1 70B for general language understanding, generation and enterprise self-hosting

Llama 3.1 70B is a Meta Llama model for general language understanding, generation and enterprise self-hosting, commonly evaluated for open ecosystems, self-hosting and customizable AI products.

Meta Llamallama-3.1

Llama 3.1 405B

Llama 3.1 405B for complex reasoning, multilingual tasks and high-quality generation

Llama 3.1 405B is a Meta Llama model for complex reasoning, multilingual tasks and high-quality generation, commonly evaluated for open ecosystems, self-hosting and customizable AI products.

Meta Llamallama-3.2

Llama 3.2 Vision

Llama 3.2 Vision for image-text understanding, visual Q&A and multimodal applications

Llama 3.2 Vision is a Meta Llama model for image-text understanding, visual Q&A and multimodal applications, commonly evaluated for open ecosystems, self-hosting and customizable AI products.

Meta Llamallama-4

Llama 4 Scout

Llama 4 Scout for multimodal, long-context and efficient reasoning workflows

Llama 4 Scout is a Meta Llama model for multimodal, long-context and efficient reasoning workflows, commonly evaluated for open ecosystems, self-hosting and customizable AI products.

128K contextMultimodalInput $0/1M tokensOutput $0/1M tokens
Meta Llamallama-4

Llama 4 Maverick

Llama 4 Maverick for complex multimodal tasks, agents and high-quality generation

Llama 4 Maverick is a Meta Llama model for complex multimodal tasks, agents and high-quality generation, commonly evaluated for open ecosystems, self-hosting and customizable AI products.

128K contextMultimodalInput $0/1M tokensOutput $0/1M tokens