Model Garden

Browse models across providers with detailed specs, pricing, and performance signals.

Model Filters

Reset Filters

Location

Europe

US

Globe Showing Europe Africa

Global

Show more

CREATORS

Open AI logo

Open AI

Alibaba logo

Alibaba

Google logo

Google

Z.ai logo

Z.ai

Cohere logo

Cohere

xAI logo

xAI

Mistral logo

Mistral

Minimax logo

Minimax

AWS logo

AWS

Meta logo

Meta

Anthropic logo

Anthropic

Jina logo

Jina

Moonshot AI logo

Moonshot AI

Show all

Models

342 models

All

Text

Image

Audio

Speech

Embedding

Moderation

Rerank

NVIDIA logo

Nvidia Nemotron 3 Nano 30b A3b

262144

Nemotron 3 Nano 30B A3B is a compact MoE model optimized for reasoning, coding, and long-context RAG workflows.

by

nvidia

·

·

262144

context

·

$0.06

/M input tokens

·

$0.24

/M output tokens

DeepSeek logo

DeepSeek Chat V3.1

164000

DeepSeek Chat V3.1, a powerful MoE language model with 164K context window. Excels at general chat, coding, and complex reasoning tasks with function calling support.

by

deepseek

·

·

164000

context

·

$0.20

/M input tokens

·

$0.80

/M output tokens

DeepSeek logo

DeepSeek V3.2

163840

DeepSeek V3.2, an advanced MoE language model with 164K context window. Features improved reasoning, function calling, and general intelligence capabilities.

by

deepseek

·

·

163840

context

·

$0.30

/M input tokens

·

$0.50

/M output tokens

DeepSeek logo

DeepSeek V4 Flash

1000000

DeepSeek V4-Flash is a 284B total / 13B active parameter chat model with long-context support and fast, cost-efficient inference.

by

deepseek

·

·

1000000

context

·

$0.25

/M input tokens

·

$0.30

/M output tokens

DeepSeek logo

DeepSeek V4 Pro

1000000

DeepSeek V4-Pro is a 1.6T total / 49B active parameter model with 1M context and top-tier agentic reasoning and coding capabilities.

by

deepseek

·

·

1000000

context

·

$1.75

/M input tokens

·

$3.50

/M output tokens

Minimax logo

MiniMax M2

196608

MiniMax M2, a versatile model with 197K context window. Optimized for coding tasks with function calling and reasoning support.

by

minimax

·

·

196608

context

·

$0.25

/M input tokens

·

$1.00

/M output tokens

Minimax logo

MiniMax M2.5

196608

MiniMax M2.5, a capable model with 197K context window. Supports function calling and reasoning for general purpose tasks.

by

minimax

·

·

196608

context

·

$0.30

/M input tokens

·

$1.20

/M output tokens

Minimax logo

MiniMax M3

1000000

MiniMax-M3 is a frontier multimodal coding model with a 1M context window, agentic reasoning, and tool use.

by

minimax

·

·

1000000

context

·

$0.40

/M input tokens

·

$2.00

/M output tokens

jina logo

Ocr V1

32768

Page-to-Markdown document parser for images and PDFs.

by

jina

·

·

32768

context

·

$0.05

/M input tokens

·

$0.00

/M output tokens

typesafe logo

Jev Latest

64000

Jev classification model: typed yes/no, choice and score questions over text or JSON state. 64k context; state plus longest question limited to 32k tokens.

by

typesafe

·

·

64000

context

·

$0.042

/M input tokens

·

$0.00

/M output tokens

anthropic logo

Claude Fable 5.1

1000000

Claude Fable 5.1

by

anthropic

·

·

1000000

context

·

$10.00

/M input tokens

·

$50.00

/M output tokens

alibaba logo

Qwen Qwen3.8 27b

131072

Qwen 3.8 27B is a multimodal reasoning model with tool use, structured outputs, and long-context support.

by

alibaba

·

·

131072

context

·

$0.99

/M input tokens

·

$1.49

/M output tokens

Cohere logo

Parse V5.0

8192

A multimodal document parsing model that extracts text, tables, lists, forms, images, captions, and layout into Markdown.

by

cohere

·

·

8192

context

·

$0.00

/M input tokens

·

$0.00

/M output tokens

Google logo

Gemini 3.8 Flash

1048576

Gemini 3.8 Flash is our most intelligent Flash model, engineered for long-horizon software engineering, autonomous agents, and complex enterprise workflows—all with the speed and cost efficiency of Flash.

by

google

·

·

1048576

context

·

$0.825

/M input tokens

·

$4.125

/M output tokens

Meta logo

Muse Voice Transcribe 1.0

Speech-to-text model for transcribing supported WAV recordings with optional language bias.

by

meta

·

·

context

·

$0.003

/M input tokens

·

$0.00

/M output tokens

Open AI logo

GPT 6 Astra

1050000

GPT-6 Astra is OpenAI's flagship reasoning model for complex professional and agentic work.

by

openai

·

·

1050000

context

·

$11.00

/M input tokens

·

$55.00

/M output tokens

Alibaba Cloud logo

Qwen Qwen3.8 Flash Next

256000

Qwen3.8 Flash-Next is an ultra-efficient multimodal model with elite coding from 6B active parameters.

by

alibaba

·

·

256000

context

·

$0.20

/M input tokens

·

$0.50

/M output tokens

Anthropic logo

Claude Sonnet 4

1000000

Claude Sonnet 4 is a significant upgrade to Claude Sonnet 3.7, delivering superior coding and reasoning while responding more precisely to your instructions.

by

anthropic

·

·

1000000

context

·

$3.00

/M input tokens

·

$15.00

/M output tokens

Meta logo

Muse Spark 1.3

1048576

Meta's multimodal reasoning model for coding and long-running agentic workflows, with a one-million-token context window, tool calling, and structured outputs.

by

meta

·

·

1048576

context

·

$1.25

/M input tokens

·

$4.25

/M output tokens

Open AI logo

GPT Image 2.5 Flare

Fast image generation and editing from text and image inputs.

by

openai

·

·

context

·

$5.50

/M input tokens

·

$33.00

/M output tokens

Open AI logo

GPT Image 2.5 Sunburst

Image generation and editing from text and image inputs, with a focus on editing precision.

by

openai

·

·

context

·

$5.50

/M input tokens

·

$33.00

/M output tokens

Z.ai logo

GLM 5.3 Flash

1000000

GLM-5.3-Flash is Z.ai's native multimodal GLM-5 model for efficient coding, agentic workflows, visual understanding, and professional document tasks.

by

zai

·

·

1000000

context

·

$0.075

/M input tokens

·

$0.25

/M output tokens

ai21labs logo

Jamba 1.5 Mini

256000

Jamba 1.5 Mini

by

ai21labs

·

·

256000

context

·

$0.20

/M input tokens

·

$0.40

/M output tokens

alibaba logo

Qwen Mt Flash

16384

Qwen-MT Flash

by

alibaba

·

·

16384

context

·

$0.30

/M input tokens

·

$0.90

/M output tokens

alibaba logo

Qwen Mt Lite

16384

Qwen-MT Lite

by

alibaba

·

·

16384

context

·

$0.60

/M input tokens

·

$1.80

/M output tokens

alibaba logo

Qwen Plus

1000000

Qwen Plus

by

alibaba

·

·

1000000

context

·

$0.40

/M input tokens

·

$1.20

/M output tokens

Alibaba Cloud logo

Qwen Qwen3 Coder 30b A3b Instruct

262000

Qwen3 Coder 30B-A3B Instruct by Alibaba, a coding-optimized MoE model with 262K context window. Specialized for code generation and function calling tasks.

by

alibaba

·

·

262000

context

·

$0.06

/M input tokens

·

$0.25

/M output tokens

Alibaba Cloud logo

Qwen Qwen3 Embedding 8b

40960

Qwen3 Embedding 8B by Alibaba, an embedding model hosted on TensorX with 32K context window.

by

alibaba

·

·

40960

context

·

$0.01

/M input tokens

·

$0.00

/M output tokens

Alibaba Cloud logo

Qwen Qwen3 Vl 235b A22b Instruct

131072

Qwen3-VL 235B-A22B

by

alibaba

·

·

131072

context

·

$0.21

/M input tokens

·

$1.90

/M output tokens

Alibaba Cloud logo

Qwen Qwen3.5 122b A10b

256000

Qwen3.5 122B-A10B by Alibaba, a multimodal vision-language MoE model with 122B total and 10B active parameters across 256 experts. Supports text, image and video inputs with a 256K context window for native multimodal agent applications.

by

alibaba

·

·

256000

context

·

$0.50

/M input tokens

·

$3.50

/M output tokens

Create an account and start building today.

Create an account and start building today.

Create an account and start building today.

Create an account and start building today.