
Model Garden
Browse models across providers with detailed specs, pricing, and performance signals.
Model Filters
Reset Filters
Location
Europe
US

Global
Show more
CREATORS
Open AI
Alibaba
Z.ai
Cohere
xAI
Mistral
Minimax
AWS

Meta
Anthropic
Jina
Moonshot AI
Show all
Models
342 models
All
Text
Image
Audio
Speech
Embedding
Moderation
Rerank

Nvidia Nemotron 3 Nano 30b A3b
262144
Nemotron 3 Nano 30B A3B is a compact MoE model optimized for reasoning, coding, and long-context RAG workflows.
by
nvidia
·
·
262144
context
·
$0.06
/M input tokens
·
$0.24
/M output tokens
DeepSeek Chat V3.1
164000
DeepSeek Chat V3.1, a powerful MoE language model with 164K context window. Excels at general chat, coding, and complex reasoning tasks with function calling support.
by
deepseek
·
·
164000
context
·
$0.20
/M input tokens
·
$0.80
/M output tokens
DeepSeek V3.2
163840
DeepSeek V3.2, an advanced MoE language model with 164K context window. Features improved reasoning, function calling, and general intelligence capabilities.
by
deepseek
·
·
163840
context
·
$0.30
/M input tokens
·
$0.50
/M output tokens
DeepSeek V4 Flash
1000000
DeepSeek V4-Flash is a 284B total / 13B active parameter chat model with long-context support and fast, cost-efficient inference.
by
deepseek
·
·
1000000
context
·
$0.25
/M input tokens
·
$0.30
/M output tokens
DeepSeek V4 Pro
1000000
DeepSeek V4-Pro is a 1.6T total / 49B active parameter model with 1M context and top-tier agentic reasoning and coding capabilities.
by
deepseek
·
·
1000000
context
·
$1.75
/M input tokens
·
$3.50
/M output tokens
MiniMax M2
196608
MiniMax M2, a versatile model with 197K context window. Optimized for coding tasks with function calling and reasoning support.
by
minimax
·
·
196608
context
·
$0.25
/M input tokens
·
$1.00
/M output tokens
MiniMax M2.5
196608
MiniMax M2.5, a capable model with 197K context window. Supports function calling and reasoning for general purpose tasks.
by
minimax
·
·
196608
context
·
$0.30
/M input tokens
·
$1.20
/M output tokens
MiniMax M3
1000000
MiniMax-M3 is a frontier multimodal coding model with a 1M context window, agentic reasoning, and tool use.
by
minimax
·
·
1000000
context
·
$0.40
/M input tokens
·
$2.00
/M output tokens

Ocr V1
32768
Page-to-Markdown document parser for images and PDFs.
by
jina
·
·
32768
context
·
$0.05
/M input tokens
·
$0.00
/M output tokens

Jev Latest
64000
Jev classification model: typed yes/no, choice and score questions over text or JSON state. 64k context; state plus longest question limited to 32k tokens.
by
typesafe
·
·
64000
context
·
$0.042
/M input tokens
·
$0.00
/M output tokens
Claude Fable 5.1
1000000
Claude Fable 5.1
by
anthropic
·
·
1000000
context
·
$10.00
/M input tokens
·
$50.00
/M output tokens

Qwen Qwen3.8 27b
131072
Qwen 3.8 27B is a multimodal reasoning model with tool use, structured outputs, and long-context support.
by
alibaba
·
·
131072
context
·
$0.99
/M input tokens
·
$1.49
/M output tokens
Parse V5.0
8192
A multimodal document parsing model that extracts text, tables, lists, forms, images, captions, and layout into Markdown.
by
cohere
·
·
8192
context
·
$0.00
/M input tokens
·
$0.00
/M output tokens
Gemini 3.8 Flash
1048576
Gemini 3.8 Flash is our most intelligent Flash model, engineered for long-horizon software engineering, autonomous agents, and complex enterprise workflows—all with the speed and cost efficiency of Flash.
by
·
·
1048576
context
·
$0.825
/M input tokens
·
$4.125
/M output tokens

Muse Voice Transcribe 1.0
Speech-to-text model for transcribing supported WAV recordings with optional language bias.
by
meta
·
·
context
·
$0.003
/M input tokens
·
$0.00
/M output tokens
GPT 6 Astra
1050000
GPT-6 Astra is OpenAI's flagship reasoning model for complex professional and agentic work.
by
openai
·
·
1050000
context
·
$11.00
/M input tokens
·
$55.00
/M output tokens
Qwen Qwen3.8 Flash Next
256000
Qwen3.8 Flash-Next is an ultra-efficient multimodal model with elite coding from 6B active parameters.
by
alibaba
·
·
256000
context
·
$0.20
/M input tokens
·
$0.50
/M output tokens
Claude Sonnet 4
1000000
Claude Sonnet 4 is a significant upgrade to Claude Sonnet 3.7, delivering superior coding and reasoning while responding more precisely to your instructions.
by
anthropic
·
·
1000000
context
·
$3.00
/M input tokens
·
$15.00
/M output tokens

Muse Spark 1.3
1048576
Meta's multimodal reasoning model for coding and long-running agentic workflows, with a one-million-token context window, tool calling, and structured outputs.
by
meta
·
·
1048576
context
·
$1.25
/M input tokens
·
$4.25
/M output tokens
GPT Image 2.5 Flare
Fast image generation and editing from text and image inputs.
by
openai
·
·
context
·
$5.50
/M input tokens
·
$33.00
/M output tokens
GPT Image 2.5 Sunburst
Image generation and editing from text and image inputs, with a focus on editing precision.
by
openai
·
·
context
·
$5.50
/M input tokens
·
$33.00
/M output tokens
GLM 5.3 Flash
1000000
GLM-5.3-Flash is Z.ai's native multimodal GLM-5 model for efficient coding, agentic workflows, visual understanding, and professional document tasks.
by
zai
·
·
1000000
context
·
$0.075
/M input tokens
·
$0.25
/M output tokens

Jamba 1.5 Mini
256000
Jamba 1.5 Mini
by
ai21labs
·
·
256000
context
·
$0.20
/M input tokens
·
$0.40
/M output tokens

Qwen Mt Flash
16384
Qwen-MT Flash
by
alibaba
·
·
16384
context
·
$0.30
/M input tokens
·
$0.90
/M output tokens

Qwen Mt Lite
16384
Qwen-MT Lite
by
alibaba
·
·
16384
context
·
$0.60
/M input tokens
·
$1.80
/M output tokens

Qwen Plus
1000000
Qwen Plus
by
alibaba
·
·
1000000
context
·
$0.40
/M input tokens
·
$1.20
/M output tokens
Qwen Qwen3 Coder 30b A3b Instruct
262000
Qwen3 Coder 30B-A3B Instruct by Alibaba, a coding-optimized MoE model with 262K context window. Specialized for code generation and function calling tasks.
by
alibaba
·
·
262000
context
·
$0.06
/M input tokens
·
$0.25
/M output tokens
Qwen Qwen3 Embedding 8b
40960
Qwen3 Embedding 8B by Alibaba, an embedding model hosted on TensorX with 32K context window.
by
alibaba
·
·
40960
context
·
$0.01
/M input tokens
·
$0.00
/M output tokens
Qwen Qwen3 Vl 235b A22b Instruct
131072
Qwen3-VL 235B-A22B
by
alibaba
·
·
131072
context
·
$0.21
/M input tokens
·
$1.90
/M output tokens
Qwen Qwen3.5 122b A10b
256000
Qwen3.5 122B-A10B by Alibaba, a multimodal vision-language MoE model with 122B total and 10B active parameters across 256 experts. Supports text, image and video inputs with a 256K context window for native multimodal agent applications.
by
alibaba
·
·
256000
context
·
$0.50
/M input tokens
·
$3.50
/M output tokens
