LLM Providers Available on Orq.ai

Orq.ai's AI Router connects you to 300+ models from 28 providers through one OpenAI-compatible endpoint. Compare providers, switch models with a string change, and run on the best model at the lowest cost — without vendor lock-in.

LLM Providers Available on Orq.ai

Orq.ai's AI Router connects you to 300+ models from 28 providers through one OpenAI-compatible endpoint. Compare providers, switch models with a string change, and run on the best model at the lowest cost — without vendor lock-in.

0+

0+

Models

0

0

Providers

0

0

Unified API

Supported LLM providers

CAPABILITIES:

All

Chat

Code

Embeddings

Image Generation

Reasoning

Speech

Vision

Anthropic

Use Claude Opus 4.7, Sonnet 4.6, Haiku 4.5, and other supported Claude models through one API.

Chat

Code

Reasoning

Vision

Models:

claude-opus-5

claude-sonnet-5

claude-fable-5

Open AI

Use OpenAI's foundation models through a single Orq.ai API. Route models such as GPT-4.1, GPT-4.1-mini, o3-mini, and GPT-4o-class models via Orq's AI Router for chat, reasoning, coding, and multimodal workloads.

Chat

Code

Image Generation

Reasoning

Speech

Vision

Models:

gpt-5.5 (EU)

gpt-5.6-luna (EU)

gpt-5.6-sol (EU)

Google AI

Use Google’s Gemini models through a single Orq.ai API. Route models such as Gemini 3.1 Pro, Gemini 2.5 Flash, and Gemini 2.0 Flash‑Lite via Orq’s AI Router for chat, reasoning, coding, and multimodal workloads.

Chat

Code

Embeddings

Image Generation

Reasoning

Speech

Vision

Models:

gemini-3.5-flash-lite (Gemini API)

gemini-3.6-flash (Gemini API)

Gemini 3 Pro Image

AWS

Use AWS Bedrock’s foundation models through a single Orq.ai API. Route models such as Amazon Nova, Amazon Titan Text, and compatible third‑party models exposed via Bedrock through Orq’s AI Router for chat, reasoning, coding, and multimodal workloads.

Chat

Code

Embeddings

Reasoning

Vision

Models:

eu.anthropic.claude-opus-5

global.anthropic.claude-opus-5

us.anthropic.claude-opus-5

Z.ai

Use Z.ai’s GLM‑5 family through a single Orq.ai integration. Route models such as GLM‑5.2, GLM‑5.1, GLM‑5, GLM‑5‑Turbo, GLM‑4.7, and GLM‑4.7‑FlashX via Orq’s AI Router for chat, reasoning, coding, multilingual tasks, vision, and cost‑efficient high‑volume workloads.

Chat

Image Generation

Reasoning

Vision

Models:

glm-5.2

glm-5.1

glm-5v-turbo

xAI

Use xAI’s Grok frontier models through a single Orq.ai integration. Route models such as Grok 4, Grok 4.20, Grok 4.1 Fast, Grok 3, Grok 3 Mini, and Grok 2 Vision via Orq’s AI Router for chat, reasoning, coding, multimodal vision, and real‑time X search‑grounded workloads.

Chat

Code

Reasoning

Vision

Models:

grok-4.5

grok-4.3

grok-build-0.1

Together AI

Use Together AI’s open‑model platform through a single Orq.ai integration. Route leading models such as Llama 4 Scout/Maverick, Llama 3.1/3.3, DeepSeek V3/R1, Mixtral, Qwen, Gemma, and many code‑tuned variants via Orq’s AI Router for chat, reasoning, coding, and multimodal workloads.

Chat

Reasoning

Vision

Models:

GLM-5.2

Kimi K2.7 Code

Kimi K2.6

Wafer

Use Wafer’s ultra‑fast open‑source LLMs through a single Orq.ai integration. Route models such as GLM‑5.2, GLM‑5.1, Kimi‑K2.6/K2.7 Code, Qwen3.5‑397B, Qwen3.6‑35B, Qwen3.7‑Max, DeepSeek V4 Pro/Flash, MiniMax M3, and other Turbo models via Orq’s AI Router for chat, reasoning, coding, and agentic workloads.

Chat

Reasoning

Vision

Models:

GLM-5.2

GLM5.2-Fast

GLM-5.1

Perplexity

Use Perplexity’s Sonar models through a single Orq.ai API. Route models such as Sonar, Sonar Pro, Sonar Reasoning Pro, and Sonar Deep Research via Orq’s AI Router for chat, reasoning, coding, and web‑augmented “ask the internet” workloads.

Chat

Vision

Models:

sonar

sonar-deep-research

sonar-pro

Scaleway

Use Scaleway’s Generative API and managed inference through a single Orq.ai integration. Route open‑source models such as Llama 3.3, Mistral Medium/Small, Qwen, Gemma, Pixtral, and audio/embedding models via Orq’s AI Router for chat, reasoning, coding, vision, and GPU‑backed workloads.

Chat

Reasoning

Speech

Vision

Models:

glm-5.2

gemma-4-26b-a4b-it

qwen3.6-35b-a3b

TensorX

Use Tensorix’s private, EU‑hosted inference API through a single Orq.ai integration. Route leading open‑source and open‑weight models such as DeepSeek V4, GLM‑5.x, Qwen3.5, Kimi K2.x, MiniMax M‑series, Nemotron, and GPT‑OSS via Orq’s AI Router for chat, reasoning, coding, and multimodal workloads.

Chat

Reasoning

Vision

Models:

GLM-5.1

GLM-5.2

Kimi-K2.7-Code

Moonshot AI

Use Moonshot AI’s Kimi models through a single Orq.ai API. Route models such as Kimi K2.6, Kimi K2.5, and Moonshot‑v1‑8k/32k/128k via Orq’s AI Router for chat, reasoning, coding, and long‑context workloads.

Chat

Reasoning

Vision

Models:

Kimi K3

kimi-k2.7-code-highspeed

moonshot-v1-128k

OpenAI-compatible

Use OpenAI‑compatible models through a single Orq.ai API. Route traffic to providers like Groq, Together AI, Mistral, Moonshot, and any custom OpenAI‑compatible deployment via Orq’s AI Router, while keeping your existing OpenAI SDK and request shapes.

Models:

No models available

Mistral

Use Mistral’s foundation models through a single Orq.ai API. Route models such as Mistral Large 3, Mistral Medium 3, Mistral Small 4, and Ministral 8B via Orq’s AI Router for chat, reasoning, coding, and multimodal workloads.

Chat

Embeddings

Reasoning

Vision

Models:

Mistral Moderation 2603

codestral-2508

codestral-embed-2505

Leonardo AI

Use Leonardo AI’s image and video generation APIs through a single Orq.ai interface. Route models such as FLUX.2 Pro, Phoenix, Lucid, and other Leonardo‑hosted styles via Orq’s AI Router for creative, marketing, product, and multimodal workloads.

Models:

leonardo-diffusion-xl

leonardo-kino-xl

leonardo-lightning-xl

Jina

Use Jina’s search and retrieval APIs through a single Orq.ai endpoint. Route Jina’s embeddings, rerankers, and Reader API via Orq’s AI Router to power RAG, semantic search, and retrieval‑heavy workloads alongside your LLM stack.

Chat

Embeddings

Speech

Models:

jina-embeddings-v5-omni-nano

jina-embeddings-v5-omni-small

jina-embeddings-v5-text-nano

Inceptron

Use Inceptron’s optimized open‑model endpoints through a single Orq.ai API. Route models such as Kimi‑K2.6, GLM‑5.1, MiniMax‑M2.5, and Llama 3.3 70B Instruct via Orq’s AI Router for high‑performance chat, reasoning, coding, and long‑context workloads.

Chat

Reasoning

Vision

Models:

Kimi-K2.7-Code

GLM-5.2

Kimi-K2.6

H Company

Use H Company’s Holo model family through a single Orq.ai API. Route models such as Holo3‑122B‑A10B and Holo3‑35B‑A3B via Orq’s AI Router for agentic chat, reasoning, coding, and “computer‑use” workflows.

Chat

Reasoning

Vision

Models:

Holo3 122B A10B

Holo3.1 35B A3B

Vertex AI

Use Google Vertex AI’s Gemini models through a single Orq.ai API. Route models such as Gemini 2.5 Pro, Gemini 2.5 Flash, and Gemini 1.5 Flash via Orq’s AI Router for chat, reasoning, coding, and multimodal workloads.

Models:

No models available

Groq

Use Groq’s LPU‑accelerated models through a single Orq.ai API. Route popular open models such as Llama 3.1, Llama 3.3, Qwen, Gemma, Mixtral, and GPT‑OSS variants via Orq’s AI Router for ultra‑fast chat, reasoning, and coding workloads.

Chat

Reasoning

Speech

Vision

Models:

qwen/qwen3.6-27b

allam-2-7b

canopylabs/orpheus-arabic-saudi

DeepSeek

Use DeepSeek’s LLMs through a single Orq.ai API. Route models such as DeepSeek V‑series (for example V3/V4) and DeepSeek R‑series (reasoning models) via Orq’s AI Router for chat, reasoning, coding, and high‑volume workloads.

Chat

Reasoning

Models:

deepseek-v4-flash

deepseek-v4-pro

deepseek-chat

ElevenLabs

Use ElevenLabs’ AI voice and audio models through a single Orq.ai API. Route text‑to‑speech (TTS), voice cloning, and conversational audio workloads through Orq’s AI Router to add speech to your LLM apps, agents, and workflows.

Chat

Speech

Models:

scribe_v2

scribe_v1

eleven_flash_v2

Fal AI

Use Fal’s generative media APIs through a single Orq.ai interface. Route image, video, and other supported models (such as FLUX, Seedream, Wan, and Kling) via Orq’s AI Router for creative, marketing, product, and multimodal workloads.

Models:

flux-2-dev

flux-2-flex

flux-2-max

Bytedance

Use ByteDance’s foundation models through a single Orq.ai API. Route models such as Doubao (and other supported ByteDance / Volcengine models) through Orq’s AI Router for chat, reasoning, coding, and multimodal workloads.

Models:

Seedream-4-5-251128

Seedream-4-0-250828

Cerebras

Use Cerebras‑hosted models through a single Orq.ai API. Route high‑performance open models such as GPT‑OSS, Llama, Qwen, and other supported models served by Cerebras Cloud through Orq’s AI Router for chat, reasoning, coding, and multimodal workloads.

Chat

Reasoning

Vision

Models:

cerebras/gemma-4-31b

zai-glm-4.7

cerebras/gpt-oss-120b

Cohere

Use Cohere’s Command family through a single Orq.ai API. Route models such as Command, Command R, and lighter Command variants via Orq’s AI Router for chat, reasoning, coding, and retrieval‑augmented workloads.

Chat

Embeddings

Reasoning

Vision

Models:

c4ai-aya-expanse-32b

c4ai-aya-vision-32b

command-r7b-arabic-02-2025

Contextual AI

Use Contextual AI’s models through a single Orq.ai API. Route models such as Generate, Rerank, and LMUnit via Orq’s AI Router for grounded chat, reasoning, evaluation, and retrieval‑augmented workloads.

Models:

No models available

Microsoft Azure

Use Azure‑hosted OpenAI models through a single Orq.ai API. Route models such as GPT‑4o, GPT‑4o‑mini, and other Azure OpenAI deployments via Orq’s AI Router for chat, reasoning, coding, and multimodal workloads.

Chat

Reasoning

Vision

Models:

gpt-5.6-luna

gpt-5.6-luna

gpt-5.6-sol

Alibaba Cloud

Use Alibaba Cloud’s Qwen model family through a single Orq.ai API. Route Qwen-Max, Qwen-Plus, Qwen-Turbo, and other supported Qwen models via Orq’s AI Router for chat, reasoning, coding, and multimodal workloads.

Chat

Reasoning

Speech

Vision

Models:

qwen3.7-max

qwen3.7-plus

qwen3.7-max-2026-06-08

Create an account and start building today.

Create an account and start building today.

Create an account and start building today.

Create an account and start building today.