Solutions
By challenge
Solutions
By challenge

LLM Providers Available on Orq.ai
Orq.ai's AI Router connects you to 300+ models from 28 providers through one OpenAI-compatible endpoint. Compare providers, switch models with a string change, and run on the best model at the lowest cost — without vendor lock-in.
LLM Providers Available on Orq.ai
Orq.ai's AI Router connects you to 300+ models from 28 providers through one OpenAI-compatible endpoint. Compare providers, switch models with a string change, and run on the best model at the lowest cost — without vendor lock-in.


0+
0+
Models
0
0
Providers
0
0
Unified API
Supported LLM providers
CAPABILITIES:
All
Chat
Code
Embeddings
Image Generation
Reasoning
Speech
Vision
Anthropic
Use Claude Opus 4.7, Sonnet 4.6, Haiku 4.5, and other supported Claude models through one API.
Chat
Code
Reasoning
Vision
Models:
claude-opus-5
claude-sonnet-5
claude-fable-5
Open AI
Use OpenAI's foundation models through a single Orq.ai API. Route models such as GPT-4.1, GPT-4.1-mini, o3-mini, and GPT-4o-class models via Orq's AI Router for chat, reasoning, coding, and multimodal workloads.
Chat
Code
Image Generation
Reasoning
Speech
Vision
Models:
gpt-5.5 (EU)
gpt-5.6-luna (EU)
gpt-5.6-sol (EU)
Google AI
Use Google’s Gemini models through a single Orq.ai API. Route models such as Gemini 3.1 Pro, Gemini 2.5 Flash, and Gemini 2.0 Flash‑Lite via Orq’s AI Router for chat, reasoning, coding, and multimodal workloads.
Chat
Code
Embeddings
Image Generation
Reasoning
Speech
Vision
Models:
gemini-3.5-flash-lite (Gemini API)
gemini-3.6-flash (Gemini API)
Gemini 3 Pro Image
AWS
Use AWS Bedrock’s foundation models through a single Orq.ai API. Route models such as Amazon Nova, Amazon Titan Text, and compatible third‑party models exposed via Bedrock through Orq’s AI Router for chat, reasoning, coding, and multimodal workloads.
Chat
Code
Embeddings
Reasoning
Vision
Models:
eu.anthropic.claude-opus-5
global.anthropic.claude-opus-5
us.anthropic.claude-opus-5
Z.ai
Use Z.ai’s GLM‑5 family through a single Orq.ai integration. Route models such as GLM‑5.2, GLM‑5.1, GLM‑5, GLM‑5‑Turbo, GLM‑4.7, and GLM‑4.7‑FlashX via Orq’s AI Router for chat, reasoning, coding, multilingual tasks, vision, and cost‑efficient high‑volume workloads.
Chat
Image Generation
Reasoning
Vision
Models:
glm-5.2
glm-5.1
glm-5v-turbo
xAI
Use xAI’s Grok frontier models through a single Orq.ai integration. Route models such as Grok 4, Grok 4.20, Grok 4.1 Fast, Grok 3, Grok 3 Mini, and Grok 2 Vision via Orq’s AI Router for chat, reasoning, coding, multimodal vision, and real‑time X search‑grounded workloads.
Chat
Code
Reasoning
Vision
Models:
grok-4.5
grok-4.3
grok-build-0.1
Together AI
Use Together AI’s open‑model platform through a single Orq.ai integration. Route leading models such as Llama 4 Scout/Maverick, Llama 3.1/3.3, DeepSeek V3/R1, Mixtral, Qwen, Gemma, and many code‑tuned variants via Orq’s AI Router for chat, reasoning, coding, and multimodal workloads.
Chat
Reasoning
Vision
Models:
GLM-5.2
Kimi K2.7 Code
Kimi K2.6
Wafer
Use Wafer’s ultra‑fast open‑source LLMs through a single Orq.ai integration. Route models such as GLM‑5.2, GLM‑5.1, Kimi‑K2.6/K2.7 Code, Qwen3.5‑397B, Qwen3.6‑35B, Qwen3.7‑Max, DeepSeek V4 Pro/Flash, MiniMax M3, and other Turbo models via Orq’s AI Router for chat, reasoning, coding, and agentic workloads.
Chat
Reasoning
Vision
Models:
GLM-5.2
GLM5.2-Fast
GLM-5.1
Perplexity
Use Perplexity’s Sonar models through a single Orq.ai API. Route models such as Sonar, Sonar Pro, Sonar Reasoning Pro, and Sonar Deep Research via Orq’s AI Router for chat, reasoning, coding, and web‑augmented “ask the internet” workloads.
Chat
Vision
Models:
sonar
sonar-deep-research
sonar-pro

Scaleway
Use Scaleway’s Generative API and managed inference through a single Orq.ai integration. Route open‑source models such as Llama 3.3, Mistral Medium/Small, Qwen, Gemma, Pixtral, and audio/embedding models via Orq’s AI Router for chat, reasoning, coding, vision, and GPU‑backed workloads.
Chat
Reasoning
Speech
Vision
Models:
glm-5.2
gemma-4-26b-a4b-it
qwen3.6-35b-a3b
TensorX
Use Tensorix’s private, EU‑hosted inference API through a single Orq.ai integration. Route leading open‑source and open‑weight models such as DeepSeek V4, GLM‑5.x, Qwen3.5, Kimi K2.x, MiniMax M‑series, Nemotron, and GPT‑OSS via Orq’s AI Router for chat, reasoning, coding, and multimodal workloads.
Chat
Reasoning
Vision
Models:
GLM-5.1
GLM-5.2
Kimi-K2.7-Code
Moonshot AI
Use Moonshot AI’s Kimi models through a single Orq.ai API. Route models such as Kimi K2.6, Kimi K2.5, and Moonshot‑v1‑8k/32k/128k via Orq’s AI Router for chat, reasoning, coding, and long‑context workloads.
Chat
Reasoning
Vision
Models:
Kimi K3
kimi-k2.7-code-highspeed
moonshot-v1-128k
OpenAI-compatible
Use OpenAI‑compatible models through a single Orq.ai API. Route traffic to providers like Groq, Together AI, Mistral, Moonshot, and any custom OpenAI‑compatible deployment via Orq’s AI Router, while keeping your existing OpenAI SDK and request shapes.
Models:
No models available
Mistral
Use Mistral’s foundation models through a single Orq.ai API. Route models such as Mistral Large 3, Mistral Medium 3, Mistral Small 4, and Ministral 8B via Orq’s AI Router for chat, reasoning, coding, and multimodal workloads.
Chat
Embeddings
Reasoning
Vision
Models:
Mistral Moderation 2603
codestral-2508
codestral-embed-2505
Leonardo AI
Use Leonardo AI’s image and video generation APIs through a single Orq.ai interface. Route models such as FLUX.2 Pro, Phoenix, Lucid, and other Leonardo‑hosted styles via Orq’s AI Router for creative, marketing, product, and multimodal workloads.
Models:
leonardo-diffusion-xl
leonardo-kino-xl
leonardo-lightning-xl
Jina
Use Jina’s search and retrieval APIs through a single Orq.ai endpoint. Route Jina’s embeddings, rerankers, and Reader API via Orq’s AI Router to power RAG, semantic search, and retrieval‑heavy workloads alongside your LLM stack.
Chat
Embeddings
Speech
Models:
jina-embeddings-v5-omni-nano
jina-embeddings-v5-omni-small
jina-embeddings-v5-text-nano
Inceptron
Use Inceptron’s optimized open‑model endpoints through a single Orq.ai API. Route models such as Kimi‑K2.6, GLM‑5.1, MiniMax‑M2.5, and Llama 3.3 70B Instruct via Orq’s AI Router for high‑performance chat, reasoning, coding, and long‑context workloads.
Chat
Reasoning
Vision
Models:
Kimi-K2.7-Code
GLM-5.2
Kimi-K2.6
H Company
Use H Company’s Holo model family through a single Orq.ai API. Route models such as Holo3‑122B‑A10B and Holo3‑35B‑A3B via Orq’s AI Router for agentic chat, reasoning, coding, and “computer‑use” workflows.
Chat
Reasoning
Vision
Models:
Holo3 122B A10B
Holo3.1 35B A3B
Vertex AI
Use Google Vertex AI’s Gemini models through a single Orq.ai API. Route models such as Gemini 2.5 Pro, Gemini 2.5 Flash, and Gemini 1.5 Flash via Orq’s AI Router for chat, reasoning, coding, and multimodal workloads.
Models:
No models available
Groq
Use Groq’s LPU‑accelerated models through a single Orq.ai API. Route popular open models such as Llama 3.1, Llama 3.3, Qwen, Gemma, Mixtral, and GPT‑OSS variants via Orq’s AI Router for ultra‑fast chat, reasoning, and coding workloads.
Chat
Reasoning
Speech
Vision
Models:
qwen/qwen3.6-27b
allam-2-7b
canopylabs/orpheus-arabic-saudi
DeepSeek
Use DeepSeek’s LLMs through a single Orq.ai API. Route models such as DeepSeek V‑series (for example V3/V4) and DeepSeek R‑series (reasoning models) via Orq’s AI Router for chat, reasoning, coding, and high‑volume workloads.
Chat
Reasoning
Models:
deepseek-v4-flash
deepseek-v4-pro
deepseek-chat
ElevenLabs
Use ElevenLabs’ AI voice and audio models through a single Orq.ai API. Route text‑to‑speech (TTS), voice cloning, and conversational audio workloads through Orq’s AI Router to add speech to your LLM apps, agents, and workflows.
Chat
Speech
Models:
scribe_v2
scribe_v1
eleven_flash_v2
Fal AI
Use Fal’s generative media APIs through a single Orq.ai interface. Route image, video, and other supported models (such as FLUX, Seedream, Wan, and Kling) via Orq’s AI Router for creative, marketing, product, and multimodal workloads.
Models:
flux-2-dev
flux-2-flex
flux-2-max
Bytedance
Use ByteDance’s foundation models through a single Orq.ai API. Route models such as Doubao (and other supported ByteDance / Volcengine models) through Orq’s AI Router for chat, reasoning, coding, and multimodal workloads.
Models:
Seedream-4-5-251128
Seedream-4-0-250828
Cerebras
Use Cerebras‑hosted models through a single Orq.ai API. Route high‑performance open models such as GPT‑OSS, Llama, Qwen, and other supported models served by Cerebras Cloud through Orq’s AI Router for chat, reasoning, coding, and multimodal workloads.
Chat
Reasoning
Vision
Models:
cerebras/gemma-4-31b
zai-glm-4.7
cerebras/gpt-oss-120b
Cohere
Use Cohere’s Command family through a single Orq.ai API. Route models such as Command, Command R, and lighter Command variants via Orq’s AI Router for chat, reasoning, coding, and retrieval‑augmented workloads.
Chat
Embeddings
Reasoning
Vision
Models:
c4ai-aya-expanse-32b
c4ai-aya-vision-32b
command-r7b-arabic-02-2025

Contextual AI
Use Contextual AI’s models through a single Orq.ai API. Route models such as Generate, Rerank, and LMUnit via Orq’s AI Router for grounded chat, reasoning, evaluation, and retrieval‑augmented workloads.
Models:
No models available
Microsoft Azure
Use Azure‑hosted OpenAI models through a single Orq.ai API. Route models such as GPT‑4o, GPT‑4o‑mini, and other Azure OpenAI deployments via Orq’s AI Router for chat, reasoning, coding, and multimodal workloads.
Chat
Reasoning
Vision
Models:
gpt-5.6-luna
gpt-5.6-luna
gpt-5.6-sol
Alibaba Cloud
Use Alibaba Cloud’s Qwen model family through a single Orq.ai API. Route Qwen-Max, Qwen-Plus, Qwen-Turbo, and other supported Qwen models via Orq’s AI Router for chat, reasoning, coding, and multimodal workloads.
Chat
Reasoning
Speech
Vision
Models:
qwen3.7-max
qwen3.7-plus
qwen3.7-max-2026-06-08
