xAI

on Orq.ai

Use xAI’s Grok frontier models through a single Orq.ai integration. Route models such as Grok 4, Grok 4.20, Grok 4.1 Fast, Grok 3, Grok 3 Mini, and Grok 2 Vision via Orq’s AI Router for chat, reasoning, coding, multimodal vision, and real‑time X search‑grounded workloads.

Capabilities:

Chat

Code

Reasoning

Vision

Models Supported:

grok-4.5

grok-4.3

grok-build-0.1

grok-4.20-0309-non-reasoning

grok-4.20-0309-reasoning

Provider HQ:

xAI, founded by Elon Musk, operating globally with Grok APIs and X integrations.

Access xAI through Orq.ai’s AI Router

xAI provides the Grok API for frontier‑tier models with advanced reasoning, tool use, voice, image generation, and real‑time search over the X social graph. Grok endpoints use an OpenAI‑style JSON format, support long context windows, and can be called directly from existing OpenAI SDKs or HTTP clients.

xAI models available on Orq.ai

Model

Type

Context / scope

Best for

Pricing tier

grok-4

Chat / reasoning / coding / search / vision

256K context

Frontier‑tier reasoning, enterprise data extraction, programming, and real‑time X‑grounded answers

Premium – $3.00 input / $15.00 output per 1M tokens

grok-4.20

Chat / reasoning / tools / search

2M context (current SKU)

Long‑context reasoning and agents with lower output cost than many competitors

Mid‑tier frontier – $2.00 input / $6.00 output per 1M tokens; cached input around $0.20 / 1M tokens on discount tiers

grok-4.1-fast / grok-4-fast

Chat / fast / cost‑efficient search

256K context

Fast, cheaper tasks such as high‑volume chat, summarisation, lightweight reasoning, and routing

Cost‑efficient – $0.20 input / $0.50 output per 1M tokens, one of the cheapest frontier‑adjacent APIs

Why use xAI through Orq.ai?

Capability

Provider / models

Direct

Through Orq.ai

Chat & reasoning

Grok 4, Grok 4.20, Grok 4-fast, Grok 3/3-mini

Call Grok models directly via the xAI API for chat, reasoning, enterprise tasks, and tool use

Use Grok models through Orq.ai’s router, adding central routing, tracing, evals, budgets, and governance around each request

Coding

Grok 4, Grok 4.20, Grok 3

Use xAI directly for code generation, debugging, and step‑by‑step reasoning over code

Route coding workloads through Orq.ai, compare Grok against other providers, and monitor cost, latency, and quality from one control layer

Vision & search

grok-2-vision, grok-2-image, Grok 4/4.20 with image + X search

Use xAI’s multimodal endpoints and real‑time X grounding directly for vision + search workflows

Use Orq.ai to orchestrate which steps call xAI (for example, X‑grounded reasoning or vision) while other steps use alternative providers, all under unified observability

This gives teams a practical way to use xAI where it performs best while centralising routing, observability, evals, and cost controls across the wider model stack.

Pricing

Plans and API access

xAI offers:

  • Grok API (developer): pay‑per‑token usage across Grok models, independent of X Premium subscriptions.

  • User plans (X app): Free, X Premium, SuperGrok, X Premium+, and SuperGrok Heavy tiers that unlock Grok access inside X, separate from API usage.

For Grok API (2026 public data):aifreeapi+2

  • Cheapest model: grok-4.1-fast and grok-4-fast at $0.20 input / $0.50 output per 1M tokens.

  • Flagship frontier: grok-4 at $3.00 input / $15.00 output per 1M tokens with 256K context.

  • Current reasoning SKU: grok-4.20 at $2.00 input / $6.00 output per 1M tokens with a 2M context window; some catalogs list a Grok 4.3 variant around $1.25 input / $2.50 output per 1M.

  • Legacy: grok-3 at $2.00 / $10.00, grok-3-mini at $0.30 / $0.50, grok-2-vision at $2.00 / $10.00.

xAI also offers limited free developer quotas and free access in some X Premium tiers, but those are separate from Orq‑routed API usage.

Check xAI’s pricing page and your Orq.ai provider configuration for current per‑model rates, cached‑input discounts, and any free‑credit offers when using xAI via Orq.ai.

Compatible frameworks and tools

Orq.ai exposes xAI through:

  • an xAI provider configuration in AI Router, where you paste your xAI API key, and

  • an OpenAI‑style integration, since Grok endpoints use similar chat/completions JSON schemas.

That means:

  • Existing backend workflows using OpenAI SDKs or OpenAI‑compatible clients can switch their base URL to Orq.ai and route some traffic to Grok models without changing request payloads.

  • Agents, code assistants, and tools that integrate with Orq.ai can be configured so that specific steps (for example, X‑grounded reasoning or long‑context Grok 4.20 analysis) are served by xAI, while other steps use different providers, all sharing the same observability and governance layer.

Check the Orq.ai integration docs for the latest supported frameworks and tools for xAI.

FAQs

Do I need a separate xAI account to use xAI through Orq.ai?

Yes. You obtain an xAI API key from the xAI console, then configure it in Orq.ai’s AI Router; Orq.ai then gives you one place to manage routing, observability, and cost controls around that Grok usage.

Can I route only some workflows to xAI and others to different providers?

Yes. You define routes per workflow in Orq.ai and choose which ones should use Grok vs other models, so you can reserve xAI for X‑grounded reasoning, long‑context analysis, or specific frontier tasks while sending other workloads to cheaper or regional providers.

Does using xAI through Orq.ai add latency?

Orq.ai is designed as a lightweight router, so the added overhead is small compared to Grok’s own reasoning and search time. You can use routing policies, caching, and provider selection to keep end‑to‑end performance within your targets while gaining visibility and control.

Alternatives to

xAI

Anthropic

Use Claude Opus 4.7, Sonnet 4.6, Haiku 4.5, and other supported Claude models through one API.

Chat

Code

Reasoning

Vision

Models:

claude-opus-5

claude-sonnet-5

claude-fable-5

Open AI

Use OpenAI's foundation models through a single Orq.ai API. Route models such as GPT-4.1, GPT-4.1-mini, o3-mini, and GPT-4o-class models via Orq's AI Router for chat, reasoning, coding, and multimodal workloads.

Chat

Code

Image Generation

Reasoning

Speech

Vision

Models:

gpt-5.5 (EU)

gpt-5.6-luna (EU)

gpt-5.6-sol (EU)

Google AI

Use Google’s Gemini models through a single Orq.ai API. Route models such as Gemini 3.1 Pro, Gemini 2.5 Flash, and Gemini 2.0 Flash‑Lite via Orq’s AI Router for chat, reasoning, coding, and multimodal workloads.

Chat

Code

Embeddings

Image Generation

Reasoning

Speech

Vision

Models:

gemini-3.5-flash-lite (Gemini API)

gemini-3.6-flash (Gemini API)

Gemini 3 Pro Image

AWS

Use AWS Bedrock’s foundation models through a single Orq.ai API. Route models such as Amazon Nova, Amazon Titan Text, and compatible third‑party models exposed via Bedrock through Orq’s AI Router for chat, reasoning, coding, and multimodal workloads.

Chat

Code

Embeddings

Reasoning

Vision

Models:

eu.anthropic.claude-opus-5

global.anthropic.claude-opus-5

us.anthropic.claude-opus-5

Orq.ai AI Gateway gradient artwork

Get your API key and start routing in minutes

€1 of free credit included. No card. Live in two minutes. The full platform is there when you need it.

Orq.ai AI Gateway gradient artwork

Get your API key and start routing in minutes

€1 of free credit included. No card. Live in two minutes. The full platform is there when you need it.