H Company

on Orq.ai

Use H Company’s Holo model family through a single Orq.ai API. Route models such as Holo3‑122B‑A10B and Holo3‑35B‑A3B via Orq’s AI Router for agentic chat, reasoning, coding, and “computer‑use” workflows.

Capabilities:

Chat

Reasoning

Vision

Models Supported:

Holo3 122B A10B

Holo3.1 35B A3B

Provider HQ:

H Company, headquartered in Paris, France, with activity across Europe and beyond.

Access H Company through Orq.ai’s AI Router

H Company builds Holo, a family of mixture‑of‑experts foundation models optimized for agents that can operate desktops, browsers, and other applications, with variants targeting both large‑scale reasoning and efficient inference. These models cover high‑end reasoning, GUI/desktop agents, and cost‑efficient task automation.

H Company models available on Orq.ai

Model

Type

Context / scope

Best for

Pricing tier

Holo3‑122B‑A10B

Chat / reasoning / computer‑use

Around 64K tokens (check H / Orq docs for current limits)

Complex reasoning, OSWorld‑style computer‑use benchmarks, and high‑stakes agent tasks where you want H Company’s flagship MoE model

Premium – priced around 0.40 USD per 1M input tokens and 3.00 USD per 1M output tokens via H Company’s API

Holo3‑35B‑A3B

Chat / reasoning / agents

Around 262K tokens (verify in H / Orq docs)

Everyday agentic workloads, desktop automation, and product features that balance quality with cost and latency

Mid‑tier – priced around 0.25 USD per 1M input tokens and 1.80 USD per 1M output tokens via H Company’s API

Holo‑1 / Holotron‑12B (where integrated)

Chat / vision / computer‑use

Multimodal context (GUI/screens; see H docs)

Lighter‑weight, multimodal “computer‑use” agents and browser automation where smaller models are sufficient

Cost‑efficient – public pricing focuses on larger Holo3 MoE models; lighter models are generally cheaper per 1M tokens when available via H’s platform

Why use H Company through Orq.ai?

Capability

Provider / models

Direct

Through Orq.ai

Chat / agents

Holo3‑122B‑A10B, Holo3‑35B‑A3B, Holo‑1

Call Holo3 and related models directly via H Company’s API for chat, reasoning, and agentic control of desktops and browsers

Use H Company models through Orq.ai’s OpenAI‑compatible endpoint, adding routing, tracing, evals, budgets, and governance controls around each request

Code

Holo3 models tuned or suitable for tooling/automation

Use H Company directly for agent workflows that include code manipulation, test automation (Tester H), and workflow orchestration (Runner H)

Route these workloads through Orq.ai, compare H Company Holo models against other providers, and monitor cost, latency, and quality from one control layer

Embeddings / tools

Embedding providers plus H’s agents (Runner H, Surfer H, Tester H)

H Company focuses on agent and “computer‑use” models; embeddings and classic vector search are typically handled by other stacks

Use Orq.ai to route embeddings to dedicated providers while keeping Holo3 for reasoning, agent steps, and GUI automation

This gives teams a practical way to use H Company where it performs best while centralising routing, observability, evals, and cost controls across the wider model stack.

Pricing

Model rates

H Company model pricing may differ depending on whether you:

  • connect your own H Company API key (BYOK), or

  • use H Company models billed through Orq.ai where available.

Key public figures include:

  • Holo3‑35B‑A3B: about 0.25 USD per 1M input tokens and 1.80 USD per 1M output tokens.

  • Holo3‑122B‑A10B: about 0.40 USD per 1M input tokens and 3.00 USD per 1M output tokens.

H’s agent products (for example Surfer H) are sometimes priced per task (for example around 0.13 USD per task at launch for certain benchmarks), but Orq.ai focuses on per‑token LLM usage when routing Holo models.

Check the Orq.ai pricing page and your workspace’s provider configuration for current per‑model rates, quotas, and plan details.

Compatible frameworks and tools

Orq.ai exposes H Company via:

  • a dedicated H Company provider configuration in the AI Router, and

  • an OpenAI‑compatible API layer for applications that expect that interface.

That means:

  • Popular AI frameworks (for example, OpenAI‑compatible clients and orchestration libraries) can talk to Holo3 via Orq’s router once the H Company provider is configured.

  • Code assistants and IDE tools that support MCP or OpenAI‑compatible APIs (Cursor, VS Code, Claude Desktop, Warp, Zed, and similar tools you’ve documented) can route through Orq.ai to H Company, depending on model and integration configuration.

Check the Orq.ai integration docs for the latest supported frameworks and tools for H Company.

FAQs

Do I need a separate H Company account to use H Company through Orq.ai?

You can either connect your own H Company API key into Orq.ai (from the H Company portal) or, where available, use H Company models billed via Orq.ai; the exact options depend on your Orq plan, region, and how H Company is configured in your workspace. In both cases, Orq.ai gives you one place to manage routing, observability, and cost controls around that H Company usage.docs.orq+3

Can I route only some workflows to H Company and others to different providers?

Yes. You define routes per workflow in Orq.ai and decide which ones should use Holo3 vs other models, so you can reserve H Company for desktop/browser agents and specific reasoning tasks while sending other workloads to different providers.orq+1

Does using H Company through Orq.ai add latency?

Orq.ai is designed as a lightweight router layer, so the added overhead is small compared to the model’s own latency. You can use routing policies, caching, and provider selection to keep end‑to‑end performance within your targets while leveraging H Company’s agent‑optimized models

Alternatives to

H Company

Anthropic

Use Claude Opus 4.7, Sonnet 4.6, Haiku 4.5, and other supported Claude models through one API.

Chat

Code

Reasoning

Vision

Models:

claude-opus-5

claude-sonnet-5

claude-fable-5

Open AI

Use OpenAI's foundation models through a single Orq.ai API. Route models such as GPT-4.1, GPT-4.1-mini, o3-mini, and GPT-4o-class models via Orq's AI Router for chat, reasoning, coding, and multimodal workloads.

Chat

Code

Image Generation

Reasoning

Speech

Vision

Models:

gpt-5.5 (EU)

gpt-5.6-luna (EU)

gpt-5.6-sol (EU)

Google AI

Use Google’s Gemini models through a single Orq.ai API. Route models such as Gemini 3.1 Pro, Gemini 2.5 Flash, and Gemini 2.0 Flash‑Lite via Orq’s AI Router for chat, reasoning, coding, and multimodal workloads.

Chat

Code

Embeddings

Image Generation

Reasoning

Speech

Vision

Models:

gemini-3.5-flash-lite (Gemini API)

gemini-3.6-flash (Gemini API)

Gemini 3 Pro Image

AWS

Use AWS Bedrock’s foundation models through a single Orq.ai API. Route models such as Amazon Nova, Amazon Titan Text, and compatible third‑party models exposed via Bedrock through Orq’s AI Router for chat, reasoning, coding, and multimodal workloads.

Chat

Code

Embeddings

Reasoning

Vision

Models:

eu.anthropic.claude-opus-5

global.anthropic.claude-opus-5

us.anthropic.claude-opus-5

Orq.ai AI Gateway gradient artwork

Get your API key and start routing in minutes

€1 of free credit included. No card. Live in two minutes. The full platform is there when you need it.

Orq.ai AI Gateway gradient artwork

Get your API key and start routing in minutes

€1 of free credit included. No card. Live in two minutes. The full platform is there when you need it.