Jina logo

Clip V2

Clip V2

Multilingual Multimodal embeddings for texts and images

MODALITIES

CONTEXT

8192

INPUT PRICE

$0.05

per 1M

PROVIDERS / REGIONS

All live provider offers for this model, with region and token pricing.

Provider

Product

Context

Location

Input/M

Cached input

No provider offers available.

Additional provider routing data is not available yet.

CAPABILITIES

Supported inputs, outputs, and model features from the live CMS record.

Text input

Image input

Vision

Text modality

Image modality

Text to speech input

Explore the full model garden

Compare providers, capabilities, and pricing across the live Orq.ai catalog.

Build on sovereign EU infrastructure