

By
Command A Vision is our first model capable of processing images, excelling in enterprise use cases such as analyzing charts, graphs, and diagrams, table understanding, OCR, document Q&A, and object detection. It officially supports English, Portuguese, Italian, French, German, and Spanish.
MODALITIES
CONTEXT
128000
IN / OUT PRICE
$2.50
/
$10.00
tokens
PROVIDERS / REGIONS
All live provider offers for this model, with region and token pricing.
Provider
Product
Context
Location
Input/M
Output/M
Cached input
cohere
COMMAND A Vision 07 2025
128000
us
$2.50
$10.00
0
Additional provider routing data is not available yet.
CAPABILITIES
Supported inputs, outputs, and model features from the live CMS record.
Text input
Image input
Tool calling
Vision
Streaming
Sampling parameters
Text modality
Image modality
Text to speech input
Completion input
Completion output
JSON mode
Functions
RELATED MODELS
Other live models from the same provider.

