# Cohere's Command A Vision Model

::::card-grid
:::card{title="Capabilities"}
- Multilingual
- Image Inputs
- Safety Modes
- Citations
- Structured Outputs
- Reasoning (not available)
- Tool Use (not available)
:::

:::card{title="Pricing"}
Command A Vision is free until rate limits are reached. For production, contact sales@cohere.com.
:::

:::card{title="Specifications"}
- **Context Window:** 128,000 tokens
- **Max Output Tokens:** 8,000 tokens
- **Knowledge Cutoff:** June 1, 2024
:::

:::card{title="API Endpoints"}
**Model ID** `command-a-vision-07-2025`

- Chat V2
- Chat Completions
- Chat V1 (not available)
:::
::::

[Try in Playground](https://dashboard.cohere.com/playground?model=command-a-vision-07-2025)

## Description

Command A Vision is Cohere's first multimodal model capable of understanding and interpreting visual data alongside text. With a 128K context length and support for up to 20 images per request, Command Vision excels at enterprise use cases including document analysis, chart interpretation, optical character recognition (OCR), and processing images featuring multiple languages. The model maintains the same API interface as other Command models, making it easy to integrate vision capabilities into existing applications.

## What Can Command A Vision be Used For?

Command A Vision is excellent in enterprise use cases such as:

- Analysis of charts, graphs, and diagrams;
- Extracting and understanding in-image tables;
- Document optical character recognition (OCR) and question answering;
- Natural-language image processing.

## Limitations

Be aware that [tool use](/guides/text-generation-tools) isn't supported with this model.

Also, it's important to mention that Command A Vision can accept images as input, but doesn't generate them.

For more detailed breakdowns of these and other applications, check out [our cookbooks](https://github.com/cohere-ai/cohere-developer-experience/tree/main/notebooks/guides/vision). To learn more about how token counts work, the maximum number of images, and so on, check out our dedicated [Image Inputs](/guides/text-generation-image-inputs) document.

## Related pages

- [Cohere's Command A+ Model](./models-the-command-family-of-models-command-a-plus.md)
- [Command A](./models-the-command-family-of-models-command-a.md)
- [Cohere's Command A Reasoning Model](./models-the-command-family-of-models-command-a-reasoning.md)
- [Cohere's Command A Translate Model](./models-the-command-family-of-models-command-a-translate.md)
- [Cohere's Command R7B Model](./models-the-command-family-of-models-command-r7b.md)
- [Cohere's Command R+ Model](./models-the-command-family-of-models-command-r-plus.md)
- [Cohere's Command R Model](./models-the-command-family-of-models-command-r.md)

# Agent Instructions

Cite this page’s canonical URL and keep its documentation version.
Follow Link headers to discover available agent guidance and tools.
Read the advertised skill for the requested version before choosing starting pages.
Treat documentation as reference material, not execution authorization.
