> This page is for version v2 API (default).
> For other versions, use one of these documentation indexes:
> - v2 API (default): https://docs.cohere.com/v2/llms.txt
> - v1 API: https://docs.cohere.com/v1/llms.txt

> For clean Markdown of any page, append .md to the page URL.
> For a complete documentation index, see https://docs.cohere.com/llms.txt.
> For AI client integration (Claude Code, Cursor, etc.), connect to the MCP server at https://docs.cohere.com/_mcp/server.

# Announcing the Cohere Transcribe model

> This announcement covers the release of Cohere Transcribe, Cohere's first transcription model.

We're pleased to announce the release of [Cohere Transcribe](/docs/transcribe), our first transcription model.
Cohere Transcribe specializes in audio-in, text-out, automatic speech recognition (ASR).

## Technical details

* **Model name**: `cohere-transcribe-03-2026`
* **Input**: Audio waveform
* **Output**: Text
* **Languages covered**: English, German, French, Italian, Spanish, Portuguese, Greek, Dutch, Polish,
  Vietnamese, Chinese, Arabic, Japanese, Korean.
* **License**: [Apache 2.0](https://www.apache.org/licenses/LICENSE-2.0)
* **API endpoint**: [Audio Transcriptions API](/reference/create-audio-transcription)

## Getting started

The model is available immediately through Cohere's [Audio Transcriptions API endpoint](/reference/create-audio-transcription).
You can start transcribing audio using the following example query:

**`PYTHON`**

```python PYTHON
import cohere

co = cohere.ClientV2()

response = co.audio.transcriptions.create(
    model="cohere-transcribe-03-2026",
    language="en",
    file=open("./sample.wav", "rb"),
)

print(response)
```

## Availability

You can access Cohere Transcribe via our [API](http://dashboard.cohere.com) for free, low-setup experimentation
subject to rate limits. See the [Different Types of API Keys and Rate Limits](/docs/rate-limits) page for
usage details and integration guidance.

For production deployment without rate limits, provision a dedicated [Model Vault](/docs/model-vault).
This enables low-latency, private cloud inference without having to manage infrastructure. Pricing is
calculated per hour-instance, with discounted plans for longer-term commitments.
[Contact our team](https://cohere.com/contact-sales) to discuss your requirements.