> This page is for version v2 API (default).
> For other versions, use one of these documentation indexes:
> - v2 API (default): https://docs.cohere.com/v2/llms.txt
> - v1 API: https://docs.cohere.com/v1/llms.txt

> For clean Markdown of any page, append .md to the page URL.
> For a complete documentation index, see https://docs.cohere.com/llms.txt.
> For AI client integration (Claude Code, Cursor, etc.), connect to the MCP server at https://docs.cohere.com/_mcp/server.

# Quickstart

> Create your first vault from the Model Vault app and make an inference request in a few minutes.

This quickstart walks you through creating your first vault from the Model Vault app and making an
inference request against it. The same flow applies to every vault; the only choice is whether to
create a **Standard** or an **Encrypted** vault.

## Choose a vault type

Go to [dashboard.cohere.com](https://dashboard.cohere.com/), open **Vaults**, and create a new vault. First
choose the vault type, **Standard** or **Encrypted** (this can't be changed after creation). Encrypted
vaults run inside hardware-backed trusted execution environments and support remote attestation. See
[Encrypted Vaults](../../docs/model-vault/encrypted).

![](/_fern-img/3131a96fb7c86d66179a62dc07970eaaf02ddba61be85153b8ad624cc63332a0.webp)

## Select a model and performance tier

In the configuration panel, name your vault, select a model, and choose a performance tier, then set
the replica range. See [Creating a Vault](../../docs/model-vault/creating-a-vault) for the full walkthrough.

![](/_fern-img/d759b712c7ecb82c96e260d046a0e952b087c61a89209428289771d85d9394ae.webp)

## Get your endpoint and model name

Once the vault status is **Ready**, open its summary page and copy the **endpoint URL** and **model
name**. See [Calling a Vault over the API](../../docs/model-vault/standard/api-access).

![](/_fern-img/7db7fbbc218722f149e912a1e96e072068748bd95e11ef81f76773c7cf7b5e41.webp)

## (Encrypted only) Verify the attestation

If you created an encrypted vault, the client verifies the deployment's remote attestation
automatically before sending any data, and refuses the connection if it does not pass. You can also inspect the
attestation results, the CPU, GPU, software, and policy checks, in the Model Vault app or through the API,
so you have proof of exactly what is running. See
[Verifying Your Deployment](../../docs/model-vault/encrypted/verifying-deployment) and
[Remote Attestation](../../docs/model-vault/encrypted/attestation).

![](/_fern-img/981791dde6cc4040fe380481c8827bf03dbb02d1326dd2b0d6a2f3203db6b159.webp)

## Make a request

Point the Cohere SDK at your vault endpoint and call it like any other Cohere model.

**`PYTHON`**

```python PYTHON
import cohere

co = cohere.ClientV2(
    api_key="<COHERE_API_KEY>",
    base_url="<YOUR_VAULT_ENDPOINT_URL>",
)

response = co.chat(
    model="<YOUR_VAULT_MODEL_NAME>",
    messages=[{"role": "user", "content": "Hello from Model Vault!"}],
)

print(response.message.content[0].text)
```

## Next steps

* [Managing Vaults](../../docs/model-vault/managing-vaults): view vault details, edit, pause/resume, and delete models.
* [Calling a Vault over the API](../../docs/model-vault/standard/api-access): full request patterns, base URL configuration, and supported endpoints.
* [Monitoring](../../docs/model-vault/monitoring): track latency, throughput, and utilization.
* Go deeper on your vault type: [Standard Vault](../../docs/model-vault/standard) or [Encrypted Vault](../../docs/model-vault/encrypted).