# Mistral Large 4 API guide: a first request and its token cost

> How to call Mistral Large 4 in Public Preview, read its response and calculate a sample token charge from official rates.

By BIG CHANGE Editorial

Published: 2026-10-06T17:46:01.996Z
Updated: 2026-10-06T17:46:01.996Z
Canonical: https://bigchange.ai/blog/mistral-large-4-api-guide

![An unbranded desktop monitor displays model: mistral-large-4, choices[0].message.content and usage above a keyboard and a separate calculator.](https://bigchange.ai/api/media/file/mistral-large-4-api-guide-hero-v1.png)
AI-generated conceptual illustration by BIG CHANGE. The field names reflect Mistral's documentation; the image does not depict a screenshot, live API response, measured usage or test.

Mistral opened a Public Preview of Large 4 on October 6, 2026. Developers can call its hosted API with the model ID `mistral-large-4`. The model card describes it as open-weight, while Mistral's company update says a full release is planned for later in October. We did not verify downloadable weights for this preview. This guide follows Mistral's documentation; BIG CHANGE has not sent an API request or measured the model's output.

## The big change

- **What changed:** A new Mistral flagship is available to developers through a priced API preview. Its documented 1 million-token context and vision encoder widen the kinds of material an application can send, though this walkthrough uses text only.
- **Why it matters:** A team can evaluate Large 4 through the hosted chat interface today. The published token rates make a small trial's variable cost calculable in advance. Output quality and latency still need measurement in the team's own application.
- **What to watch:** Mistral says a full release is planned for later this month. Teams interested in self-deployment will need to verify whether downloadable weights become available and under what terms. The current Public Preview can receive silent updates and has no guaranteed route to general availability, so an integration decision also depends on lifecycle tolerance.

## Send a text request

Mistral's [API quickstart](https://docs.mistral.ai/resources/cookbooks/quickstart) says to create a Studio account, activate payments to enable API keys, and use a key for the chat endpoint. It does not state a minimum purchase in those instructions. Store the key as `MISTRAL_API_KEY` in your environment. The [model card](https://docs.mistral.ai/models/mistral-large-4-0) identifies the preview model as `mistral-large-4`; the quickstart's older `mistral-large-latest` example names a different alias. Mistral's [lifecycle policy](https://docs.mistral.ai/inference/model-lifecycle) says Public Preview models do not get `-latest` aliases.

The request below uses the [documented chat completions endpoint](https://docs.mistral.ai/api). Its small task is to extract two facts from the model card text. Replace the prompt with material from your application after checking its accuracy and permissions.

```bash
curl https://api.mistral.ai/v1/chat/completions \
  -H "Content-Type: application/json" \
  -H "Authorization: Bearer $MISTRAL_API_KEY" \
  -d '{
    "model": "mistral-large-4",
    "messages": [
      {
        "role": "user",
        "content": "Extract the model name and context size from this text: Mistral Large 4. Context: 1M tokens. Answer in one sentence."
      }
    ]
  }'
```

The endpoint takes a `model` string and a `messages` array. Here the sole message has the `user` role and text content. On a successful, ordinary chat completion, Mistral documents a response with `choices`; the generated text is in `choices[0].message.content`. The response also contains `model` and `usage`. You should expect an answer identifying Mistral Large 4 and its 1 million-token context, but its exact wording is generated and has not been checked by BIG CHANGE. Inspect the returned `usage` for the actual token counts rather than treating the prompt's visible words as tokens.

The model card lists structured outputs, function calling, document Q&A, chat completions, batching, agents and built-in tools. It describes a vision encoder, but this text request does not establish an image payload format or test visual understanding. Consult the relevant feature documentation before adding files, images or tools to an application.

## Estimate the variable charge

As listed on the [October 6 model card](https://docs.mistral.ai/models/mistral-large-4-0), standard input costs **$0.68 per million tokens**, cached input **$0.07 per million**, and output **$2.09 per million**. For an illustrative request with 1,000 uncached input tokens and 200 output tokens, the token charge is `(1,000 × $0.68 + 200 × $2.09) ÷ 1,000,000 = $0.001098`. Those token counts are assumptions for arithmetic, not a measured run of the request above. Account plan terms and any other charges are outside this example.

The 1 million-token context is the maximum space for input and output combined. Mistral's [known limitations](https://docs.mistral.ai/resources/known-limitations) say an overlong request returns `400 Bad Request`. The platform also applies organization rate limits that vary by plan and model, returning `429 Too Many Requests` when exceeded. Check your account's available models and limits before building around the preview.

Mistral's [lifecycle policy](https://docs.mistral.ai/inference/model-lifecycle) defines Public Preview as near-final and subject to silent updates. It gives no guaranteed path to general availability and a one-month deprecation notice. Keep the model ID explicit in your integration, record the returned model and token usage for each request, and rerun your own application checks when the preview changes. These are practical precautions drawn from the published lifecycle terms, not results from BIG CHANGE testing.

## Sources & further reading

- [Mistral Large 4 model card](https://docs.mistral.ai/models/mistral-large-4-0), dated October 6, 2026. Gives the `mistral-large-4` ID, Public Preview Open v26.10 status, 1 million-token context, listed features and current token prices.
- [Mistral's company update](https://gc.linkedin.com/company/mistralai), October 6, 2026, says the model is available to preview now and a full release is planned for later in the month. Its [launch announcement](https://mistral.ai/news/mistral-large-4/) was unavailable to our retrieval tool. We did not verify a downloadable weight file or an exact weight-release date.
- [API quickstart](https://docs.mistral.ai/resources/cookbooks/quickstart) and [chat endpoint reference](https://docs.mistral.ai/api). Explain payment activation for API keys, the request fields and the response structure. The code above adapts their general chat example to the Large 4 model ID.
- [Model lifecycle policy](https://docs.mistral.ai/inference/model-lifecycle) and [known limitations](https://docs.mistral.ai/resources/known-limitations). Explain preview updates, retirement notice, context errors and rate limits.

## Sources

- [Mistral Large 4 model card](https://docs.mistral.ai/models/mistral-large-4-0) — Primary source for Public Preview Open v26.10, model ID, context, listed features and prices.
- [Mistral API quickstart](https://docs.mistral.ai/resources/cookbooks/quickstart) — Primary source for account and payment activation requirements and the general chat quickstart.
- [Mistral chat completions API reference](https://docs.mistral.ai/api) — Primary source for the documented chat completion request and response fields.
- [Mistral model lifecycle](https://docs.mistral.ai/inference/model-lifecycle) — Primary source for Public Preview update, alias and deprecation terms.
- [Mistral known limitations](https://docs.mistral.ai/resources/known-limitations) — Primary source for context-window and rate-limit errors.
- [Mistral AI company update](https://gc.linkedin.com/company/mistralai) — Mistral's dated company update states preview availability now and a full release later in October; it does not confirm a downloadable-weights date.
