25 signals
HOlO V1 IS LIVEone ranked AI digest a day, scored in publicREAD HOW IT WORKS →HOlO V2 STARTSyour X account, your signals, every day

Model registry / Mistral AI / Devstral / Devstral Medium

Devstral Medium

Legacy model retained for compatibility with older integrations

Line

Devstral

Weights

Open

Released

2025-07-10

Context

128K

Input

$0.40 / 1M

Output

$2.00 / 1M

Max output

128K

Coverage

not covered yet

Overview

Devstral Medium is a legacy model in the devstral line, retained for compatibility with older integrations. It is the first version of its line in the registry, released on 2025-07-10. The model has a context window of 128000 tokens and a maximum output of 128000 tokens. It accepts text and returns text. It supports tool calling but has no reasoning mode. The weights are open. It is served by 2 hosts. Pricing starts at $0.40 per million input tokens and $2.00 per million output tokens. Its knowledge cutoff is 2025-05.

Specs

Released2025-07-10
LineMistral AI · Devstral
WeightsOpen
Context128K tokens
Max output128K tokens
Inputtext
Outputtext
ReasoningNo
Tool callingYes
Knowledge cutoff2025-05
API iddevstral-medium-2507

Pricing

Input$0.40 / 1M tokens
Cached inputnot stated
Output$2.00 / 1M tokens

First version of its line in the registry.

Strengths

  • 128k token context window
  • 128k token max output
  • Supports tool calling
  • Open weights for self-hosting
  • Text-to-text processing

Best for

  • Reach for it for maintaining legacy integrations
  • Reach for it for text-based tool calling workflows
  • Reach for it for tasks needing a 128k context window

How to access

ProviderModel id
Merge Gatewaydevstral-medium-2507
Mistraldevstral-medium-2507

2 hosts serve this model at the price above · the maker's documentation

Devstral: every version

VersionReleasedContextInput
Devstral 2Current2025-12-09262K$0.40
Devstral Small 22025-12-09256K$0.00
Devstral Medium2025-07-10128K$0.40
Devstral Small2025-07-10128K$0.10

The registry keeps the six most recent versions of a line. Older ones stay in the file and can be surfaced without a rebuild.

FAQ

What is the context window size?
The context window is 128000 tokens, and the maximum output is also 128000 tokens.
Does it support reasoning?
No, it does not have a reasoning mode. It supports tool calling and processes text only.
Are the weights open?
Yes, the weights are open. The model is served by 2 hosts and pricing starts at $0.40 per million input tokens and $2.00 per million output tokens.
Open weights128K context