25 signals
HOlO V1 IS LIVEone ranked AI digest a day, scored in publicREAD HOW IT WORKS →HOlO V2 STARTSyour X account, your signals, every day

Model registry / Mistral AI / Devstral / Devstral Small

Devstral Small

Legacy model retained for compatibility with older integrations

Line

Devstral

Weights

Open

Released

2025-07-10

Context

128K

Input

$0.10 / 1M

Output

$0.30 / 1M

Max output

128K

Coverage

not covered yet

Overview

Devstral Small is a legacy model from Mistral AI, retained for compatibility with older integrations. It belongs to the devstral line, which also includes Devstral Medium. It has a context window of 128000 tokens and a maximum output of 128000 tokens. The weights are open. It supports tool calling but not reasoning mode. It accepts text and returns text. Pricing starts at $0.10 per million input tokens and $0.30 per million output tokens. It is served by 2 hosts.

Specs

Released2025-07-10
LineMistral AI · Devstral
WeightsOpen
Context128K tokens
Max output128K tokens
Inputtext
Outputtext
ReasoningNo
Tool callingYes
Knowledge cutoff2025-05
API iddevstral-small-2507

Pricing

Input$0.10 / 1M tokens
Cached inputnot stated
Output$0.30 / 1M tokens

First version of its line in the registry.

Strengths

  • 128k token context window
  • 128k token max output
  • Open weights for self-hosting
  • Supports tool calling
  • Low input price at $0.10 per million

Best for

  • Reach for it for compatibility with existing integrations
  • Reach for it for large-context text generation
  • Reach for it for tool-calling workflows
  • Reach for it for cost-sensitive text processing

How to access

ProviderModel id
Merge Gatewaydevstral-small-2507
Mistraldevstral-small-2507

2 hosts serve this model at the price above · the maker's documentation

Devstral: every version

VersionReleasedContextInput
Devstral 2Current2025-12-09262K$0.40
Devstral Small 22025-12-09256K$0.00
Devstral Medium2025-07-10128K$0.40
Devstral Small2025-07-10128K$0.10

The registry keeps the six most recent versions of a line. Older ones stay in the file and can be surfaced without a rebuild.

FAQ

What is the context window and max output?
Both the context window and the maximum output are 128000 tokens.
Does it support tool calling and reasoning?
It supports tool calling, but it does not have a reasoning mode.
Is the model open weights and what is the price?
Yes, the weights are open. Pricing starts at $0.10 per million input tokens and $0.30 per million output tokens.
Open weights128K context