For the complete documentation index, see llms.txt. This page is also available as Markdown.

Mistral Small 24B

Summary: Mistral Small 24B is a high-performance multilingual and multimodal model from Mistral, built for conversational agents and virtual assistants. It supports a wide range of European languages, is optimized for speed and quality, and can process text, images, and mixed input. Images can be provided as URLs or base64-encoded strings, allowing the model to analyze and respond to visual content in addition to text. This model is particularly effective for applications requiring fast and reliable chat completions, making it suitable for real-time interactions in various domains. The model supports advanced capabilities like streaming and tool calling, enhancing its usability in interactive applications.

Intelligence

Speed

Input

Output

Intelligence active Intelligence active

Speed active Speed active

Text active Model icon Audio inactive

Text active Image inactive Audio inactive

Moderate

High

Text, Image

Text

Central parameters

Description: Medium-sized model with 24B parameters featuring a 128k token context window. Images are processed as input tokens, so including images in your requests will increase the total input token count. Token billing is based on the total number of tokens processed, including both input and output tokens.

Model identifier: mistralai/Mistral-Small-24B-Instruct

IONOS CLOUD AI Model Hub Lifecycle and Alternatives

IONOS start date

End of Life

Alternative

Successor

August 1, 2025

N/A

Origin

Provider

Country

License

Flavor

Release

France

Instruct

June 25, 2025

Technology

Context window

Parameters

Quantization

Multilingual

Further details

128k

24B

fp8

Yes

Modalities

Text

Image

Audio

Input and output

Input

Not supported

Endpoints

Chat Completions

Embeddings

Image generation

v1/chat/completions

Not supported

Not supported

Vision model example

Base64 image example

When providing a base64 encoded image, use the image_url type and map the url field to your data URI. Example: data:image/jpeg;base64,...

Response:

Features

Streaming

Reasoning

Tool calling

Supported

Not supported

Supported

Rate limits

Rate limits ensure fair usage and reliable access to the AI Model Hub. In addition to the contract-wide rate limits, no model-specific limits apply.

Last updated

Was this helpful?