All models

Voxtral Small

by Mistral AI

Chat

Voxtral Small is Mistral AI's speech-understanding model: it takes spoken audio directly as input alongside text, so you can ask questions about a recording rather than transcribing it first. Built on the Mistral Small 24B backbone, it handles a 131K-token context and targets voice assistants and audio analysis.

Publisher
Mistral AI
API model name
voxtral-small:24b
Context window
128K tokens
Sizes
24B
Licence
Apache 2.0
Model card on Hugging Face

Run Voxtral Small

Install it on a Pendra worker, then call it through the OpenAI-compatible API with a pdr_sk_ key.

from pendra import Pendra

client = Pendra(api_key="pdr_sk_...")

response = client.chat.completions.create(
    model="voxtral-small:24b",
    messages=[{"role": "user", "content": "Hello!"}],
)

print(response.choices[0].message.content)

Run Voxtral Small on your own UK infrastructure

Deploy a worker, install Voxtral Small, and start serving it through one sovereign API endpoint.