A 12B parameter model with a 128k token context length built by Mistral in collaboration with NVIDIA. The model is multilingual, supporting English, French, German, Spanish, Italian, Portuguese, Chinese, Japanese,...
by Mistralai|131K context|$0.02/M input tokens|$0.03/M output tokens
Endpoints
Available providers for this model, with details on pricing, context limits, and real-time health metrics.