Loading...
Loading...
A 12B parameter model with a 128k token context length built by Mistral in collaboration with NVIDIA. The model is multilingual, supporting English, French, German, Spanish, Italian, Portuguese, Chinese, Japanese,...
No benchmark data available.
Check your models and usage for free. See the relevant official changes, estimated cost, and a practical next step.