
Loading comments…
Achievement
Project Info
Product Keywords
Mistral 3 is the next generation of open multimodal and multilingual AI models from Mistral. This release includes three state-of-the-art small, dense models (14B, 8B, and 3B parameters) and Mistral Large 3 – the most capable model to date – a sparse mixture-of-experts (MoE) model trained with 41B active and 675B total parameters. All models are released under the Apache 2.0 license, empowering developers with open access to frontier AI capabilities.
Mistral Large 3 is Mistral's first mixture-of-experts model since the Mixtral series, trained from scratch on 3,000 NVIDIA H200 GPUs. It achieves parity with the best instruction-tuned open-weight models on general prompts while demonstrating image understanding and best-in-class performance on multilingual conversations.
The Ministral 3 series comes in three sizes (3B, 8B, and 14B) with base, instruct, and reasoning variants – each with native multimodal and multilingual capabilities. These models deliver the best cost-to-performance ratio of any open-source model in their category.
Every Mistral 3 model – from the smallest 3B variant to the massive Large 3 – is released under the Apache 2.0 license. This includes both base and instruction fine-tuned versions, providing a strong foundation for further customization across enterprise and developer communities.
Mistral partnered with NVIDIA, vLLM, and Red Hat to deliver efficient inference support. Large 3 runs on a single 8×A100 or 8×H100 node using vLLM, while Ministral models deploy seamlessly on DGX Spark, RTX PCs, laptops, and Jetson devices – from data center to robot.
"Mistral Large 3 joins the ranks of frontier instruction-fine-tuned open-source models while the Ministral series offers the best performance-to-cost ratio in its category."
This dual release is remarkable because it covers the entire spectrum of AI deployment – from massive cloud-based MoE models to tiny edge-optimized dense models – all under one permissive license. Mistral Large 3 debuts at #2 in the OSS non-reasoning models category on the LMArena leaderboard, while the Ministral 3 series sets a new standard for cost-efficient local AI.
You need open-weight models that balance performance, cost, and deployment flexibility. Mistral 3 is ideal if you're building multilingual applications, deploying AI on edge devices, or want a frontier MoE model you can customize freely. The Apache 2.0 license and broad hardware support make it a strong choice for both enterprise and community projects.
Other tools you might consider
Okara lets you use 30+ powerful open-source AI models without dealing with infrastructure setup. The best models like Kimi and DeepSeek are too big to run on your laptop, we handle that for you. Switch between models, search Google, Reddit, X, YouTube in your chats, analyze files, generate images, and work with your team. Everything's encrypted and we never train on your data
TranslateGemma is a new suite of open AI translation models built on Google’s Gemma 3. It enables high-quality communication across 55 languages, combining strong accuracy with exceptional efficiency. Designed to run on mobile, local devices, and cloud environments without compromising performance.
We introduce PersonaPlex, a full-duplex conversational AI model that enables natural conversations with customizable voices and roles. PersonaPlex handles interruptions and backchannels while maintaining any chosen persona, outperforming existing systems on conversational dynamics and task adherence.
Whats 1Code? An app to run your Claude Code agents in parallel that works on Mac and Web. On Mac - run locally, with or without worktrees. On Web - run in remote sandboxes with live previews of your app, mobile included, so you can check on agents from anywhere. Running multiple Claude Codes in parallel dramatically sped up how we build features.
Maker
moonbyte
Loading comments…