Mistral AI
Released: 2023-12-11
Mixtral 8x7B
Model Specifications
Context Window 32k tokens
Parameters 47B
Pricing (Input) $0.70 / M tokens
Pricing (Output) $2.10 / M tokens
What is Mixtral 8x7B?
Mixtral 8x7B is a Mixture of Experts (MoE) model released in December 2023. It features 47 billion parameters (activating 13 billion parameters per token), providing the performance of a large model with the speed and cost of a smaller one.
It was one of the first MoE models to be open-sourced under Apache 2.0, setting a standard for open-weight efficiency.
Key Capabilities
- Fast token generation: Sparsity ensures low latency during serving.
- Open-source flexibility: Highly customizable base for fine-tuning.
- Strong reasoning: Out-performs standard 7B and 13B models.
Ideal Use Cases
- Lightweight chatbots: Providing cost-effective user interactions.
- Document classification: Sorting logs and user requests.
- Local model serving: Running models on single-GPU hardware nodes.
Historical figures, architectures, and capabilities are for informational purposes only. Not technical, professional, legal, or financial advice. Sources: Benchmark evaluations derived from public developer statements.