
Natively multimodal MoE model from MiniMax with sparse attention (MSA), up to 1M-token context and open weights; ~428B parameters (23B active).
โ Activeโ Public accessโ Open weightsLLMMultimodalReasoning modelTool-using model
Context window
1M
tokens
Parameters
428B (23B active)
parameters
Access:APIDownloadHostedDeployment:โ Cloud๐ป Local
Overview
Classification
LLMMultimodalReasoning modelTool-using model
Access & deployment
APIDownloadHosted
CloudLocal
Weights: Open weights
Key parameters
๐ Context: 1M
๐งฉ Parameters: 428B (23B active)
โ Tools
๐ฅ Input: text, image
Technical specification
Context window
1M
tokens
Parameters
428B (23B active)
parameters
License
minimax-community
Features:โ Tool use
Modalities
โฌ Input
textimage
โฌ Output
text
Capabilities and applications
Native model capabilities
Long context
Processing very long inputs (tens to hundreds of thousands of tokens) while maintaining coherence.
Category: language
Coding
Generating, completing, explaining and debugging code across multiple programming languages.
Category: coding
Reasoning
The model's ability to perform multi-step logical inference, solve complex problems and decompose tasks into steps.
Category: reasoning
Interleaved Multimodal Input
Category: reasoning
Function Calling
Category: planning
Technical architecture
Core Architecture
Model Form