M

MiniMax: minimax-m2-her

minimax-m2-her
204.8K contesto2.0K outputRagionamentoStrumentiCache
Rilasciato Jan 23, 2026Limite delle conoscenze Jun 2025Aggiornato Sep 24, 2026

MiniMax-M2 is a high-efficiency Mixture-of-Experts (MoE) model architected specifically for coding and agentic workflows. While boasting 230B total parameters, it only activates 10B per token, ensuring low latency and high throughput. Its standout feature is "Interleaved Thinking," where the model uses internal reasoning tags (<think>) to plan and self-correct during complex tasks. This makes it exceptionally robust for multi-step tool use, terminal-based coding, and long-horizon planning. It ranks as a top-tier open-weight model, rivaling proprietary giants in agentic benchmarks like VIBE and SWE-bench.

Modalità chatTokenizer Other

Prezzi

Prezzo di input
$0.60/ 1 M token
Prezzo di output
$2.40/ 1 M token
Finestra di contesto 204.8K tokenEndpoint compatibili openaiProvider MiniMax

Disponibilità

Performance

Caricamento dati di performance...

Utilizzo e classifica

Caricamento utilizzo...

Parametri supportati

Tutti i provider = supportato da ogni upstream che serve questo modello. Alcuni provider = dipende dall'upstream che gestisce la richiesta. Predefinito = il valore inviato quando non lo imposti.

ParametroProviderPredefinito
frequency_penaltyAlcuni providerNon inviato per impostazione predefinita
max_tokensTutti i provider-
temperatureTutti i provider1
top_pTutti i provider0.95

Domande frequenti

Modelli simili