M

MiniMax: minimax-m2-her

minimax-m2-her
204.8K Kontext2.0K AusgabeReasoningToolsCache
Veröffentlicht Jan 23, 2026Wissensstand Jun 2025Aktualisiert Sep 24, 2026

MiniMax-M2 is a high-efficiency Mixture-of-Experts (MoE) model architected specifically for coding and agentic workflows. While boasting 230B total parameters, it only activates 10B per token, ensuring low latency and high throughput. Its standout feature is "Interleaved Thinking," where the model uses internal reasoning tags (<think>) to plan and self-correct during complex tasks. This makes it exceptionally robust for multi-step tool use, terminal-based coding, and long-horizon planning. It ranks as a top-tier open-weight model, rivaling proprietary giants in agentic benchmarks like VIBE and SWE-bench.

Modus chatTokenizer Other

Preise

Input-Preis
$0.60/ 1 Mio. Tokens
Output-Preis
$2.40/ 1 Mio. Tokens
Kontextfenster 204.8K TokensKompatible Endpunkte openaiAnbieter MiniMax

Verfügbarkeit

Leistung

Lade Leistungsdaten...

Nutzung & Rang

Nutzung wird geladen...

Unterstützte Parameter

Alle Anbieter = jeder Upstream, der dieses Modell bereitstellt, unterstützt den Parameter. Einige Anbieter = hängt vom Upstream ab, der die Anfrage bearbeitet. Standard = der Wert, der gesendet wird, wenn nichts gesetzt ist.

ParameterAnbieterStandard
frequency_penaltyEinige AnbieterStandardmäßig nicht gesendet
max_tokensAlle Anbieter-
temperatureAlle Anbieter1
top_pAlle Anbieter0.95

Häufige Fragen

Ähnliche Modelle