all the models — AI benchmark observatory
← Models

MiniMax M2.7

MiniMax · open weight · minimax-m2.7

MiniMax M2.7 features model self-improvement driving productivity innovation. It builds complex agent harnesses independently to accomplish highly complex productivity tasks. M2.7 demonstrates excellent performance in real-world software engineering including end-to-end project delivery, log analysis, code security, and ML tasks. On SWE-Pro it scores 56.22%, nearly matching Opus. It excels in professional office domains achieving the highest ELO among open-source models on GDPval-AA (1495), with significant improvement in complex editing for Office Suite. M2.7 maintains 97% skill adherence on 40 complex skills cases.

Not enough scores for a fingerprint yet.

Benchmark scores

BenchmarkScore
SWE-bench Multilingual0.77

Pricing

  • MiniMax$0.30 / $1.20
  • Novita$0.30 / $1.20
  • Fireworks$0.30 / $1.20
  • DeepInfra$0.38 / $1.70

Input / output per 1M tokens

AA metrics

No Artificial Analysis link yet.

Arena Elo

  • No Arena snapshot linked.