The MiniMax release page reports 56.22% on SWE-Pro, 76.5 on SWE Multilingual, and 52.7 on Multi-SWE-Bench, alongside a self-feedback workflow.
MiniMax News / Early Echoes of Self-Evolution · Read evidenceMiniMax M2.7 · Reviews and evidence
Which MiniMax M2.7 conclusions hold up?
Browse public evaluations by topic, source identity, and evidence type. Different versions, tiers, and harnesses are not treated as directly comparable.
This is a third-party source navigator, not a Tabbit test. Use the original source for live metrics; unknown values remain unknown.
Editorial takeaways
Editorial takeaways
BenchLM aggregates 18 public pieces of evidence about MiniMax M2.7 and flags different dates, versions, and harnesses; it is an evidence index, not one unified score.
BenchLM · Read evidenceA Reddit user reports coding and tool-calling experience across about 1,000 prompts, useful for finding parsing failures; it is not a controlled comparison.
Reddit / r/MiniMaxAI · Read evidenceSelected evidence
MiniMax M2.7 Official Release: SWE-Pro, VIBE-Pro, and Agent Workflow Benchmarks
The MiniMax release page reports 56.22% on SWE-Pro, 76.5 on SWE Multilingual, and 52.7 on Multi-SWE-Bench, alongside a self-feedback workflow.
Unverified: the original source could not be rechecked.
- Condition
- Model/version: MiniMax M2.7; release page 2026-03.
- Condition
- Harness/sample: three SWE benchmarks and official workflow; full prompts and repeats unknown.
- Condition
- Date: page reopened 2026-09-20.
18 Pieces of Public Evidence for MiniMax M2.7 on BenchLM
BenchLM aggregates 18 public pieces of evidence about MiniMax M2.7 and flags different dates, versions, and harnesses; it is an evidence index, not one unified score.
Unverified: the original source could not be rechecked.
- Condition
- Model/version: MiniMax M2.7 as labeled by BenchLM; page 2026-08-17.
- Condition
- Harness/sample: 18 items use mixed settings; not one test set.
- Condition
- Date: page reopened 2026-09-20.
Reddit Users' 1,000 Prompts and Coding/Tool-Calling Experience with MiniMax M2.7
A Reddit user reports coding and tool-calling experience across about 1,000 prompts, useful for finding parsing failures; it is not a controlled comparison.
Unverified: the original source could not be rechecked.
- Condition
- Model/version: source identifies the discussed model; client and runtime are not normalized.
- Condition
- Harness/sample: personal report without fixed tasks or repeat rule.
- Condition
- Date: source reopened 2026-09-20.
All sources
All sources
MiniMax M2.7 Official Release: SWE-Pro, VIBE-Pro, and Agent Workflow Benchmarks
The MiniMax release page reports 56.22% on SWE-Pro, 76.5 on SWE Multilingual, and 52.7 on Multi-SWE-Bench, alongside a self-feedback workflow.
Unverified: the original source could not be rechecked.
- Condition
- Model/version: MiniMax M2.7; release page 2026-03.
- Condition
- Harness/sample: three SWE benchmarks and official workflow; full prompts and repeats unknown.
- Condition
- Date: page reopened 2026-09-20.
18 Pieces of Public Evidence for MiniMax M2.7 on BenchLM
BenchLM aggregates 18 public pieces of evidence about MiniMax M2.7 and flags different dates, versions, and harnesses; it is an evidence index, not one unified score.
Unverified: the original source could not be rechecked.
- Condition
- Model/version: MiniMax M2.7 as labeled by BenchLM; page 2026-08-17.
- Condition
- Harness/sample: 18 items use mixed settings; not one test set.
- Condition
- Date: page reopened 2026-09-20.
Reddit Users' 1,000 Prompts and Coding/Tool-Calling Experience with MiniMax M2.7
A Reddit user reports coding and tool-calling experience across about 1,000 prompts, useful for finding parsing failures; it is not a controlled comparison.
Unverified: the original source could not be rechecked.
- Condition
- Model/version: source identifies the discussed model; client and runtime are not normalized.
- Condition
- Harness/sample: personal report without fixed tasks or repeat rule.
- Condition
- Date: source reopened 2026-09-20.
MiniMax M2.7
Compare MiniMax M2.7 in Tabbit
Model access, features, and permissions depend on your current client account.