TestsBenchmarks · Agentic
Claw-Eval
Personal-assistant / OpenClaw-style agent tasks (pass^3 unless noted).
Measured by independent testersIndependent Reported by the makerVendor-reported
| Thinking levelSetting | |||||||
|---|---|---|---|---|---|---|---|
| 1 | MiniMax M3 | MiniMax | 74.5% | Defaultsetting not stated | Maker's own figureVendor-reportedMiniMax ↗ | — | 1 Jun 2026 |