TestsBenchmarks · Agentic
WildClawBench
Agent tasks in the OpenClaw-style personal agent harness.
Measured by independent testersIndependent Reported by the makerVendor-reported
| Thinking levelSetting | |||||||
|---|---|---|---|---|---|---|---|
| 1 | Muse Glimmer 30B | Meta (Meta Superintelligence Labs, Muse) | 47.6% | HighHigh reasoning | Maker's own figureVendor-reportedMeta ↗ | — | 10 Aug 2026 |