TestsBenchmarks · Agentic
Finance Agent
Vals AI financial-analysis agent benchmark (version noted per row).
Measured by independent testersIndependent Reported by the makerVendor-reported
| Thinking levelSetting | |||||||
|---|---|---|---|---|---|---|---|
| 1 | GPT-5.5 | OpenAI | 60.0% | Extra highreasoning effort=xhigh | Maker's own figureVendor-reportedOpenAI ↗ | — | 23 Apr 2026 |
| 2 | Claude Opus 4.8 | Anthropic | 53.9% | Maxadaptive thinking, effort=max | Maker's own figureVendor-reportedAnthropic ↗ | — | 28 May 2026 |
| 3 | Claude Opus 4.7 | Anthropic | 51.5% | Maxadaptive thinking, effort=max | Maker's own figureVendor-reportedAnthropic ↗ | — | 28 May 2026 |