CapitalBench Score
A score of 30 means the model earned 30% of the best possible return across these rounds. Calculation
A score of 30 means the model earned 30% of the best possible return across these rounds. Calculation
Monthly comparison set
Monthly comparison set automatically opened when the Aug 13 official roster first required a new equal-run benchmark group across 8 models.
Every ranked model in this set is scored only on rounds that all 8 listed models completed. If one model misses a resolved round, that round is excluded from this set for everyone.
Every ranked model in this set completed the same 3 monthly rounds.
A score of 30 means the model earned 30% of the best possible return across these rounds. Calculation
A score of 30 means the model earned 30% of the best possible return across these rounds. Calculation
Average portfolio return across the same finished rounds.
Average realized return and frozen portfolio risk across the same 3 shared monthly rounds.
Return leaderGrok 4.6 led the models at -0.38% average return with a 68.0/100 risk score.
Benchmark testGrok 4.5, Claude Opus 4.8, Grok 4.3, and Gemini 3.1 Pro beat the S&P 500 while taking no more allocation risk.
Claude Fable 5 ranks first in Aug 19 Monthly. Grok 4.6 ranks first in Aug 13 Monthly. The groups share 3 completed rounds. Aug 19 Monthly includes 6 more rounds. Claude Opus 4.8 appears only in Aug 13 Monthly.
Aug 19 Monthly is the main published ranking. Aug 13 Monthly also has enough rounds, so compare them to see whether the results hold across different model groups.
Compare these groupsThis roster stays fixed so the set can keep growing as a clean equal-run comparison.
anthropic-claude-fable-5
3 shared rounds in this set Anthropic Claude Opus 4.8anthropic-claude-opus-4-8
3 shared rounds in this set Anthropic Claude Opus 5anthropic-claude-opus-5
3 shared rounds in this set Google Gemini 3.1 Progoogle-gemini-3-1-pro
3 shared rounds in this set OpenAI GPT-5.6 Solopenai-gpt-5-6-sol
3 shared rounds in this set xAI Grok 4.3xai-grok-4-3
3 shared rounds in this set xAI Grok 4.5xai-grok-4-5
3 shared rounds in this set xAI Grok 4.6xai-grok-4-6
3 shared rounds in this setIncluded rounds count toward the score. Excluded rounds are resolved rounds inside this comparison history where at least one set model was missing.
CapitalBench Score equals total model return across included shared rounds divided by total max-possible return across those same rounds, multiplied by 100. Max possible is the best eligible asset in each included round in hindsight.