We read every piece of feedback, and take your input very seriously.
To see all available qualifiers, see our documentation.
1 parent 6f162bf commit b31febcCopy full SHA for b31febc
1 file changed
README.md
@@ -89,6 +89,8 @@ CoreAI ships a game-creation benchmark: it scores how well an LLM builds and cha
89
90
**Top-tier models — v2 frontier sweep (2026-07-11):**
91
92
+<img src="Docs/Images/benchmark_v2_frontier.svg" alt="CoreAI Game-Creation Benchmark v2 — top-tier frontier-model comparison ranked by suite score" width="900">
93
+
94
| # | Model | Suite | Pass-rate |
95
|---:|---|---:|---:|
96
| 1 | `gpt-5.6-sol` | **96.6** | 85.7% |
0 commit comments