Epoch's work is free to use, distribute, and reproduce provided the source and authors are credited under the Creative Commons BY license.
Learn more about this graph
We compare four recent models (GPT-6 Astra, Claude Fable 5.1, GPT-5.6 Sol and Kimi K3) using the Epoch Capabilities Index (ECI), as well as domain-specific ECI variants meant to measure math and software capabilities. Domain-specific ECIs retain the benchmark difficulty parameters from the general ECI fit and re-estimate each model’s capability from only a subset of those benchmarks from a specific domain, making the resulting scores comparable to general ECI scores. See our domain-specific ECI documentation for details on methodology.
Data
Assumptions and limitations
Download this data
Explore this data
Benchmark results featuring the performance of leading AI models on challenging tasks.



