Loading the page.
Documentation
Journal
Frontend
Backend
Admin
CI/CD
Loading the page.
AA Coding Agent Index · artificialanalysis.ai · captured 23.09.2026 · 6 pairs
aa index + task price/time - artificialanalysis.ai (23.09) · t-bench 4.0 - tbench.ai (23.09) · deepswe v1.1 (22.09) · rubench r1 - /rubench (09.07) · "—" - pairing was not run · * - run with fallback model substitution, analysis - /rubench
How to read the table
the pair and the model are on different tabs
AI agents are measured as an agent+model pair (config in the badge), models on their own; the best config for each, one row.
the benchmark is selected on the right
the bar and rating column follow the selected benchmark; «—» - not run on it, the row moves down. Sort by clicking a column.
price and time - real
averages per task: agents - artificialanalysis measurement, models - the actual deepswe run bill. Not derived from tokens.