AI Battle > Rankings > GPT-6 Astra (xhigh)
GPT-6 Astra (xhigh), published by OpenAI, appears in 15 of the 129 datasets AI Battle tracks. Its strongest published placement is #1 of 20 in Terminal-Bench v4.0: Score.
Last updated: . Every placement below comes from the named public dataset; AI Battle joins them by slug and changes nothing else.
| Ranking | Rank | Value |
|---|---|---|
| Terminal-Bench v4.0: Score | #1 of 20 | 0.60 |
| Humanity's Last Exam: Output Tokens per Task | #2 of 20 | 280 |
| AA-Omniscience Index: Score | #3 of 20 | 43.4 |
| MMMU-Pro: Token Usage | #3 of 19 | 1000367 |
| MMMU-Pro: Score | #3 of 19 | 0.86 |
| Humanity's Last Exam: Time per Task | #4 of 20 | 1.52 |
| Terminal-Bench v4.0: Output Tokens per Task | #4 of 20 | 17246 |
| AA-Omniscience Index: Token Usage | #4 of 20 | 740990 |
| AA-Omniscience Accuracy | #6 of 20 | 0.62 |
| Humanity's Last Exam: Score | #7 of 20 | 0.55 |
| Terminal-Bench v4.0: Time per Task | #9 of 20 | 13.1 |
| AA-Omniscience Hallucination Rate | #13 of 20 | 0.48 |
| MMMU-Pro: Total Cost to Run | #14 of 19 | 10.0 |
| Terminal-Bench v4.0: Cost per Task | #17 of 20 | 0.00 |
| Humanity's Last Exam: Cost per Task | #20 of 20 | 0.00 |
Each row is a separate measurement in a separate dataset — the position reflects the metric of that ranking, and a higher number is not "better" for every dataset (times and prices are costs). AI Battle deliberately does not average placements into a single score: datasets differ in coverage, so a total would hide more than it shows.
Machine-readable profile: /data/models/gpt-6-astra-xhigh.json.
Placements, joins and this page are computed by AI Battle from public datasets — see the methodology for sources, limitations and corrections.