Wire flash
TechGPT-6 Astra tops Terminal Bench 4.0 at half the cost of second-place model
Editorial responsibility
- No named human review is recorded for this page.
- Source reporting is collected, normalized, translated or condensed automatically when needed.
- Automatically published source-backed update
A post on X by user thsottiaux announces that GPT-6 Astra has achieved the number one ranking on Terminal Bench 4.0 using the Codex harness. The post highlights that GPT-6 Astra accomplishes this at 50% of the cost of the second-place model. Terminal Bench 4.0 is a benchmark for evaluating AI model performance, and the Codex harness refers to a specific testing framework. The claim positions GPT-6 Astra as both the top performer and a cost-efficient option in the competitive AI landscape. The post includes a link to further details, though the full context of the benchmark and the identity of the second-place model are not provided in the post itself. This development underscores ongoing advancements in AI model efficiency and performance.
Source report
GPT-6 Astra has secured the #1 position on the Terminal Bench 4.0 benchmark, utilizing the Codex harness. Notably, it achieves this top performance at 50% of the cost of the second-ranked model.
Source
Source
thsottiauxNeutral / independent
Part of this Story
GPT-6 Astra tops multiple benchmarks at lower cost but faces price hike and mixed results