model releaseOpenAI

OpenAI's GPT-6 Astra Reportedly Automates AI Engineering Tasks at Under $6 an Hour, According to Latent Space Testing

TL;DR

A Latent Space report describes GPT-6 Astra, a new OpenAI model the blog says can autonomously handle AI engineering tasks—training models, labeling data, deploying systems—at an estimated cost of under $6 per hour. The claims, including 97.6% on FrontierMath and 99.9% on ARC-AGI-3, come from independent blog testing rather than an official OpenAI announcement.

2 min read
0

What happened

AI-focused blog Latent Space published a report describing a model it calls "GPT-6 Astra," which it says OpenAI released and which the blog tested extensively—burning what it describes as more than 20 billion tokens across a range of practical tasks. According to the report, Astra can perform work the author compares to hiring a junior AI engineer: selecting and training models, labeling data, monitoring pipelines, reading logs, deploying and debugging systems, and coordinating fleets of subagents (including agents running other models) over what the blog says are threads spanning billions of tokens.

It's important to note this article is based entirely on a single third-party blog's early-access testing. There is no independent OpenAI announcement text included in the source material, and no system card details beyond a reference to one existing. Benchmark numbers, pricing, and capability claims below should be read as reported by Latent Space, not independently verified by us.

The numbers, as reported

Latent Space claims Astra achieves 97.6% on the hardest version of FrontierMath and 99.9% on ARC-AGI-3, which the blog describes as "completely saturating" both benchmarks. The blog also claims Astra beats a competing model it calls "Fable 5.1" on unspecified metrics, and that Artificial Analysis independently confirmed Astra is more token-efficient than models it refers to as "Sol" and "Fable." None of these comparison models appear in confirmed OpenAI or third-party model registries at this time, and the naming suggests these may be code names or placeholders rather than confirmed product names.

On pricing, the blog states it observed throughput of 33 tokens per second at a maximum rate of $50 per million tokens, and extrapolates a real-world cost of under $6 per hour for typical engineering workloads. The report does not break out separate input and output token pricing, nor does it specify context window size. Both figures should be treated as pricing not yet disclosed in verifiable form.

What OpenAI has confirmed

The source material contains no direct OpenAI statement, press release, or pricing page. Quotes attributed to OpenAI figures Greg (likely Greg Brockman) and Jakub (likely Jakub Pachocki) are secondhand, with Greg reportedly saying "AGI is here" and Jakub reportedly calling Astra the "Automated AI Research Intern" he wanted. Neither quote is sourced to an official transcript in the material provided.

What this means

If accurate, a model capable of autonomously running engineering workflows—data labeling, subagent orchestration, deployment, debugging—at single-digit hourly cost would represent a meaningful drop in the cost of AI-assisted software and research work. But every figure here originates from one blog's self-reported, unaudited testing, not from OpenAI's own materials, a published system card, or a benchmark leaderboard we can verify independently. Readers should treat the FrontierMath and ARC-AGI-3 scores, the pricing model, and the capability claims as unconfirmed until OpenAI publishes official documentation or a system card that independent evaluators can check against. We will update this article once verifiable pricing, context window specifications, and benchmark methodology become available.

Related Articles

model release

OpenAI Launches GPT-6 Astra, Claims State-of-the-Art Computer Use and 98% on FrontierMath Tier 4

OpenAI has launched GPT-6 Astra, claiming state-of-the-art results on computer use, coding, and scientific reasoning benchmarks. The model is rolling out to a limited set of organizations first, with general ChatGPT availability expected within days.

model release

OpenAI Releases GPT-6 Astra, First Model to Cross 'Critical' Cybersecurity Threshold

OpenAI has begun rolling out GPT-6 Astra, the first model to reach the company's internal 'Critical' cybersecurity threshold. Access is being phased, with companies in OpenAI's Daybreak cybersecurity program getting priority following added safeguards after a prior model containment breach.

model release

OpenAI Launches GPT-6 Astra, Matches Claude Fable Pricing at $10/$50 per Million Tokens

OpenAI has begun rolling out GPT-6 Astra, priced at $10/million input and $50/million output tokens to match Claude Fable. The model claims a 99.9% score on ARC-AGI 3 using a custom harness and leads on security and long-context benchmarks, though it trails Fable on general intelligence rankings.

changelog

OpenAI Python SDK v3.8.0 Reveals Reference to Unannounced 'gpt-6-astra' Model

The openai-python SDK v3.8.0 release notes, dated September 3, 2026, list a feature addition for 'gpt-6-astra' — a model name not previously confirmed by OpenAI. No official announcement, pricing, or specifications have been released.

Comments

Loading...