Build Interactive 3D Worlds With GPT-6 & Blender

TL;DR
Three frontier launches in 48 hours repriced the agentic workhorse tier: Grok 4.7 at $2/$6 (Sep 21), Claude Opus 5.5 at $4/$20 with $0.20 cache reads (Sep 22), and GPT-6 Sol at $2/$10 with Luna at $0.10/$0.50 (Sep 22). Same-day-verified rates, honest benchmark attribution, and a decision guide.
Direct answer
Three frontier launches in 48 hours repriced the agentic workhorse tier: Grok 4.7 at $2/$6 (Sep 21), Claude Opus 5.5 at $4/$20 with $0.20 cache reads (Sep 22), and GPT-6 Sol at $2/$10 with Luna at $0.10/$0.50 (Sep 22). Same-day-verified rates, honest benchmark attribution, and a decision guide.
Best for
Developers comparing real tool tradeoffs before choosing a stack.
Covers
Verdict, tradeoffs, pricing signals, workflow fit, and related alternatives.
Last updated: September 26, 2026
Three frontier-class models launched inside 48 hours, and all three repriced the tier they entered. SpaceXAI's Grok 4.7 opened on September 21 at $2/$6 per million tokens. Anthropic's Claude Opus 5.5 arrived on September 22 at $4/$20 with cache reads cut to $0.20. OpenAI answered the same day with GPT-6 Sol at $2/$10 and GPT-6 Luna at $0.10/$0.50. Within hours the comparison charts were pointing at each other's benchmark tables and disagreeing by double digits. This post compares the three launches head to head with same-day-verified prices from the vendors' own pages, separates vendor-reported benchmarks from third-party work, and works through who should move to what.
All prices below were verified September 26, 2026 against the vendors' live pages:
| Vendor | Pricing page | Announcement | Notes |
|---|---|---|---|
| OpenAI | developers.openai.com/api/docs/pricing | Introducing GPT-6 Sol and Luna | Announcement returned HTTP 403 on re-fetch September 26; rates taken from the live pricing page |
| Anthropic | platform.claude.com/docs/en/about-claude/pricing | Claude Opus 5.5 | Cache multipliers per the pricing docs footnotes |
| SpaceXAI | Grok 4.7 announcement | Grok 4.7 | Rates and benchmarks from the announcement; no full public rate card |
| DeepSeek | api-docs.deepseek.com/quick_start/pricing | V4.1-Flash release | Context for the open-weights floor |
Grok 4.7 (September 21). SpaceXAI positions it as its most capable coding and knowledge-work model, trained with a longer RL run weighted toward tasks that take hours, with a new safeguard stack (3.3% of risky dual-use prompts pass HackerBench v0.3, per the announcement). It serves at the same price and speed as Grok 4.6: $2 input / $6 output per MTok, with a fast variant at double the output speed for double the price. Available in Cursor, Grok Build, and the Grok API.
Claude Opus 5.5 (September 22). Anthropic's first model in the new Claude 5.5 family (Sonnet 5.5 and Haiku 5.5 are promised "in the coming weeks") performs at Fable 5.1 level on most work at 40 percent lower cost than Opus 5: $4 input / $20 output, cache reads at $0.20 (a 0.05x multiplier, down from the standard 0.1x), and fast mode at $8/$40. It set new highs on Terminal-Bench 4.0 (66.4%) and GDPval-AA v2.1 (1846 Elo). Our full release guide has the benchmarks, the effort-level caveat (at max effort it can over-think into the 128K output ceiling), and the OpenCode setup.
GPT-6 (September 22). OpenAI's new family ships three tiers: Astra ($10/$50, max capability), Sol ($2/$10, the coding-and-agentic flagship), and Luna ($0.10/$0.50, high-volume workhorse), all with a 1,050,000-token context window. Sol at $2/$10 is the first frontier-class model priced below the mid tier, and Luna is the cheapest model OpenAI has ever shipped - cheaper than DeepSeek on every axis. GPT-5.6 stays live: Sol's $4/$20 promo runs at least through November 21, 2026.
Standard rates per million tokens, verified September 26, 2026:
| GPT-6 Sol | Claude Opus 5.5 | Grok 4.7 | Claude Fable 5.1 (reference) | |
|---|---|---|---|---|
| Launch date | Sep 22 | Sep 22 | Sep 21 | Sep 1 |
| Input | $2.00 | $4.00 | $2.00 | $10.00 |
| Cached input read | $0.20 | $0.20 | not published | $0.25 |
| Output | $10.00 | $20.00 | $6.00 | $50.00 |
| Long context | $4.00 / $15.00 (over 272K) | flat to 1M | not published | flat to 1M |
| Context window | 1.05M | 1M | not published | 1M |
| Fast mode | no | $8 / $40 | 2x speed at 2x price | no |
| Position | coding-and-agentic flagship | workhorse of the 5.5 family | coding and knowledge work | Anthropic's max-capability |
Benchmarks, with honest attribution (none of these are independent third-party runs; each vendor published its own table):
| Benchmark | GPT-6 Sol | Claude Opus 5.5 | Grok 4.7 | Fable 5.1 |
|---|---|---|---|---|
| CursorBench 4.0 | 41.7% (published by SpaceXAI's chart; OpenAI has not published one) | roughly 11 points above 41.7% (Anthropic) | 46.3% (SpaceXAI) | 51.8% (SpaceXAI's chart) |
| Terminal-Bench 4.0 | not published by OpenAI | 66.4% (Anthropic) | 37.6% (SpaceXAI) | 55.8% (Anthropic's Opus 5.5 announcement) |
| DeepSWE (v1.1 where noted) | not published by OpenAI | not published | 71.0% high effort (SpaceXAI) | 70.0% (SpaceXAI) |
| GDPval (Elo) | 1542 (SpaceXAI's chart, Astra) | 1846 (AA GDPval-AA v2.1, Anthropic) | 1695 (SpaceXAI) | 1735 (SpaceXAI) |
The cross-vendor numbers clash on purpose: SpaceXAI's chart scores GPT-6 Sol at 41.7% CursorBench and Fable 5.1 at 51.8%, Anthropic's chart scores Opus 5.5 roughly 11 points above Sol's best, and Terminal-Bench 4.0 splits the field three ways with Opus 5.5 clearly on top (66.4%) and Grok 4.7 far below (37.6%). The only third-party workflow data we have seen this month puts GPT-6 Sol at 79.1% at $0.2152 per run on TypeSafe's workflow evals (fixed code-defined workflows, benchmarked against the average of GPT-6 Astra and Fable 5.1). Treat all of it as directional: run your own golden set before you switch, and note that DeepSeek's V4.1-Flash at $0.15/$0.60 off-peak is still the open-weights reference floor the frontier pricing tracker tracks.
From the archive
Sep 25, 2026 • 10 min read
Sep 23, 2026 • 8 min read
Sep 22, 2026 • 7 min read
Sep 21, 2026 • 8 min read
Per-token rates understate what agent loops actually spend, so here is the same task across the field: 200K input / 20K output, standard rates, no cache.
| Model | Uncached task | With 90% cached input |
|---|---|---|
| GPT-6 Sol | $0.40 + $0.20 = $0.60 | $0.036 + $0.04 + $0.20 = $0.276 |
| Claude Opus 5.5 | $0.80 + $0.40 = $1.20 | $0.036 + $0.08 + $0.40 = $0.516 |
| Grok 4.7 | $0.40 + $0.12 = $0.52 | no published cache rate |
| Fable 5.1 | $2.00 + $1.00 = $3.00 | $0.045 + $0.20 + $1.00 = $1.245 |
| GPT-5.6-Sol (promo) | $0.80 + $0.40 = $1.20 | $0.036 + $0.08 + $0.40 = $0.516 |
Two takeaways. First, the gap between "cheap" and "frontier" collapsed: GPT-6 Sol's $0.60 uncached task is within 15% of Grok 4.7's $0.52 and half of Opus 5.5's $1.20, and it is the most capable coding model of the three by every published agent benchmark except the ones Anthropic disputes. Second, caching is where Opus 5.5 fights back: its 0.05x cache multiplier means a typical agent loop (over 90% of input tokens cached) pays $0.036 in cache reads versus Sol's identical $0.036, but Opus 5.5's uncached input at 2x Sol's rate. If your loops are cache-dominant and quality-sensitive, Opus 5.5 at $0.516 per task is the cheapest frontier-grade option with a published cache rate.
Coding agents on the OpenAI stack. GPT-6 Sol, unambiguously. It is the designated coding-and-agentic flagship, inherits Codex, the Responses API, and structured outputs, and its $10 output rate means output-heavy loops no longer need routing out of the flagship tier. If your workload is high-volume and simple, GPT-6 Luna at $0.10/$0.50 does the same job for a third of the input cost - see the budget tier comparison.
Long-horizon, quality-critical agent runs. Opus 5.5 is the strongest terminal-side story on the table (Terminal-Bench 4.0 66.4%) with the published cache economics to back it, and its fast mode at $8/$40 is a price-for-speed option alongside Grok's (2x output speed at 2x price). The caveat from early testers: stay on default effort levels, because max effort can burn the output ceiling.
Price-performance maximalists with router-friendly workloads. Grok 4.7 at $2/$6 is the cheapest flagship output rate on the market, with DeepSWE 71.0% (high effort) that matches Fable 5.1 and slots far above both GPT-6 Sol and Opus 5.5 on that axis. Its Terminal-Bench 4.0 score (37.6%) is the honest counterweight. Routing cheap-by-default with escalation handles exactly this shape of tradeoff - the routing strategies guide has the framework.
Anthropic shops. If you standardized on Claude for tooling and eval surface, note the pricing reality: Opus 5.5 at $4/$20 is 2.5x cheaper on input and output than Fable 5.1, and Anthropic's own claim is that it performs at Fable 5.1 level on most work. Buy the step down to Opus 5.5 unless your eval set shows the ceiling is what you pay for.
Cost floors. For any task where quality is secondary, the month's real news is GPT-6 Luna at $0.10/$0.50: it undercut DeepSeek on every axis, so the "cheapest frontier callable" label moved from open weights to a closed provider for the first time. Budget models built on the old DeepSeek math need a re-basing pass - the DeepSeek migration guide covers the September rename and rates.
GPT-6 Sol at $2/$10 per MTok versus Opus 5.5 at $4/$20 - Sol is 2x cheaper on both axes. On cached agent loops the gap narrows to $0.276 versus $0.516 per representative task, because Opus 5.5's cache reads ($0.20) match Sol's ($0.20) at a 0.05x multiplier versus Sol's 0.1x.
Yes, verified September 26, 2026 from the official announcement: $2 per million input tokens and $6 per million output tokens, with a fast variant at double the output speed for double the price.
Astra ($10/$50) is the max-capability model, used for long-horizon research runs and flagpole work; Sol ($2/$10) is the coding-and-agentic flagship and the sensible default for developers; Luna ($0.10/$0.50) is the high-volume workhorse. All three share a 1,050,000-token context window, and long-context rates rise to $20/$75, $4/$15, and $0.20/$0.75 respectively.
It is the successor and replaces it as the coding-and-agentic flagship at half the price ($2/$10 versus $4/$20). GPT-5.6-Sol's promotional pricing runs at least through November 21, 2026, so the legacy tier remains available while teams migrate.
Yes, at $8/$40 per MTok, available on the Claude API only (not Bedrock or other platforms), applied across the full context window. It is the same 2x-price structure as OpenAI's fast mode on the GPT-5.6 family.
The 5.5 family is announced, not shipped: Anthropic's Opus 5.5 release notes say Sonnet 5.5 and Haiku 5.5 will follow "in the coming weeks" (verified September 26, 2026 on the pricing page, which lists neither yet). If you run Sonnet 5 at $2/$10 today, there is no announced successor price to wait on - Sonnet 5's rates are permanent.
Read next
Anthropic shipped Claude Opus 5.5 on September 22, 2026: Fable 5.1-level performance on most work at $4/$20 per million tokens (40% cheaper than Opus 5), cache reads down 60% to $0.20, output 30% faster. Benchmarks, decision guide, and the OpenCode setup.
8 min readThe 5-7 AI developer stories that actually mattered this week - ranked, linked, and cut for builders.
11 min readSame-day-verified llm api pricing september 2026: Claude Fable 5.1 and Opus 5.5, GPT-6 Astra/Sol/Luna, Grok 4.7, Claude Sonnet 5, and DeepSeek V4.1-Flash compared per million tokens, plus the caveats that change the math.
11 min readTechnical content at the intersection of AI and development. Building with AI agents, Claude Code, and modern dev tools - then showing you exactly how it works.
Anthropic's recommended default for complex work, released May 28, 2026. 1M context, 128K output, $5/$25 per million tok...
View ToolAnthropic's agentic coding CLI. Runs in your terminal, edits files autonomously, spawns sub-agents, and maintains memory...
View ToolAnthropic's AI. Opus 4.6 for hard problems, Sonnet 4.6 for speed, Haiku 4.5 for cost. 200K context window. Best coding m...
View ToolAnthropic's flagship reasoning model. Best-in-class for coding, long-context analysis, and agentic workflows. 1M token c...
View ToolEvery coding agent in one window. Stop alt-tabbing between Claude, Codex, and Cursor.
View AppTurn a one-liner into a working Claude Code skill. From idea to installed in a minute.
View AppUnlock pro skills and share private collections with your team.
View AppDeep comparison of the top AI agent frameworks - LangGraph, CrewAI, Mastra, CopilotKit, AutoGen, and Claude Code.
AI AgentsUse opus, sonnet, haiku, and best to switch models easily.
Claude CodeHybrid mode: Opus for planning, Sonnet for execution.
Claude Code
Anthropic released Claude Opus 5, described as a thoughtful, proactive model approaching frontier intelligence at about half the price of Fable, and the video reviews the announcement, benchmarks, and...

Anthropic Releases Claude Opus 4.7: Benchmarks, Vision Upgrades, Memory, Pricing & New Claude Code Features Anthropic has released Opus 4.7, and the video covers the announcement, benchmark results, ...

Exploring Claude Opus 4.6: Features, Benchmarks, Anthropic's Latest Frontier Model In this video, I delve into the details of Claude Opus 4.6, highlighting key features and performance benchmarks. Th

Anthropic shipped Claude Opus 5.5 on September 22, 2026: Fable 5.1-level performance on most work at $4/$20 per million...

The 5-7 AI developer stories that actually mattered this week - ranked, linked, and cut for builders.

Same-day-verified llm api pricing september 2026: Claude Fable 5.1 and Opus 5.5, GPT-6 Astra/Sol/Luna, Grok 4.7, Claude...

The cheap coding tier repriced again: GPT-6 Luna opened at $0.10/$0.50 and took the floor from DeepSeek, whose V4.1-Flas...

OpenAI slashes GPT-5.6 Luna by 80% to $0.20/M input tokens, cuts Terra by 20%, adds Sol Fast mode at 2.5x speed, and rev...

A verified directory of the frontier AI models in July 2026 - Claude Fable 5, Opus 5, GPT-5.6 Sol/Terra/Luna, Sonnet 5,...

New tutorials, open-source projects, and deep dives on coding agents - delivered weekly.