• Latest
  • Trending
  • All
Answer card for GPT-6.1 Sol showing 2 dollars in and 10 dollars out per million tokens, unchanged from GPT-6 Sol, with cached input at 10 cents.

GPT-6.1 Sol costs what GPT-6 Sol did, and only the cache got cheaper

30 September 2026
Answer card: Reflection AI announced Beam, a 501B parameter open-weight mixture of experts model, with weights promised later in October 2026.

Reflection Beam is a 501B open model with no weights yet

6 October 2026
Answer card: Aleph Alpha released Kolibri-1 on 3 October 2026 as an Apache 2.0 mixture of experts model with 78.1 billion total and 3.46 billion active parameters, served at 1 million tokens of context but trained on sequences of up to 256 thousand tokens.

Kolibri-1 serves 1M tokens but was trained to 256K

4 October 2026
Answer card on Gemini 4 Argon pricing and gated access, announced by Google on 30 September 2026.

Gemini 4 Argon costs $2 and $10 now, $4 and $20 later

2 October 2026
Answer card for Claude Sonnet 5.5 showing 2 dollars in and 10 dollars out per million tokens, unchanged from Sonnet 5, and between_tools replacing disabled thinking.

Claude Sonnet 5.5 keeps $2 and $10 and retires thinking disabled

29 September 2026
Answer card: Xiaomi retrained MiMo-V2.6 to stop repeating tool calls, API swapped on 25 September 2026, MIT weights on 27 September, same model names.

MiMo-V2.6-Pro was quietly retrained to stop looping on tool calls

28 September 2026
Answer card: Anthropic committed $11.6 billion over seven years to Akamai Cloud for CPU workloads only, with revenue from the second half of 2027.

Anthropic’s $11.6B Akamai deal buys CPUs, not GPUs

27 September 2026
Answer card: Google Suncatcher MVP satellite, four Trillium TPUs on about one kilowatt, launching on SpaceX Transporter-18, reported for 1 October 2026.

Google’s first Suncatcher satellite flies four TPUs on 1 kW

25 September 2026
Answer card for Claude Opus 5.5: 4 dollars in and 20 dollars out per million tokens, down from 5 and 25, with cache reads at 20 cents.

Claude Opus 5.5 drops to $4 and $20, and breaks four Opus 5 habits

23 September 2026
The official xAI announcement card for Grok 4.7, white type on a dark grey and navy gradient.

Grok 4.7 keeps $2 and $6, and its gains over 4.6 are xhigh versus high

22 September 2026
Answer card stating that Qwen-Image-2.1, released on 20 September 2026, ships open weights with a 7 billion parameter diffusion transformer, a Qwen3-VL 8B text encoder and an RGBA VAE totalling about 33 gigabytes in BF16, under the Qwen Research License that limits use to research or evaluation and requires a separate commercial licence, unlike the Apache 2.0 licence of Qwen-Image 1.0.

Qwen-Image-2.1 brings the weights back, but not the Apache licence

21 September 2026
Answer card stating that Ternary Bonsai 2 27B, released by PrismML on 17 September 2026 under Apache 2.0, packs Qwen3.8 27B into 5.95 gigabytes at 1.72 bits per weight, keeps 98.2 percent of the 14-benchmark average, about 75 percent on SWE-bench Verified and Terminal-Bench 2.1, and needs PrismML's llama.cpp fork to run.

Does Bonsai 2 27B really keep 98% of Qwen3.8 in 5.95 GB?

20 September 2026
Answer card stating that Jev 1.13 from TypeSafe AI is a decision model in early access since 15 September 2026 that returns typed probabilities instead of text, priced at 42 dollars per billion input tokens with output tokens free, answering in 70 to 500 milliseconds, with a 64K token request budget, text input only, and a documented list of things it does badly, including counting and dates.

Jev 1.13 bills $42 a billion tokens, and it can’t count

19 September 2026
  • About
  • Contact
  • Privacy
  • Legal
Wednesday, October 7, 2026
  • Login
Packet Nebula
  • Home
  • Articles
    • Security
    • Network
    • Dev
    • Sysadmin
    • SEO
    • Email & DNS
  • Tools
    • Network tools: free, fast, no signup
    • Security tools: free, fast, no signup
    • Developer tools: free, fast, no signup
    • Sysadmin tools: free, fast, no signup
    • SEO tools: free, fast, no signup
    • Email & DNS tools: free, fast, no signup
  • Download
  • About
No Result
View All Result
Packet Nebula
No Result
View All Result
Home Dev

GPT-6.1 Sol costs what GPT-6 Sol did, and only the cache got cheaper

by stephane
30 September 2026
in Dev
0
Answer card for GPT-6.1 Sol showing 2 dollars in and 10 dollars out per million tokens, unchanged from GPT-6 Sol, with cached input at 10 cents.
495
SHARES
1.4k
VIEWS
Share on FacebookShare on Twitter

The headline says a fifth of the price. Of GPT-6 Astra, sure. Against the model it actually replaces, GPT-6.1 Sol costs exactly what GPT-6 Sol did: $2 per million input tokens and $10 per million output. The one line that moved on OpenAI's pricing page is cached input, which halves from $0.20 to $0.10. We don't think that's a small thing, but it isn't the story the launch was sold on.

The short answer

GPT-6.1 Sol launched at DevDay on 29 September 2026 as gpt-6.1-sol. It's $2 in and $10 out per million tokens, like GPT-6 Sol, with cached input cut to $0.10. The window is 1.05M tokens, output tops out at 128K, and prompts over 272K input tokens bill at $4 and $15. OpenAI says it gets close to GPT-6 Astra on coding and computer use.

$2 / $10per 1M, same as GPT-6 Sol
$0.10cached input, down from $0.20
1.05Mcontext window, 128K output
Answer card for GPT-6.1 Sol: the model id gpt-6.1-sol costs 2 dollars in and 10 dollars out per million tokens, unchanged from GPT-6 Sol, with cached input at 10 cents and higher prices past 272K input tokens.
Pricing and limits as listed on OpenAI's API pages on 30 September 2026.

What the price card really says

Here's the full card, read off OpenAI's pricing page this morning. Standard tier: $2 input, $0.10 cached input, $10 output. Cache writes are listed at $2.50, the same as GPT-6 Sol. Past 272K input tokens the whole thing steps up to $4 input, $0.20 cached and $15 output, with cache writes at $5. Batch halves everything. The fast tier doubles it, so $4 and $20. Regional processing adds 10% where it's offered. And one oddity: GPT-6 Sol appears in the Flex table, GPT-6.1 Sol doesn't. Maybe it's coming. Right now it isn't listed.

So where does "a fifth" come from? GPT-6 Astra lists at $10 and $50, which we covered when it launched. Divide by five, you land on Sol. That's true, and it's a fair comparison if you were paying Astra rates for work Sol can now handle. If you were already on GPT-6 Sol, your input and output bill doesn't change by a cent.

The cache cut is where the real saving hides, and it's bigger than it looks for agent loops. A coding agent that re-sends the same 150K token repo context forty times a session is mostly paying cached input. On that shape of workload, halving cached input can move the invoice more than a cheaper output rate would. Honestly, I'd rather have this than a flashy list price drop. Just don't expect it to help a chat app with short prompts and no reuse.

Bar chart of standard cached input prices per million tokens on OpenAI's pricing page: GPT-6 Astra 1 dollar, GPT-5.6 Sol 40 cents, GPT-6 Sol 20 cents, GPT-6.1 Sol 10 cents.
Cached input rates from OpenAI's API pricing page, 30 September 2026. Our chart.

Close to Astra, on OpenAI's numbers

The performance pitch is that 6.1 Sol delivers "nearly the same level of intelligence as GPT-6 Astra" for agentic coding, computer use and professional work. The figures OpenAI published back that up in places. On DeepSWE v1.1 it matches Astra and sits 6.4 points above GPT-6 Sol. On OSWorld 2.0 it lands within 2.1 points of Astra, at roughly a seventh of the cost per task. On Terminal-Bench Science 0.1 the average task cost is $5.47, against $23.80 for Astra. Factual errors at low effort fall from 11.4% to 7.7%.

All vendor numbers, and none of them independent yet. Simon Willison ran his usual quick visual test across the GPT-6 family and found the outputs "not notably different", which tells you about as much as a quick test can. We'd wait for third party evals before moving anything expensive.

The system card addendum is less flattering, and we think it's the part worth reading. In OpenAI's own tests, 6.1 Sol tried to get around restrictions it had been given in 23.5% of cases. That's a big drop from GPT-6 Sol's 64.4%, but still above Astra's 17.4%. On a biology troubleshooting benchmark it scores 47.96% against 63.46% for Astra, a 15.5 point gap. And the model that was supposed to ship alongside it, GPT-6.1 Astra, didn't. TechCrunch, citing The Wall Street Journal, reports it was shelved after safety testing. OpenAI hasn't given a date for it.

Before you switch the model string

It isn't a drop-in swap for every integration. The model page lists reasoning efforts low, medium (the default), high, xhigh and max. There's no none and no minimal, so if your code sends either one, check what you get back before a deploy does it for you. Tool calling needs the Responses API. Chat Completions still works, just without tools. Batch is supported. Knowledge cutoff is 30 April 2026. The window is 1,050,000 tokens and max output is 128,000.

In ChatGPT it's live for Plus, Pro, Business, Enterprise and Edu in ChatGPT Work and in Codex, but not yet in the regular chat. An Ultrafast version for Codex is promised "in the coming days", without a price so far. Against Opus 5.5 at $4 and $20, OpenAI claims 2.2 points more on AutomationBench at about a third of the cost. Again, their number.

Our take: if you're on GPT-6 Sol, switch once your effort settings and tool calls pass a test run, because the cache cut is free money on long agent loops. If you're on Astra, run your own evals first. Two points on a benchmark can be the two points your task needed. A quick check that your key sees the model:

Linux
curl -s https://api.openai.com/v1/models/gpt-6.1-sol -H "Authorization: Bearer $OPENAI_API_KEY"

Sources

Prices for every GPT-6 family model, cache writes, the 272K threshold and the Flex table are from OpenAI's API pricing page. Model ID, window, output limit, cutoff, effort levels and endpoints are from the GPT-6.1 Sol model page. The launch claims are from Introducing GPT-6.1 Sol. Benchmark and cost per task figures are as reported by The Next Web, and ChatGPT availability plus the GPT-6.1 Astra report are from TechCrunch. The quick visual test is from Simon Willison. The figures are ours, built from those pages.

Frequently asked questions

How much does GPT-6.1 Sol cost?

$2 per million input tokens and $10 per million output, with cached input at $0.10 and cache writes at $2.50. Prompts over 272K input tokens bill at $4 and $15. Batch is half price. Those are OpenAI's list prices on 30 September 2026.

Is GPT-6.1 Sol cheaper than GPT-6 Sol?

Only on cached input, which drops from $0.20 to $0.10 per million tokens. Input, output and cache write prices are identical. Workloads that reuse a large prompt many times will see the difference, short one-off requests won't.

Does GPT-6.1 Sol support reasoning effort none?

No. The model page lists low, medium, high, xhigh and max, with medium as the default. The none and minimal settings aren't supported, so check any code that sends them before switching.

Is GPT-6.1 Sol as good as GPT-6 Astra?

Close on OpenAI's own figures for coding and computer use, and clearly behind on its biology troubleshooting test. None of the numbers are independent yet, so we'd test both on your own tasks.

Tags: AI pricinggptllmnewsopenai
Share198Tweet124
stephane

stephane

  • Trending
  • Comments
  • Latest
Answer card: Proton Lumo 2.0 is private by policy, not by locality. Saved history is locked so even Proton cannot read it, but the prompt is decrypted on a Proton EU server to answer it, then forgotten.

Proton Lumo 2.0 review: how private is it, really?

3 September 2026
The Agentic Coding section of the official Hy4 preview benchmark appendix published by Tencent, a table comparing Hy3 and Hy4 preview against DeepSeek V4 Pro 0813, Qwen 3.8 Max, GLM 5.3, Kimi K3, GPT 5.6 Sol and Claude Opus 5 across SWE-bench Multilingual, SWE-bench Pro, DeepSWE, three SWE Atlas tasks, SWE-Marathon, Terminal-Bench 2.1, NL2Repo-Bench, CyberGym, ProgramBench, PostTrainBench and Harbor-Index.

Tencent’s 770B Hy4 tops one benchmark row in 46

3 September 2026
Answer card: Qwen 3.7 Max is API-only and cannot run locally yet; the open Qwen models (Qwen 3.6 27B, qwen3:8b to 32b) run offline via Ollama.

Qwen 3.7 local: what you can actually run offline

22 June 2026
Answer card: JWTs are not encrypted, anyone can read them; the signature proves who issued the token, not who may read it.

Are JWTs encrypted? No, and the difference will bite you

0
Answer card: a random 8 character password falls in under 2 hours offline, while 16 random characters hold for 1.4 trillion years at the same speed.

How long does it take to crack a password in 2026?

0
Answer card: three DNS records decide if your mail lands or bounces; SPF lists allowed senders, DKIM signs messages, DMARC sets the failure policy.

SPF, DKIM and DMARC explained: the records your email needs

0
Answer card: Reflection AI announced Beam, a 501B parameter open-weight mixture of experts model, with weights promised later in October 2026.

Reflection Beam is a 501B open model with no weights yet

6 October 2026
Answer card: Aleph Alpha released Kolibri-1 on 3 October 2026 as an Apache 2.0 mixture of experts model with 78.1 billion total and 3.46 billion active parameters, served at 1 million tokens of context but trained on sequences of up to 256 thousand tokens.

Kolibri-1 serves 1M tokens but was trained to 256K

4 October 2026
Answer card on Gemini 4 Argon pricing and gated access, announced by Google on 30 September 2026.

Gemini 4 Argon costs $2 and $10 now, $4 and $20 later

2 October 2026
  • About
  • Contact
  • Privacy
  • Legal

Copyright © 2026 Stephane Cardon.

Welcome Back!

Login to your account below

Forgotten Password?

Retrieve your password

Please enter your username or email address to reset your password.

Log In
No Result
View All Result
  • Home
  • Articles
    • Security
    • Network
    • Dev
    • Sysadmin
    • SEO
    • Email & DNS
  • Tools
    • Network tools: free, fast, no signup
    • Security tools: free, fast, no signup
    • Developer tools: free, fast, no signup
    • Sysadmin tools: free, fast, no signup
    • SEO tools: free, fast, no signup
    • Email & DNS tools: free, fast, no signup
  • Download
  • About

Copyright © 2026 Stephane Cardon.