• Latest
  • Trending
  • All
Answer card: Meta Muse Spark 1.1 is a multimodal agentic model with a self-managed 1M-token context, launched July 9 at $1.25 input and $4.25 output per million tokens.

Meta ships Muse Spark 1.1 at $1.25 per million tokens

3 September 2026
Answer card stating that Ternary Bonsai 2 27B, released by PrismML on 17 September 2026 under Apache 2.0, packs Qwen3.8 27B into 5.95 gigabytes at 1.72 bits per weight, keeps 98.2 percent of the 14-benchmark average, about 75 percent on SWE-bench Verified and Terminal-Bench 2.1, and needs PrismML's llama.cpp fork to run.

Does Bonsai 2 27B really keep 98% of Qwen3.8 in 5.95 GB?

20 September 2026
Answer card stating that Jev 1.13 from TypeSafe AI is a decision model in early access since 15 September 2026 that returns typed probabilities instead of text, priced at 42 dollars per billion input tokens with output tokens free, answering in 70 to 500 milliseconds, with a 64K token request budget, text input only, and a documented list of things it does badly, including counting and dates.

Jev 1.13 bills $42 a billion tokens, and it can’t count

19 September 2026
Answer card stating that Qwen3.8-Omni-Flash launched on 17 September 2026 as an API only model on Alibaba Cloud Model Studio, taking text, images, audio and video in a 1M token context and returning text only, priced at 0.15 dollars per million input tokens for every modality and 0.47 dollars per million output tokens in the international regions, with no open weights published and the Qwen-Live Harness GitHub repository returning 404.

Qwen3.8-Omni-Flash bills audio at $0.15 and ships no weights

18 September 2026
Answer card stating that on 15 September 2026 AWS said it is unable to restore access to resources and data hosted exclusively in the Middle East Bahrain region me-south-1 and in the mec1-az2 zone of the UAE region, because the damage spanned multiple Availability Zones and exceeded what multi-AZ services are designed to withstand.

AWS can’t restore me-south-1, six months after the drone strikes

17 September 2026
Answer card stating that Google released Gemini 3.8 Live and Gemini 3.8 Live Extended Thinking on 15 September 2026 at 3 dollars per million audio input tokens and 12 dollars out, that the thinking model requires asynchronous tools, and that Artificial Analysis scores it 82.6 on its Speech to Speech Quality Index.

Gemini 3.8 Live Extended Thinking rejects any tool that blocks

16 September 2026
Answer card summarising the Atria Dawn Preview release: 744B GLM-5.2 base, MIT licence, 1.5 TB BF16 and 756 GB FP8 checkpoints, 256K context, top on five of sixteen benchmark rows and trailing on SWE-bench Pro.

Atria Dawn Preview is 744B under MIT, and the BF16 weighs 1.5 TB

15 September 2026
Answer card stating that OpenAI released the Agents API in public beta on 10 September 2026 with no separate fee, billed through model tokens, tool calls and hosted sandbox time, with a choice of OpenAI hosted, self hosted or partner sandboxes, US only data residency and no Zero Data Retention support.

OpenAI’s Agents API has no fee, no ZDR and a one hour sandbox clock

14 September 2026
Answer card: Sakana Fugu Max at $2 and $6 per million tokens, Fugu Ultra v2 unchanged at $5 and $30, and Sakana saying Ultra v2 scores without Fable 5 or GPT-6 Astra in its pool.

Fugu Max costs $2 and $6 while Fugu Ultra v2 runs without Fable 5

13 September 2026
Answer card stating that DeepSeek released DeepSeek-V4.1-Flash on 10 September 2026 as a 552 billion parameter mixture of experts model with a new causal encoder decoder architecture that activates 8 billion parameters on input and 16 billion on output, with native vision, a one million token context and MIT licensed weights, that the API model name is now deepseek-flash at 0.15 dollars per million input tokens and 0.60 dollars per million output tokens off peak, and that DeepSeek announced V4 Pro would be routed to V4.1-Flash from 14 September and reversed that on 11 September.

DeepSeek V4.1-Flash arrived, and the V4 Pro retirement lasted a day

12 September 2026
Answer card stating that Cognition released SWE-2 on 10 September 2026, a coding model post-trained from Kimi K3, scoring 50.0 percent on FrontierCode 1.1 Main against 50.9 percent for Claude Fable 5.1 and 27.3 percent on Terminal-Bench 4 against 55.8 percent, available only inside Devin.

SWE-2 trails Fable 5.1 by one point, and by 28 on Terminal-Bench 4

11 September 2026
Answer card for Meta Muse, free to 100 million tokens a week then $20 a month, launched 8 September 2026 for United States adults only, running in a dedicated per user virtual machine.

Does Meta Muse do enough to earn your inbox and a card on file?

9 September 2026
Answer card stating that the public download pages for the VMware Virtual Disk Development Kit on developer.broadcom.com began returning 404 errors on 25 August 2026 with no announcement or deprecation notice, that Broadcom support tells customers the kit is no longer available for use or download, and that release lines 7.0.3.1, 8.x and 9.x are all affected.

Broadcom pulled VDDK 8.0 and 9.0, and the 404 is the only notice

8 September 2026
  • About
  • Contact
  • Privacy
  • Legal
Sunday, September 20, 2026
  • Login
Packet Nebula
  • Home
  • Articles
    • Security
    • Network
    • Dev
    • Sysadmin
    • SEO
    • Email & DNS
  • Tools
    • Network tools: free, fast, no signup
    • Security tools: free, fast, no signup
    • Developer tools: free, fast, no signup
    • Sysadmin tools: free, fast, no signup
    • SEO tools: free, fast, no signup
    • Email & DNS tools: free, fast, no signup
  • Download
  • About
No Result
View All Result
Packet Nebula
No Result
View All Result
Home Dev

Meta ships Muse Spark 1.1 at $1.25 per million tokens

by stephane
3 September 2026
in Dev
0
Answer card: Meta Muse Spark 1.1 is a multimodal agentic model with a self-managed 1M-token context, launched July 9 at $1.25 input and $4.25 output per million tokens.
496
SHARES
1.4k
VIEWS
Share on FacebookShare on Twitter

Meta spent years as the lab that gave its weights away, so watching it stand up a paid API is a small plot twist. On July 9 it shipped Muse Spark 1.1, and here's the one-liner we'd give a teammate: it's a multimodal reasoning model built for agents. A 1M-token context it manages itself. Multi-agent orchestration, plus computer use across apps. The hook isn't a benchmark, though. It's the invoice. $1.25 per million input tokens and $4.25 output, which lands well under OpenAI and Anthropic for the same rough tier. Meta's own numbers put it top of the tool-use charts, past Claude Opus 4.8 there, while it trails on the hard long-horizon coding test. So the real question isn't whether Meta can ship a frontier-ish model. It clearly can. It's whether the price makes the gaps worth living with, and whether you can even get at it yet.

The short answer

Muse Spark 1.1 landed July 9 as Meta’s first paid model API: a multimodal agentic model with a self-managed 1M-token context, multi-agent orchestration and computer use. It tops Meta’s tool-use benchmarks and beats Opus 4.8 there, but trails GPT-5.5 on long-horizon coding. The price is the pitch. The catch, for most of us: the developer API is a US-only preview at launch.

$1.25 / $4.25per 1M tokens
1Mtoken context, self-managed
US previewAPI, for now
Answer card: Meta Muse Spark 1.1 is a multimodal agentic model with a self-managed 1M-token context, launched July 9 at $1.25 input and $4.25 output per million tokens.
Meta's first proper paid API. The headline number is a price, not a score.

The pitch is the price

Meta building a paid model API at all is the story under the story. This is the company whose whole AI identity was open weights you could download and run yourself. Muse Spark 1.1 keeps a free consumer face, it runs in Thinking mode inside the Meta AI app and on meta.ai, but the new thing is a metered endpoint you pay for by the token. Joel Kaplan, Meta’s policy chief, framed it as “high intelligence at one of the best price points in the market,” which is marketing, sure, but the numbers back the price part.

$1.25 per million input tokens. $4.25 output. For a model pitched at agent work, that sits comfortably under the frontier offerings from OpenAI and Anthropic, and Meta throws in $20 of free credits to get you in the door. It is not the absolute cheapest thing on the market. It is cheap for what it claims to be, which is a different and more interesting position.

Bar chart of Meta-reported scores: MCP Atlas tool use, Muse Spark 1.1 88.1 vs GPT-5.5 75.3; DeepSWE 1.1 long-horizon coding, Muse Spark 1.1 53.3 vs GPT-5.5 67.0.
Same model, two very different results depending on the task. Meta-reported.

What “agentic” actually means here

Every model calls itself agentic now, so the word has gone soft. Muse Spark 1.1 attaches some concrete things to it. First, that context window. A million tokens is not new on its own, plenty of models quote big numbers. What Meta claims is that the model actively manages the window: it recalls actions from early in a run and compacts the history so the important steps survive instead of scrolling off the top. For a long agent loop, that management is the part that usually breaks.

Second, orchestration. Muse Spark 1.1 is trained to run as a lead agent that plans and delegates across parallel subagents, and also to be a subagent itself, one that does a defined job and knows when to escalate. And computer use: it works across multiple apps with data changing under it, choosing between writing a script to automate something and just driving the interface directly.

On paper that lines up with where its benchmarks are strong. Meta reports 88.1 on MCP Atlas, its tool-use test, against 82.2 for Claude Opus 4.8 and 75.3 for GPT-5.5. Tool use is exactly what an orchestration model needs to nail, so the shape of the results at least matches the pitch.

Meta's official launch artwork for Muse Spark 1.1: blue particle streams sweeping across a black field, with the Muse Spark 1.1 wordmark.

Image: Meta. Official Muse Spark 1.1 launch artwork.

Where it trails, and the catch

Now the honest column. On DeepSWE 1.1, the long-horizon coding test that measures resolving real, multi-step issues, Muse Spark 1.1 scores 53.3 against GPT-5.5’s 67.0. That is a real gap, not rounding. So the model that tops the tool-use chart sits well down the field on the hardest sustained coding. Read together, the picture is a model that’s excellent at driving tools and agents, and merely okay when the code change itself is deep and sprawling.

Two more things temper the excitement. The benchmarks are Meta-reported, run at launch, and nobody outside has reproduced them yet. Treat them as a starting claim, not a verdict. And the practical blocker: the developer API shipped as a public preview for US developers. If you’re building from Europe or most of the rest of the world, the paid endpoint may not be open to you yet, whatever the price says. The free consumer app is wider, but you can’t build a product on the chat window.

Output is text only, too. Image and audio go in, only text comes back, so this is not the model for generating media. And there’s no full model card at launch, which is a bit rich for a company that built its reputation on openness.

Checklist of Muse Spark 1.1 strengths and catches: cheap pricing, self-managed 1M context, multi-agent orchestration, tops MCP Atlas tool use; trails on DeepSWE long-horizon coding, US-only preview, text-only output, no full model card.
The short version: strong value and real agent skills, with a few gaps that matter.

So should you switch?

If you run agents at volume and you’re in the US, Muse Spark 1.1 is worth an afternoon on your own workload, because the combination it offers, cheap tokens plus genuine tool-use and orchestration strength, is rare and the price makes the experiment cheap. Just don’t take the coding claims on faith. Run it on the messy, multi-file task you actually care about before you move anything real onto it.

If your bottleneck is the hardest coding, the deep sprawling change, the numbers still point elsewhere for now. The cheaper end of the field is getting crowded fast, and Meta just walked in from a direction nobody was watching. We wrote up the other budget disruptor, xAI’s coding-first model, in xAI Grok 4.5 is here, and if you’re weighing the Anthropic side on price versus power, our Sonnet 5 vs Opus 4.8 breakdown covers that trade. The pattern across all of them is the same: the interesting pressure on the frontier is coming from underneath, on price.

Sources: Meta’s Muse Spark 1.1 announcement on the Meta AI blog, plus pricing, benchmark and availability reporting from DataCamp, PPC Land and Crypto Briefing, July 9 to 13 2026. All benchmark figures are Meta-reported and not yet independently reproduced; pricing, regional availability and context limits may change.

Frequently asked questions

How much does Muse Spark 1.1 cost?

Through the Meta Model API it is $1.25 per million input tokens and $4.25 per million output. New sign-ups get $20 in free credits before pay-as-you-go kicks in. Meta positions that above GPT-5 mini and Claude Haiku 4.5 on price, but below Claude Sonnet 4.8, so it is a mid-tier rate for a model aimed at agent work.

Is Muse Spark 1.1 good at coding?

It depends which coding. On tool use and computer-use tasks it leads: Meta reports 88.1 on MCP Atlas against 82.2 for Claude Opus 4.8. On the long-horizon coding test DeepSWE 1.1 it scores 53.3, behind GPT-5.5's 67.0. So it is strong at driving tools and agents, weaker on sprawling multi-step code changes. All those numbers are Meta-reported and not yet independently reproduced.

What does the 1M-token context actually do?

Muse Spark 1.1 actively manages a context window of up to a million tokens. It remembers earlier actions, pulls back information from much earlier in a run, and compacts the history so the steps that matter for later work survive. For a long agent run that reads files, calls tools and edits code, that is the difference between staying coherent and losing the thread halfway through.

Can I use Muse Spark 1.1 in Europe?

The consumer version runs free in Thinking mode inside the Meta AI app and at meta.ai with a Meta login. The developer API, though, launched July 9 as a public preview for US developers. If you are outside the US and want to build on it, you may be waiting until Meta widens access.

What is Muse Spark 1.1 built for?

Agentic work rather than chat. It is trained to orchestrate multiple agents, acting as a lead that plans and delegates to parallel subagents, or as a subagent that does one job and reports back. It also handles computer use, picking between writing a script and clicking through an interface. Text, image and audio go in; only text comes out.

Tags: agentsaillmmetanewspricing
Share198Tweet124
stephane

stephane

  • Trending
  • Comments
  • Latest
Answer card: Proton Lumo 2.0 is private by policy, not by locality. Saved history is locked so even Proton cannot read it, but the prompt is decrypted on a Proton EU server to answer it, then forgotten.

Proton Lumo 2.0 review: how private is it, really?

3 September 2026
The Agentic Coding section of the official Hy4 preview benchmark appendix published by Tencent, a table comparing Hy3 and Hy4 preview against DeepSeek V4 Pro 0813, Qwen 3.8 Max, GLM 5.3, Kimi K3, GPT 5.6 Sol and Claude Opus 5 across SWE-bench Multilingual, SWE-bench Pro, DeepSWE, three SWE Atlas tasks, SWE-Marathon, Terminal-Bench 2.1, NL2Repo-Bench, CyberGym, ProgramBench, PostTrainBench and Harbor-Index.

Tencent’s 770B Hy4 tops one benchmark row in 46

3 September 2026
Answer card: Qwen 3.7 Max is API-only and cannot run locally yet; the open Qwen models (Qwen 3.6 27B, qwen3:8b to 32b) run offline via Ollama.

Qwen 3.7 local: what you can actually run offline

22 June 2026
Answer card: JWTs are not encrypted, anyone can read them; the signature proves who issued the token, not who may read it.

Are JWTs encrypted? No, and the difference will bite you

0
Answer card: a random 8 character password falls in under 2 hours offline, while 16 random characters hold for 1.4 trillion years at the same speed.

How long does it take to crack a password in 2026?

0
Answer card: three DNS records decide if your mail lands or bounces; SPF lists allowed senders, DKIM signs messages, DMARC sets the failure policy.

SPF, DKIM and DMARC explained: the records your email needs

0
Answer card stating that Ternary Bonsai 2 27B, released by PrismML on 17 September 2026 under Apache 2.0, packs Qwen3.8 27B into 5.95 gigabytes at 1.72 bits per weight, keeps 98.2 percent of the 14-benchmark average, about 75 percent on SWE-bench Verified and Terminal-Bench 2.1, and needs PrismML's llama.cpp fork to run.

Does Bonsai 2 27B really keep 98% of Qwen3.8 in 5.95 GB?

20 September 2026
Answer card stating that Jev 1.13 from TypeSafe AI is a decision model in early access since 15 September 2026 that returns typed probabilities instead of text, priced at 42 dollars per billion input tokens with output tokens free, answering in 70 to 500 milliseconds, with a 64K token request budget, text input only, and a documented list of things it does badly, including counting and dates.

Jev 1.13 bills $42 a billion tokens, and it can’t count

19 September 2026
Answer card stating that Qwen3.8-Omni-Flash launched on 17 September 2026 as an API only model on Alibaba Cloud Model Studio, taking text, images, audio and video in a 1M token context and returning text only, priced at 0.15 dollars per million input tokens for every modality and 0.47 dollars per million output tokens in the international regions, with no open weights published and the Qwen-Live Harness GitHub repository returning 404.

Qwen3.8-Omni-Flash bills audio at $0.15 and ships no weights

18 September 2026
  • About
  • Contact
  • Privacy
  • Legal

Copyright © 2026 Stephane Cardon.

Welcome Back!

Login to your account below

Forgotten Password?

Retrieve your password

Please enter your username or email address to reset your password.

Log In
No Result
View All Result
  • Home
  • Articles
    • Security
    • Network
    • Dev
    • Sysadmin
    • SEO
    • Email & DNS
  • Tools
    • Network tools: free, fast, no signup
    • Security tools: free, fast, no signup
    • Developer tools: free, fast, no signup
    • Sysadmin tools: free, fast, no signup
    • SEO tools: free, fast, no signup
    • Email & DNS tools: free, fast, no signup
  • Download
  • About

Copyright © 2026 Stephane Cardon.