• Latest
  • Trending
  • All
Answer card: Alibaba published the Qwen3.8-27B weights on 14 August 2026 under Apache 2.0 as a dense native vision language model with 262,144 tokens of native context, while the 2.4 trillion parameter flagship weights published two days earlier carry a custom qwen3.8-max licence and are text only.

Qwen3.8-27B ships Apache 2.0 with vision, the 2.4T doesn’t

15 August 2026
Answer card stating that Mullvad announced on 3 September 2026 that it is shutting down its public encrypted domain name system servers on 2 November 2026 and sponsoring the Quad9 Foundation instead, with 194.242.2.2 and its five sibling addresses all going away, and virtual private network customers unaffected.

Mullvad’s DNS servers go dark on 2 November, and Quad9 blocks no ads

5 September 2026
OpenAI announcement image for GPT-6 Astra, a spiral galaxy of white, blue and amber points of light curling around a bright core on a near black star field.

GPT-6 Astra lists at $10 and $50, 2.5x what GPT-5.6 Sol costs

6 September 2026
Google's official announcement image for the release, reading Introducing Gemini 3.8 Flash and 3.8 Flash Cyber in black type over a pale blue background with a blurred white chevron and the four colour Gemini spark below.

Gemini 3.8 Flash keeps the price and the 1 January cliff

3 September 2026
Answer card stating that Anthropic announced Enterprise Frontier Safeguards on 1 September 2026, that activity data used for misuse monitoring moves into cloud storage the customer controls under the customer own encryption keys, that Anthropic charges nothing for the feature while the cloud provider bills storage and egress, and that the phased rollout starts later in autumn 2026 with interim zero data retention on Fable 5 and Fable 5.1 for eligible customers.

Anthropic moves retention into your own cloud, for 30 days

3 September 2026
Official Google diagram of a client connection in three numbered steps: a DNS lookup with a query and an address, a TLS ClientHello and ServerHello, then a content exchange with a website. A callout on the DNS step reads 25% of global web traffic is now protected by encrypted DNS, and a callout beside an Android phone on the ClientHello step reads Android 17 supports ECH GREASE by default.

Android 17 hides the SNI, not your DNS or destination

3 September 2026
Still frame from the Claude Fable 5.1 launch video showing model-designed protein binders in orange docked against twelve grey target proteins, rendered as ESMFold2 structure predictions.

Claude Fable 5.1 breaks forced tool use, cuts cache 75%

1 September 2026
Answer card stating that on 31 August 2026 the European Commission designated ChatGPT a Very Large Online Search Engine under the Digital Services Act, the first conversational AI service classified that way, because it answers user prompts and queries including by searching the web, with OpenAI having declared roughly 159.1 million average monthly users in the European Union for ChatGPT search.

The EU now calls ChatGPT a very large search engine

3 September 2026
Answer card stating that on 31 August 2026 the Department of War added OpenAI ChatGPT Mil and Starshield AI Grok for Government to the GenAI.mil portal alongside Google Gemini, all three accredited at Impact Level 5 for Controlled Unclassified Information, with 1.7 million unique users onboarded out of roughly 3 million eligible personnel, and ChatGPT Mil currently serving GPT-5.4 Terra with GPT-5.6 Terra said to be rolling out.

ChatGPT Mil and Grok reached IL5 on GenAI.mil

3 September 2026
Answer card stating that Anthropic opened a research preview of the Model Hardware Standard on 27 August 2026, standardising the driver layer between an operating system and a laboratory instrument with read and write primitives plus discovery and safety limits, reachable through MCP as well as a command line and code files, with no public specification published.

Anthropic’s Model Hardware Standard is gated, and sits under MCP

3 September 2026
Official Cohere key art for the Parse 5 launch: the Cohere mark and the wordmark Parse with a superscript 5 in white, centred on a soft out of focus gradient of deep blue, violet and amber curves.

Cohere Parse 5 is $1.50 per 1,000 pages, on three of five dimensions

3 September 2026
Title card from the OpenAI announcement video: a man sits on a blue sofa in a loft with tall windows and potted plants, a laptop open on the coffee table in front of him, with the words WebMCP in ChatGPT in large white type across the lower left.

WebMCP in ChatGPT needs GPT-5.6 Sol or Terra

3 September 2026
The Agentic Coding section of the official Hy4 preview benchmark appendix published by Tencent, a table comparing Hy3 and Hy4 preview against DeepSeek V4 Pro 0813, Qwen 3.8 Max, GLM 5.3, Kimi K3, GPT 5.6 Sol and Claude Opus 5 across SWE-bench Multilingual, SWE-bench Pro, DeepSWE, three SWE Atlas tasks, SWE-Marathon, Terminal-Bench 2.1, NL2Repo-Bench, CyberGym, ProgramBench, PostTrainBench and Harbor-Index.

Tencent’s 770B Hy4 tops one benchmark row in 46

3 September 2026
  • About
  • Contact
  • Privacy
  • Legal
Sunday, September 6, 2026
  • Login
Packet Nebula
  • Home
  • Articles
    • Security
    • Network
    • Dev
    • Sysadmin
    • SEO
    • Email & DNS
  • Tools
    • Network tools: free, fast, no signup
    • Security tools: free, fast, no signup
    • Developer tools: free, fast, no signup
    • Sysadmin tools: free, fast, no signup
    • SEO tools: free, fast, no signup
    • Email & DNS tools: free, fast, no signup
  • Download
  • About
No Result
View All Result
Packet Nebula
No Result
View All Result
Home Dev

Qwen3.8-27B ships Apache 2.0 with vision, the 2.4T doesn’t

by stephane
15 August 2026
in Dev
0
Answer card: Alibaba published the Qwen3.8-27B weights on 14 August 2026 under Apache 2.0 as a dense native vision language model with 262,144 tokens of native context, while the 2.4 trillion parameter flagship weights published two days earlier carry a custom qwen3.8-max licence and are text only.
491
SHARES
1.4k
VIEWS
Share on FacebookShare on Twitter

Open the two model cards side by side and the surprise is which one you would rather self host. Qwen3.8-27B went up on Hugging Face on 14 August under Apache 2.0: dense, with a vision encoder nobody had been promised, 262,144 tokens of native context, and a reasoning effort dial you can actually turn down. The 2.4 trillion parameter flagship whose weights landed two days earlier carries a bespoke licence called qwen3.8-max, and its own card says multimodal inputs are not supported and thinking cannot be disabled. So the small one is the permissive one. It is also the multimodal one. We went looking for the catch, and the catch is mostly VRAM.

The short answer

Alibaba shipped two sets of Qwen3.8 weights inside three days. The 2.4T flagship came first, text only, under a bespoke licence. Then the 27B landed under Apache 2.0 with a native vision encoder. If you self host, the second one is the interesting release, and it is the one that fits on hardware you might already own.

Apache 2.0on the 27B, custom licence on the 2.4T
visionon the 27B only
~28 GBFP8 weights, one 48 GB card
Answer card: Alibaba published the Qwen3.8-27B weights on Hugging Face on 14 August 2026 under Apache 2.0, a dense native vision language model with 262,144 tokens of native context extensible to one million, while the 2.4 trillion parameter flagship weights published two days earlier carry a custom qwen3.8-max licence and are text only.
Two weight drops, two days apart, and the small one has the better terms.

What actually landed

Back on 4 August we wrote that Qwen3.8-Max was not open source, that it was API only, and that the weights were promised for the following week with no licence named. Both halves of that promise have now been kept, sort of.

The flagship weights went up around 12 August as Qwen/Qwen3.8-2.4T-A95B. Then on 14 August the 27B appeared, and it is the one that got the attention.

Here is the shape of it. Qwen3.8-27B is dense, not sparse. Sixty four layers, built as sixteen repeats of a block that mixes Gated DeltaNet with Gated Attention. Native context of 262,144 tokens, extensible to a million. And a vision encoder, which is the part nobody had been trailing: the card calls it a native vision language model that understands images and videos.

Side by side comparison of the two Qwen3.8 open weight releases: the 27B under Apache 2.0 with a native vision encoder and configurable reasoning effort, against the 2.4T mixture of experts flagship under the custom qwen3.8-max licence, text only and with thinking forced on.
Same family, same context window, and a licence field that diverges hard.

The licence is the line I would read first. Apache 2.0 on the 27B, which is boring in the best way: your legal team already has a position on it. The flagship card names qwen3.8-max instead, a bespoke licence, and reporting around that launch describes a revenue share aimed at large commercial users with neither the threshold nor the percentage published. Unpublished terms are not terms you can plan around, so I would treat the flagship weights as unusable for anything commercial until the actual text is out.

The bit that reads backwards

Normally the flagship gets the features and the small model gets the leftovers. Not here.

The 2.4T card is explicit: multimodal inputs are not supported, and thinking cannot be disabled. Both of those are real constraints. Forced thinking means every call pays the reasoning tokens whether the task needs them or not, which is exactly the problem we flagged on GLM-5.3 last week when Z.ai removed non-thinking mode. The 27B keeps the dial, with reasoning_effort at xhigh, medium or low.

So the downloadable flagship is a text only model that always thinks, and the downloadable 27B sees images and lets you turn the thinking down. Honestly, for most self hosted work that is the wrong way round from what the parameter counts suggest, and it is a good outcome.

The numbers Alibaba published

First real benchmark table for the Qwen3.8 family, incidentally. The July preview shipped with a tweet and no card at all.

On the 27B: SWE-bench Pro 61.7, Terminal Bench 2.1 73.0, LiveCodeBench v6 90.3, IFBench 79.5. On the vision side, OSWorld-Verified 84.3 and WebArena-Verified 64.8, with MathVision at 94.6.

The flagship posts higher where you would expect: SWE-bench Pro 67.7, Terminal Bench 2.1 86.6, GPQA Diamond 92.6, PaperBench 93.0. Six points of SWE-bench Pro between a 27B dense model and a 2.4 trillion parameter mixture of experts is a narrower gap than the parameter counts imply, and the usual caveat applies with full force. These are vendor numbers on vendor runs. Nobody outside Alibaba has reproduced any of them yet.

What I would take from the table is the shape rather than the digits. A 27B that lands in the low sixties on SWE-bench Pro and above eighty on OSWorld is a genuinely capable local agent model, and OSWorld is the computer-use benchmark, which is where the vision encoder earns its place.

Whether you can run it

This is the part that decides it for most people.

Bar chart of approximate VRAM for the Qwen3.8-27B weights at three precisions: about 56 GB at BF16 needing an 80 GB card, about 28 GB at FP8 fitting a 48 GB card, and about 14 to 17 GB at 4-bit fitting a 24 GB consumer card.
Weights only. The KV cache is the line item that will actually catch you out.

Alibaba publishes an FP8 build itself, Qwen3.8-27B-FP8, using fine grained quantisation at block size 128, and says the metrics come out nearly identical to the original. That build is compatible with Transformers, vLLM and SGLang. Community quantisations for llama.cpp, Ollama, LM Studio and Jan appeared within a day, as they always do.

Rough weight footprints: about 56 GB at BF16, about 28 GB at FP8, somewhere in the 14 to 17 GB range at 4-bit. Those are community estimates, not vendor figures, and they cover the weights only.

Then the cache. A 262k native window sounds like a gift until you price the KV cache that fills it, and at long context with a few concurrent requests the cache can outweigh the weights. If you are sizing a box on the 4-bit number because it fits a 24 GB card, budget for a short context or a single stream. That is the constraint people keep discovering after the hardware arrives, and it is the reason we keep saying the same thing about local models: the parameter count tells you almost nothing about what you need.

Would we run it

If you have a 48 GB card, yes, the FP8 build is the obvious thing to try this week. Apache 2.0, vision included, and a coding score that is respectable for the size.

If you were waiting on the 2.4T weights so you could self host the frontier model, the honest answer is that what shipped is not what the hosted product is. Text only, thinking forced, licence unpublished in its details. It exists, you can download it, and I am not sure who it is for yet.

And if you are picking a local model for agent work specifically, the OSWorld number plus the vision encoder is the argument for this one over a text only competitor of similar size. Worth a weekend. We will run it against Qwen3.7 on local hardware once there is something to compare beyond Alibaba’s own table.

Sources

Alibaba, Qwen3.8-27B-FP8 model card on Hugging Face, for the architecture, the Apache 2.0 licence, the 262,144 token native context, the native vision language description, the reasoning effort levels, the FP8 block size 128 quantisation and the benchmark table. Alibaba, Qwen3.8-2.4T-A95B model card, for the qwen3.8-max licence field, the 95B active parameters out of 512 experts, the statement that multimodal inputs are not supported and that thinking cannot be disabled, and its own benchmark figures. Yotta Labs, Qwen 3.8 27B specs and hardware requirements, for the VRAM estimates by precision. ExplainX, Qwen3.8-Max open weights are live, for the 12 August flagship weights date and the reporting on the revenue share clause, whose threshold and percentage are not published.

Frequently asked questions

What licence is Qwen3.8-27B under?

Apache 2.0. The licence field on the Qwen/Qwen3.8-27B-FP8 model card reads apache-2.0, which allows commercial use, modification and redistribution with attribution and no revenue share. That is a different licence from the flagship weights: the Qwen/Qwen3.8-2.4T-A95B card names a custom licence called qwen3.8-max instead.

Can Qwen3.8-27B read images?

Yes. Alibaba describes it on the card as a native vision language model that understands images and videos, and publishes vision results alongside the text ones, including 84.3 on OSWorld-Verified and 64.8 on WebArena-Verified. The 2.4T flagship weights cannot: that card states plainly that multimodal inputs are not supported.

How much VRAM does Qwen3.8-27B need?

Roughly 56 GB at BF16, about 28 GB at FP8 and somewhere around 14 to 17 GB at 4-bit, for the weights alone. The KV cache is extra and it grows with context length and concurrency, which matters here because the native window is 262,144 tokens. Alibaba does not publish VRAM figures itself, so treat these as the community estimates they are.

Is Qwen3.8-27B dense or mixture of experts?

Dense. The card describes 64 layers built as 16 repeats of a block mixing Gated DeltaNet and Gated Attention with feed forward layers. The 2.4T flagship is the sparse one, a fine grained mixture of experts with 95B parameters active per token out of 512 experts.

What context length does it support?

The card says 262,144 tokens natively, extensible up to 1,000,000. Both Qwen3.8 open weight releases quote the same pair of numbers, so context is not what separates them. The licence and the vision support are.

Tags: aialibaballmnewsopen-weightsqwen
Share196Tweet123
stephane

stephane

  • Trending
  • Comments
  • Latest
Answer card: Proton Lumo 2.0 is private by policy, not by locality. Saved history is locked so even Proton cannot read it, but the prompt is decrypted on a Proton EU server to answer it, then forgotten.

Proton Lumo 2.0 review: how private is it, really?

3 September 2026
Google's official announcement image for the release, reading Introducing Gemini 3.8 Flash and 3.8 Flash Cyber in black type over a pale blue background with a blurred white chevron and the four colour Gemini spark below.

Gemini 3.8 Flash keeps the price and the 1 January cliff

3 September 2026
Answer card stating that Anthropic announced Enterprise Frontier Safeguards on 1 September 2026, that activity data used for misuse monitoring moves into cloud storage the customer controls under the customer own encryption keys, that Anthropic charges nothing for the feature while the cloud provider bills storage and egress, and that the phased rollout starts later in autumn 2026 with interim zero data retention on Fable 5 and Fable 5.1 for eligible customers.

Anthropic moves retention into your own cloud, for 30 days

3 September 2026
Answer card: JWTs are not encrypted, anyone can read them; the signature proves who issued the token, not who may read it.

Are JWTs encrypted? No, and the difference will bite you

0
Answer card: a random 8 character password falls in under 2 hours offline, while 16 random characters hold for 1.4 trillion years at the same speed.

How long does it take to crack a password in 2026?

0
Answer card: three DNS records decide if your mail lands or bounces; SPF lists allowed senders, DKIM signs messages, DMARC sets the failure policy.

SPF, DKIM and DMARC explained: the records your email needs

0
Answer card stating that Mullvad announced on 3 September 2026 that it is shutting down its public encrypted domain name system servers on 2 November 2026 and sponsoring the Quad9 Foundation instead, with 194.242.2.2 and its five sibling addresses all going away, and virtual private network customers unaffected.

Mullvad’s DNS servers go dark on 2 November, and Quad9 blocks no ads

5 September 2026
OpenAI announcement image for GPT-6 Astra, a spiral galaxy of white, blue and amber points of light curling around a bright core on a near black star field.

GPT-6 Astra lists at $10 and $50, 2.5x what GPT-5.6 Sol costs

6 September 2026
Google's official announcement image for the release, reading Introducing Gemini 3.8 Flash and 3.8 Flash Cyber in black type over a pale blue background with a blurred white chevron and the four colour Gemini spark below.

Gemini 3.8 Flash keeps the price and the 1 January cliff

3 September 2026
  • About
  • Contact
  • Privacy
  • Legal

Copyright © 2026 Stephane Cardon.

Welcome Back!

Login to your account below

Forgotten Password?

Retrieve your password

Please enter your username or email address to reset your password.

Log In
No Result
View All Result
  • Home
  • Articles
    • Security
    • Network
    • Dev
    • Sysadmin
    • SEO
    • Email & DNS
  • Tools
    • Network tools: free, fast, no signup
    • Security tools: free, fast, no signup
    • Developer tools: free, fast, no signup
    • Sysadmin tools: free, fast, no signup
    • SEO tools: free, fast, no signup
    • Email & DNS tools: free, fast, no signup
  • Download
  • About

Copyright © 2026 Stephane Cardon.