• Latest
  • Trending
  • All
Answer card: Tencent open-sourced Hy3, an Apache 2.0 licensed 295B mixture-of-experts model that runs only 21B active parameters per token, with 256K context, free on OpenRouter until July 21.

Is Tencent’s 295B Hy3 really free?

3 September 2026
Answer card stating that Mullvad announced on 3 September 2026 that it is shutting down its public encrypted domain name system servers on 2 November 2026 and sponsoring the Quad9 Foundation instead, with 194.242.2.2 and its five sibling addresses all going away, and virtual private network customers unaffected.

Mullvad’s DNS servers go dark on 2 November, and Quad9 blocks no ads

5 September 2026
OpenAI announcement image for GPT-6 Astra, a spiral galaxy of white, blue and amber points of light curling around a bright core on a near black star field.

GPT-6 Astra lists at $10 and $50, 2.5x what GPT-5.6 Sol costs

6 September 2026
Google's official announcement image for the release, reading Introducing Gemini 3.8 Flash and 3.8 Flash Cyber in black type over a pale blue background with a blurred white chevron and the four colour Gemini spark below.

Gemini 3.8 Flash keeps the price and the 1 January cliff

3 September 2026
Answer card stating that Anthropic announced Enterprise Frontier Safeguards on 1 September 2026, that activity data used for misuse monitoring moves into cloud storage the customer controls under the customer own encryption keys, that Anthropic charges nothing for the feature while the cloud provider bills storage and egress, and that the phased rollout starts later in autumn 2026 with interim zero data retention on Fable 5 and Fable 5.1 for eligible customers.

Anthropic moves retention into your own cloud, for 30 days

3 September 2026
Official Google diagram of a client connection in three numbered steps: a DNS lookup with a query and an address, a TLS ClientHello and ServerHello, then a content exchange with a website. A callout on the DNS step reads 25% of global web traffic is now protected by encrypted DNS, and a callout beside an Android phone on the ClientHello step reads Android 17 supports ECH GREASE by default.

Android 17 hides the SNI, not your DNS or destination

3 September 2026
Still frame from the Claude Fable 5.1 launch video showing model-designed protein binders in orange docked against twelve grey target proteins, rendered as ESMFold2 structure predictions.

Claude Fable 5.1 breaks forced tool use, cuts cache 75%

1 September 2026
Answer card stating that on 31 August 2026 the European Commission designated ChatGPT a Very Large Online Search Engine under the Digital Services Act, the first conversational AI service classified that way, because it answers user prompts and queries including by searching the web, with OpenAI having declared roughly 159.1 million average monthly users in the European Union for ChatGPT search.

The EU now calls ChatGPT a very large search engine

3 September 2026
Answer card stating that on 31 August 2026 the Department of War added OpenAI ChatGPT Mil and Starshield AI Grok for Government to the GenAI.mil portal alongside Google Gemini, all three accredited at Impact Level 5 for Controlled Unclassified Information, with 1.7 million unique users onboarded out of roughly 3 million eligible personnel, and ChatGPT Mil currently serving GPT-5.4 Terra with GPT-5.6 Terra said to be rolling out.

ChatGPT Mil and Grok reached IL5 on GenAI.mil

3 September 2026
Answer card stating that Anthropic opened a research preview of the Model Hardware Standard on 27 August 2026, standardising the driver layer between an operating system and a laboratory instrument with read and write primitives plus discovery and safety limits, reachable through MCP as well as a command line and code files, with no public specification published.

Anthropic’s Model Hardware Standard is gated, and sits under MCP

3 September 2026
Official Cohere key art for the Parse 5 launch: the Cohere mark and the wordmark Parse with a superscript 5 in white, centred on a soft out of focus gradient of deep blue, violet and amber curves.

Cohere Parse 5 is $1.50 per 1,000 pages, on three of five dimensions

3 September 2026
Title card from the OpenAI announcement video: a man sits on a blue sofa in a loft with tall windows and potted plants, a laptop open on the coffee table in front of him, with the words WebMCP in ChatGPT in large white type across the lower left.

WebMCP in ChatGPT needs GPT-5.6 Sol or Terra

3 September 2026
The Agentic Coding section of the official Hy4 preview benchmark appendix published by Tencent, a table comparing Hy3 and Hy4 preview against DeepSeek V4 Pro 0813, Qwen 3.8 Max, GLM 5.3, Kimi K3, GPT 5.6 Sol and Claude Opus 5 across SWE-bench Multilingual, SWE-bench Pro, DeepSWE, three SWE Atlas tasks, SWE-Marathon, Terminal-Bench 2.1, NL2Repo-Bench, CyberGym, ProgramBench, PostTrainBench and Harbor-Index.

Tencent’s 770B Hy4 tops one benchmark row in 46

3 September 2026
  • About
  • Contact
  • Privacy
  • Legal
Sunday, September 6, 2026
  • Login
Packet Nebula
  • Home
  • Articles
    • Security
    • Network
    • Dev
    • Sysadmin
    • SEO
    • Email & DNS
  • Tools
    • Network tools: free, fast, no signup
    • Security tools: free, fast, no signup
    • Developer tools: free, fast, no signup
    • Sysadmin tools: free, fast, no signup
    • SEO tools: free, fast, no signup
    • Email & DNS tools: free, fast, no signup
  • Download
  • About
No Result
View All Result
Packet Nebula
No Result
View All Result
Home Dev

Is Tencent’s 295B Hy3 really free?

by stephane
3 September 2026
in Dev
0
Answer card: Tencent open-sourced Hy3, an Apache 2.0 licensed 295B mixture-of-experts model that runs only 21B active parameters per token, with 256K context, free on OpenRouter until July 21.
491
SHARES
1.4k
VIEWS
Share on FacebookShare on Twitter

Another giant model landed on Hugging Face, and this one is worth stopping for. On July 6 Tencent open-sourced Hy3, a 295B mixture-of-experts model under a plain Apache 2.0 license, and the one-line version we'd give a colleague is this: it's genuinely open, genuinely large, and only lights up 21B of its 295B parameters on any given token. It's free to try on OpenRouter through July 21. It also isn't running on your gaming PC, and that gap is the whole story. On paper it edges past GLM-5.1 in Tencent's own testing and posts strong reasoning scores. So the question isn't whether it looks good. It's whether it earns a slot in your stack.

The short answer

Tencent open-sourced Hy3, a 295B mixture-of-experts model with 21B active parameters and 256K context, under a permissive Apache 2.0 license. It posts strong reasoning scores and, by Tencent’s own blind test, edges past GLM-5.1. The catch is hardware: Tencent points at 8 big GPUs to serve it. Free to try on OpenRouter right now, free to self-host if you have the metal.

295B21B active per token
Apache 2.0weights, commercial ok
July 21free on OpenRouter until
Answer card: Tencent open-sourced Hy3, an Apache 2.0 licensed 295B mixture-of-experts model that runs only 21B active parameters per token, 256K context, free on OpenRouter until July 21.
The one-card version. Open weights, huge on disk, but only a slice of it runs per token.

What Tencent actually shipped

Weights, and a license that matters. On July 6 the Hunyuan team put Hy3 on Hugging Face (also ModelScope, GitCode and CNB) under Apache 2.0. That last part is the news as much as the size. A 295B model you can download, fine-tune and ship commercially with no gate is not something you get every week, and it lifts the geographic limits that hemmed in the April preview.

The shape of it: 295B total parameters, 21B active. It’s a mixture-of-experts, so each layer holds 192 routed experts plus one that’s always on, and a router picks the top 8 per token. 80 layers, a 256K context window, BF16 weights, and an FP8 checkpoint for teams who want a smaller footprint. There’s also a 3.8B multi-token-prediction layer bolted on, which lets the model guess more than one token per step so decoding runs faster. Recommended sampling, if you’re wiring it up, is temperature 0.9 and top_p 1.0.

Diagram: Hy3 holds 295B total parameters across 192 experts, but a router lights up only the top 8 experts per token, so 21B parameters actually run, giving 21B-model speed with 295B-model knowledge.
Why the two numbers both matter. It weighs 295B on disk and runs like a 21B model per token.

The benchmarks, and how far to trust them

Hy3’s own scorecard is genuinely strong on hard reasoning. GPQA Diamond at 90.4. USAMO 2026 at 72.0. IMOAnswerBench at 90.0. HLE, with tools, at 53.2. Those are the kind of numbers that used to belong to closed flagships, and they’re coming from open weights you can pull tonight.

The comparison people will care about is GLM-5.1, the other strong open Chinese model. Tencent ran a blind test: 270 experts, 312 real workflow comparisons, and Hy3 came out at 2.67 out of 4 against GLM-5.1’s 2.51, strongest in frontend and CI/CD work. It also claims a real reliability jump over its own April preview, with hallucinations down from 12.5% to 5.4% and commonsense errors from 25.4% to 12.7%.

Now the honest part. Every one of those figures is Tencent’s own, published alongside the model it’s selling. That’s normal, and it doesn’t make them wrong, but a vendor’s blind test is a starting point, not a verdict. We’d wait for independent runs (and honestly, run our own on the workloads we actually care about) before calling it a GLM-5.1 killer. If you’re weighing the open Chinese models against each other, our look at GLM-5.2 versus GPT-5.5 and Opus 4.8 lays out how these benchmark claims tend to hold up once you pressure them.

The catch: what it takes to run

Here’s where “open” stops meaning “yours”. Tencent’s own guidance is to serve Hy3 on 8 GPUs with large memory, naming the H20-3e class, and it ships the FP8 variant precisely because BF16 at this size is brutal on VRAM. It runs on vLLM (with the MTP layer doing speculative decoding) or SGLang (with EAGLE). Real inference stack, real data-center hardware. Not a homelab afternoon.

So for most of us, “using Hy3” this month means the free OpenRouter endpoint, tencent/hy3:free, which runs through July 21 before standard pricing kicks in. That’s the low-friction way to see whether its answers actually beat what you’re already paying for. If your interest is running an open model on your own box, Hy3 is the wrong size for that, and something like Qwen 3.7 offline is a far more honest fit for a single machine.

Checklist for Tencent Hy3: Apache 2.0 lets you self-host and ship commercially, it is free on OpenRouter until July 21, and only 21B params fire per token, but it needs about 8 large GPUs and the GLM-5.1 win is Tencent's own test.
The honest split: what the open license buys you, and where the hardware and the self-reported numbers pull back.

The honest read

Hy3 is a real gift to anyone building on open weights, and the Apache 2.0 license is the part that’ll still matter in a year. A 295B MoE with 21B active is a smart design, the reasoning scores are legitimately high, and the free window means you can judge it yourself before the pricing meter starts. What we wouldn’t do is treat the GLM-5.1 comparison as settled, or assume “open” means you’ll be running this on your own metal any time soon. Pull it up on OpenRouter this week, throw your own hard prompts at it, and see if it holds. That’s the only benchmark that pays your bills.

Sources: Tencent’s official Hy3 repository and model card on Hugging Face, with reporting and specs via MarkTechPost and Simon Willison, July 2026. Benchmark and blind-test figures are Tencent’s own and are not yet independently reproduced.

Frequently asked questions

What is Tencent Hy3?

Hy3 is Tencent Hunyuan's open large language model, released July 6, 2026. It is a 295B-parameter mixture-of-experts (MoE) model with 21B active parameters per token, a 256K context window, and Apache 2.0 licensed weights. Tencent aims it at reasoning and agent work, and ships both BF16 and FP8 checkpoints.

Is Hy3 really open source and free to use commercially?

The weights are released under Apache 2.0, which lets you download, self-host, fine-tune and ship the model in commercial products without paying a license fee. That is a genuinely permissive license, not a research-only or gated one. You still supply your own hardware to run it, or pay a host that serves it.

Can I run Hy3 locally?

Not on a normal desktop. Tencent recommends serving it on 8 GPUs with large memory, naming the H20-3e class, and ships an FP8 variant to cut the footprint. It runs on vLLM or SGLang. For a laptop or a single-GPU box, a much smaller open model is the realistic choice, not Hy3.

Is Hy3 better than GLM-5.1?

In Tencent's own blind evaluation, 270 experts scored Hy3 at 2.67 out of 4 against GLM-5.1 at 2.51 across 312 real workflow comparisons, with the biggest gains in frontend and CI/CD tasks. That is a self-reported test from the model's maker, so treat it as a promising signal and verify on your own workloads before trusting it.

How much does Hy3 cost?

The weights are free under Apache 2.0. To try it without hardware, OpenRouter runs a free endpoint at tencent/hy3:free through July 21, 2026, after which standard per-token pricing applies. Self-hosting has no license cost but does carry the GPU and power bill of serving a 295B model.

Tags: aillmnewsopen-sourceself-hostingtencent
Share196Tweet123
stephane

stephane

  • Trending
  • Comments
  • Latest
Answer card: Proton Lumo 2.0 is private by policy, not by locality. Saved history is locked so even Proton cannot read it, but the prompt is decrypted on a Proton EU server to answer it, then forgotten.

Proton Lumo 2.0 review: how private is it, really?

3 September 2026
Google's official announcement image for the release, reading Introducing Gemini 3.8 Flash and 3.8 Flash Cyber in black type over a pale blue background with a blurred white chevron and the four colour Gemini spark below.

Gemini 3.8 Flash keeps the price and the 1 January cliff

3 September 2026
Answer card stating that Anthropic announced Enterprise Frontier Safeguards on 1 September 2026, that activity data used for misuse monitoring moves into cloud storage the customer controls under the customer own encryption keys, that Anthropic charges nothing for the feature while the cloud provider bills storage and egress, and that the phased rollout starts later in autumn 2026 with interim zero data retention on Fable 5 and Fable 5.1 for eligible customers.

Anthropic moves retention into your own cloud, for 30 days

3 September 2026
Answer card: JWTs are not encrypted, anyone can read them; the signature proves who issued the token, not who may read it.

Are JWTs encrypted? No, and the difference will bite you

0
Answer card: a random 8 character password falls in under 2 hours offline, while 16 random characters hold for 1.4 trillion years at the same speed.

How long does it take to crack a password in 2026?

0
Answer card: three DNS records decide if your mail lands or bounces; SPF lists allowed senders, DKIM signs messages, DMARC sets the failure policy.

SPF, DKIM and DMARC explained: the records your email needs

0
Answer card stating that Mullvad announced on 3 September 2026 that it is shutting down its public encrypted domain name system servers on 2 November 2026 and sponsoring the Quad9 Foundation instead, with 194.242.2.2 and its five sibling addresses all going away, and virtual private network customers unaffected.

Mullvad’s DNS servers go dark on 2 November, and Quad9 blocks no ads

5 September 2026
OpenAI announcement image for GPT-6 Astra, a spiral galaxy of white, blue and amber points of light curling around a bright core on a near black star field.

GPT-6 Astra lists at $10 and $50, 2.5x what GPT-5.6 Sol costs

6 September 2026
Google's official announcement image for the release, reading Introducing Gemini 3.8 Flash and 3.8 Flash Cyber in black type over a pale blue background with a blurred white chevron and the four colour Gemini spark below.

Gemini 3.8 Flash keeps the price and the 1 January cliff

3 September 2026
  • About
  • Contact
  • Privacy
  • Legal

Copyright © 2026 Stephane Cardon.

Welcome Back!

Login to your account below

Forgotten Password?

Retrieve your password

Please enter your username or email address to reset your password.

Log In
No Result
View All Result
  • Home
  • Articles
    • Security
    • Network
    • Dev
    • Sysadmin
    • SEO
    • Email & DNS
  • Tools
    • Network tools: free, fast, no signup
    • Security tools: free, fast, no signup
    • Developer tools: free, fast, no signup
    • Sysadmin tools: free, fast, no signup
    • SEO tools: free, fast, no signup
    • Email & DNS tools: free, fast, no signup
  • Download
  • About

Copyright © 2026 Stephane Cardon.