Dispatch
OpenAI Announces $200B Valuation Round   •   EU AI Act Compliance Deadline Extended to 2027   •   Google DeepMind Releases Gemini Ultra 3.0   •   Y Combinator S26 Batch: 60% of Startups Are AI-Native   •   MarTech Consolidation: Salesforce Acquires MadTech Pioneer   •   LLM Token Costs Drop 80% Year-Over-Year   •   Meta Llama 4 Released Under Permissive Commercial Licence   •   Anthropic's Claude Achieves New Benchmarks on Reasoning Tasks   •   Venture Capital Flows to AI Infrastructure Exceed $4B in Q2   •   Adobe GenStudio Reaches 500,000 Enterprise Users   •   OpenAI Announces $200B Valuation Round   •   EU AI Act Compliance Deadline Extended to 2027   •   Google DeepMind Releases Gemini Ultra 3.0   •   Y Combinator S26 Batch: 60% of Startups Are AI-Native   •   MarTech Consolidation: Salesforce Acquires MadTech Pioneer   •   LLM Token Costs Drop 80% Year-Over-Year   •   Meta Llama 4 Released Under Permissive Commercial Licence   •   Anthropic's Claude Achieves New Benchmarks on Reasoning Tasks   •   Venture Capital Flows to AI Infrastructure Exceed $4B in Q2   •   Adobe GenStudio Reaches 500,000 Enterprise Users
Est. MMXXV — Independent Digital PressWednesday, 2 September 2026Vol. I — No. 195
MarTech • Startups • LLMs • Digital Strategyterekhindigital.comMorning Edition

Terekhin Digital Media

Rigorous Journalism at the Frontier of Digital Commerce & Machine Intelligence

Wednesday, 2 September 2026Issue No. 195
LLMs

Tencent Releases Hy4-Preview: 770B MoE Frontier Model, Apache Licence, 1M Context

A 770B parameter open-weight model from Tencent — Apache 2.0 licensed, competitive on SWE Bench Pro at 65.7 — arrives as the open-source consolidation wave absorbs the infrastructure it was built on.

Tencent released Hy4-Preview on Hugging Face and ModelScope on Thursday — a 770 billion parameter Mixture-of-Experts model activating 49 billion parameters per token across 78 layers with 256 routed experts. The model's custom attention mechanism, Gated DeepSeek Sparse Attention with IndexCache, enables cross-layer sparse index reuse and represents a notable architectural departure from standard transformer attention. Benchmark scores — 92.3 on GPQA Diamond, 65.7 on SWE Bench Pro, 64.3 on Deep SWE — place it at the competitive open-source frontier tier, particularly for software engineering tasks. One million token context window, Apache 2.0 licence, available in standard and FP8-quantized versions. Tencent explicitly describes the release as a preview with "real headroom left in both pre-training and post-training." The timing — released into a week in which the primary open-weight hosting and model development infrastructure is being absorbed by Nvidia and Stripe via the Hugging Face and Poolside acquisitions — is worth noting. Hy4-Preview lands on a commons that may, within months, no longer be a commons.

TencentHy4open-sourceMoELLMopen-weightHugging Face
← Return to Front Page
Related Articles
© MMXXVI Terekhin Digital Media — All Rights Reserved — An Independent Digital Publication