Geo News
Community curated by people like you
LatestAICryptoHealthWorld AffairsUS Politics

Mixture of experts (MoE) stories

Aug 28, 2026

Tencent releases Hy4 preview, open-source AI model with self-improvement capabilities

Tencent has unveiled Hy4 Preview, a 770-billion-parameter open-source Mixture-of-Experts model that uses a recursive self-improvement loop to boost inference throughput by 31.8%. The release, trained on Tencent's vast product ecosystem, positions the company back in the top tier of China's AI competition.

Aug 28, 2026·3 sources
00
Aug 26, 2026

Zhipu AI reveals GLM-5.3-Flash model ran on 100,000 Chinese chips

China's Zhipu AI confirmed that its viral Ox Alpha model, now revealed as GLM-5.3-Flash, is a 320B-parameter mixture-of-experts AI system that ran entirely on 100,000 domestically produced Chinese chips during launch. The geopolitically significant claim about Chinese chip capabilities remains unverified by independent sources.

Aug 26, 2026·4 sources
00
Jul 30, 2026

Thinking Machines Releases Inkling-Small, a 276B Parameter Mixture-of-Experts Model

Thinking Machines has released Inkling-Small, an open-weights Mixture-of-Experts transformer model with 276 billion parameters that achieves performance comparable to its larger predecessor Inkling while being one-quarter the size.

Jul 30, 2026·1 source
00
Jul 15, 2026

Thinking Machines Lab Releases Inkling, 975B-Parameter Open-Weights Multimodal AI Model

AI startup Thinking Machines Lab, led by former OpenAI CTO Mira Murati, has released Inkling, a 975-billion parameter mixture-of-experts model with open weights designed for enterprise customization and on-premises deployment. The company explicitly positions the model as focused on customizability rather than benchmark dominance.

Jul 15, 2026·7 sources
00
Jun 22, 2026

Tokyo's Sakana AI launches Fugu Ultra, multi-model orchestration system matching frontier AI performance

Sakana AI unveiled Fugu and Fugu Ultra on June 22, 2026, a multi-agent orchestration system that coordinates multiple AI models to match performance of leading models like Anthropic's Fable 5 and Mythos without training its own frontier model. The approach aims to reduce single-vendor dependency in enterprise AI deployments.

Jun 22, 2026·8 sources
00
May 7, 2026

Zyphra Releases ZAYA1-8B Reasoning Model Trained on AMD Hardware

Zyphra launched ZAYA1-8B, a Mixture of Experts reasoning model with only 760M active parameters that reportedly outperforms much larger open-weight models on math and coding benchmarks, trained entirely on AMD hardware.

May 7, 2026·1 source
00

Top claims

  • ▪A Cloud Security Alliance threat analysis warns that when AI participates in building its successor, the training pipeline becomes a high-value target where compromises could produce invisible behaviors.
  • ▪Goldman Sachs analysts stated that Tencent's product-plus-model strategy enabled it to collect user data from its product suite to feed back into subsequent training rounds.
  • ▪Tencent's previous generation model, Hy3, ranked 34th on the Code Arena WebDev benchmark as of September 1, 2026, highlighting the coding gains made by the Hy4 preview.

People involved

Mira Murati

Subtopics

Open-source AI4AI foundation models3AI startups3Business & enterprise AI2Compute, chips & AI infrastructure2Large language models (LLMs)2AGI arms race & geopolitics1AI1AI agents1AI inference (scaling)1AI research & benchmarks1China AI regulations1China tech & industrial strategy1Multimodal models1Open weights vs closed models1Reasoning models1Recursive self-improvement1

Related timelines

AI Data Center Gold Rush

101 stories

Congress

108 stories

Crypto hacks

100 stories

Ebola outbreak

58 stories

Iran War

209 stories

Featured Stories

Tokyo's Sakana AI launches Fugu Ultra, multi-model orchestration system matching frontier AI performance

Tokyo's Sakana AI launches Fugu Ultra, multi-model orchestration system matching frontier AI performance

Jun 22, 2026·8 sources