• All tags
  • mixture-of-experts
  • llm
  • open-source
  • ai-inference
  • kimi-k3

Mixture of Experts

Tag·428 stories

Mixture of Experts news covering a model architecture that routes each input to a subset of specialised subnetworks instead of activating all parameters. Readers can learn about routing and gating mechanisms, sparse activation and parameter counts, training stability.

Researchers from Princeton and Meta AI Introduce ‘Lory’: A Fully-Differentiable MoE Model Designed for Autoregressive Language Model Pre-TrainingThis AI Paper by DeepSeek-AI Introduces DeepSeek-V2: Harnessing Mixture-of-Experts for Enhanced AI PerformanceMixture of Expert Architecture. Definitions and Applications included Google’s Gemini and Mixtral 8x7BBringing MegaBlocks to DatabricksHow do mixture-of-experts layers affect transformer models?Alibaba Releases Qwen1.5-MoE-A2.7B: A Small MoE Model with only 2.7B Activated Parameters yet Matching the Performance of State-of-the-Art 7B models like Mistral 7BCreate Mixtures of Experts with MergeKitUnderstanding the Sparse Mixture of Experts (SMoE) Layer in MixtralMixtral-8x7B: Overview and Benchmarks with Combining Mixtral and Flash Attention 2grok-1
Posts by WinPosts by Andrew MPosts by GeekLuffy

Recommended Mixture of Experts stories

Who to follow for Mixture of Experts

feigle's user avatar
Win
@feigle
Joined Jul 7. 2026
40

Head of Growth @ Vast.ai

andrewma's user avatar
Andrew M
@andrewma
Joined Jun 19. 2025
1.6K

full time overthinker

geekluffy's user avatar
GeekLuffy
@geekluffy
Joined Apr 3. 2026
160

Full-Stack & AI Engineer building scalable web apps & high-performance edge solutions.

Top sources covering Mixture of Experts

Most upvoted Mixture of Experts posts

Best discussed Mixture of Experts posts

All posts about Mixture of Experts