GoPenAI
Read post

Mistral 8x7B 32k model stats

The Mistral 8x7B 32k model is a Mixture of Experts (MoE) model with 995 tensors, including token embedding, output norm, and output tensors. The model has 32 blocks of attention and ffn. During inference, two experts are used per token, resulting in a faster speed as if using a 12B model. The model has 47B parameters because the FFN layers are treated as individual experts. The model can be run on CPU if there is not enough VRAM on the GPU.

    #llama#llama-cpp#llm#local-ai#machine-learning#mistral-ai#mixture-of-experts
Dec 14, 2023•3m read time•From blog.gopenai.com
Post cover image
GoPenAI's image
GoPenAI

GOOpenAI is a blog or publication that focuses on exploring and discussing advancements, research, a...

693 Followers

•

4K Upvotes

Would you recommend this post?

Copy link
WhatsApp
Facebook
X
New Squad
  • © 2026 Daily Dev Ltd.
  • Guidelines
  • Explore
  • Tags
  • Sources
  • Squads
  • Leaderboard