Simon Willison

@simonw

What's the best open weight Mixture-of-Experts LLM for coding that fits in less than 60GB of RAM? I think MoE might be necessary to get reasonably interactive speeds on the hardware I have access to - I want something faster than 12 tokens/second
打开原帖#511482
  1. Industry

    Michael Truell: @zeeg Fix coming soon!
  2. Industry

    Elon Musk: @MichaelDell Compound growth is the most powerful force in the Universe
  3. Industry

    Elon Musk: Grok @Bot can make a sim of anything https://t.co/EKeFBpDTrW