• FaceDeer@fedia.io
    link
    fedilink
    arrow-up
    1
    ·
    6 days ago

    Sadly they don’t seem to have 36B-A3B models on their roadmap any more, they didn’t do one for 3.8 either. I agree it was a nice sweet spot between speed and capability, I still use the 3.6 version of 36B-A3B for larger-scale local work. Maybe someone else will aim for that. Or Qwen4 will do something new with the architecture that makes it unnecessary.

    • percent@infosec.pub
      link
      fedilink
      English
      arrow-up
      1
      ·
      6 days ago

      My Hermes instance periodically checks the overall open-weight LLM landscape, and it recently recommended trying Ornith 1.5 35B-A3B.

      I haven’t tried it yet, but on paper, it sounds like it has some potential.

      • FaceDeer@fedia.io
        link
        fedilink
        arrow-up
        1
        ·
        6 days ago

        Heh, I downloaded that one just recently, I read that it was good at natural prose and I’ve been working on a little pet project to make a framework for auto-writing short stories based on a simple premise. Haven’t tested it extensively yet though.