• Mwa@thelemmy.club
    link
    fedilink
    English
    arrow-up
    4
    ·
    1 day ago

    We even got open weight models that’s 27B + 1-bit (and it still has good performance)

    • brucethemoose@lemmy.world
      link
      fedilink
      English
      arrow-up
      3
      ·
      1 day ago

      Bonsai? Or whatever it’s called? It’s a con, so far; it’s not better than smaller models quantized to 3-4 bits.

      I love, love the idea of bitnet, but it only seems to work with models trained from scratch, which no one has done at scale yet.