• Pennomi@lemmy.world
    link
    fedilink
    English
    arrow-up
    11
    ·
    1 year ago

    Pretty easy to roll your own with Kobold.cpp and various open model weights found on HuggingFace.

    • TipRing@lemmy.world
      link
      fedilink
      English
      arrow-up
      8
      ·
      1 year ago

      Also for an interface, I’d recommend KoboldLite for writing or assistant and SillyTavern for chat/RP.

      • exu@feditown.com
        link
        fedilink
        English
        arrow-up
        4
        ·
        1 year ago

        You’ll want to use a quantised model on your GPU. You could also use the CPU and offload some parts to the GPU with llama.cpp (an option in oobabooga). Llama.cpp models are in the GGUF format.