LocalLLaMA

2221 readers

1 users here now

Community to discuss about LLaMA, the large language model created by Meta AI.

This is intended to be a replacement for r/LocalLLaMA on Reddit.

founded 1 year ago

MODERATORS

pax@sh.itjust.works

SkySyrup@sh.itjust.works

noneabove1182@sh.itjust.works

Guide on setting up a local GGML model? (lemmy.world)

submitted 1 year ago* (last edited 1 year ago) by Magiwarriorx@lemmy.world to c/localllama@sh.itjust.works

11 comments fedilink hide all child comments

I've been messing around with GPTQ models with ExLlama in ooba, and have gotten 33b models @ 3k running smoothly, but was looking to try something bigger than my VRAM can hold.

However, I'm clearly doing something wrong, and the koboldcpp.exe documentation isn't clear to me. Does anyone have a good setup guide? My understanding is koboldcpp.exe is preferable for GGML, as ooba's llama.cpp doesn't support GGML at >4k context yet.

no comments (yet)

sorted by: hot top controversial new old

there doesn't seem to be anything here