LocalLLaMA

2249 readers

1 users here now

Community to discuss about LLaMA, the large language model created by Meta AI.

This is intended to be a replacement for r/LocalLLaMA on Reddit.

founded 1 year ago

MODERATORS

SkySyrup@sh.itjust.works

pax@sh.itjust.works

noneabove1182@sh.itjust.works

WizardLM/WizardCoder-33B-V1.1 released! (huggingface.co)

submitted 10 months ago by noneabove1182@sh.itjust.works to c/localllama@sh.itjust.works

11 comments fedilink hide all child comments

Based off of deepseek coder, the current SOTA 33B model, allegedly has gpt 3.5 levels of performance, will be excited to test once I've made exllamav2 quants and will try to update with my findings as a copilot model

you are viewing a single comment's thread
view the rest of the comments

[–] noneabove1182@sh.itjust.works 2 points 10 months ago (1 children)

The 3060 is a nice cheap one for running okay sized models, but if you can find a way to stretch for a 3090 or a 7900 XTX you'll be able to run these 33B models with decent quant levels

[–] stsquad@lemmy.ml 3 points 10 months ago (1 children)

I was hoping to avoid Nvidia's binary drivers although I don't know what the driver/support status of dedicated AI accelerators are like on Linux._

[–] noneabove1182@sh.itjust.works 3 points 10 months ago

I run my Nvidia stuff in containers to not have to deal with all the stupid shenanigans