LocalLLaMA

2237 readers

1 users here now

Community to discuss about LLaMA, the large language model created by Meta AI.

This is intended to be a replacement for r/LocalLLaMA on Reddit.

founded 1 year ago

MODERATORS

SkySyrup@sh.itjust.works

pax@sh.itjust.works

noneabove1182@sh.itjust.works

WizardLM/WizardCoder-33B-V1.1 released! (huggingface.co)

submitted 10 months ago by noneabove1182@sh.itjust.works to c/localllama@sh.itjust.works

11 comments fedilink hide all child comments

Based off of deepseek coder, the current SOTA 33B model, allegedly has gpt 3.5 levels of performance, will be excited to test once I've made exllamav2 quants and will try to update with my findings as a copilot model

you are viewing a single comment's thread
view the rest of the comments

[–] stsquad@lemmy.ml 2 points 9 months ago (1 children)

I've generally tried to avoid Nvidia cards because binary blob drivers are a pain (especially as a FLOSS developer I occasionally need to build newer kernels). I believe the recent firmware changes mean the nouveau driver can now control clocking but I've no idea what the status is for CUDA which I assume you need to run the models.

They do look pretty affordable though 😀

[–] noneabove1182@sh.itjust.works 1 points 9 months ago

If you go for it and need any help lemme know I've had good results with Linux and Nvidia lately :)