WizardLM/WizardCoder-33B-V1.1 released!

noneabove1182@sh.itjust.works · 10 months ago

WizardLM/WizardCoder-33B-V1.1 released!

noneabove1182@sh.itjust.works · edit-2 9 months ago

Btw I know this is old and you may have already figured out your hardware and setup, but p40s and p100s go for super cheap on eBay.

P40 is an amazing $/GB deal, only issue is the fp16 performance is abysmal so you’ll want to run either full fp32 models or use llama.cpp which is able to cast up to that size

The p100 has less VRAM but really good fp16 performance which makes it ideal for exllamav2 usage. I picked up one of each recently, p40 was failed to deliver and p100 was delivered while I’m away, but once I have both on hand I’ll probably post a comparison to my 3090 for interests sake

Also I run all my stuff on Linux (Ubuntu 22.04) with no issues

Alex@lemmy.ml · 9 months ago

I’ve generally tried to avoid Nvidia cards because binary blob drivers are a pain (especially as a FLOSS developer I occasionally need to build newer kernels). I believe the recent firmware changes mean the nouveau driver can now control clocking but I’ve no idea what the status is for CUDA which I assume you need to run the models.

They do look pretty affordable though 😀

noneabove1182@sh.itjust.works · 9 months ago

If you go for it and need any help lemme know I’ve had good results with Linux and Nvidia lately :)

WizardLM/WizardCoder-33B-V1.1 released!

WizardLM/WizardCoder-33B-V1.1 released!

WizardLMTeam/WizardCoder-33B-V1.1 · Hugging Face