So, I was reading the privacy notice and the terms of use and I did read some sketchy stuff about it (data used in advertising, getting keystroke). How bad is it? Is it like chatgpt or worse? Anything I can do about it?
So, I was reading the privacy notice and the terms of use and I did read some sketchy stuff about it (data used in advertising, getting keystroke). How bad is it? Is it like chatgpt or worse? Anything I can do about it?
I have a 950m
It looks like that has 4GB of RAM, so depending on the rest of your system you might actually be able to run some quantised models! You’ll need to use software that supports “offloading” the operations to system RAM which don’t fit in your GPU’s VRAM, like LMStudio (https://lmstudio.ai/).
I recommend checking out Unsloth’s models, they specifically try to fine-tune for use on older/slower hardware like yours. Their HuggingFace is here: https://huggingface.co/unsloth
This is the version of Deepseek you’ll wanna try: https://huggingface.co/unsloth/DeepSeek-R1-Distill-Llama-8B-GGUF
Click one of the options on the right hand side to download the model file:
Very basically speaking, the lower the “bits”, the smaller the file (and the dumber the model), and therefore the less VRAM and system RAM you’ll need to run it. If you get one of the 2-bit versions, you might be able to fit the whole thing inside your GPU - the 2-bit models are only ~3.2GB! You can probably run 4-bit though, even on your hardware.
Wow, that’s a thorough explanation. Thanks! I also have 16 gigs of ram and an i7 6th gen
No problem - and, that’s not thorough, that’s the cut down version haha!
Yeah, that hardware’s a little old so the token generation might be slow-ish (your RAM speed will make a big difference, so make sure you have the fastest RAM the system will support), but you should be able to run smaller models without issue 😊 Glad to help, I hope you manage to get something up and running!