Hacker Newsnew | past | comments | ask | show | jobs | submitlogin

I am running Deepseek R1 on my AMD Ryzen 7 PRO 5850U integrated GPU. While my experience will R1 doesn't make me think well of it, it is impressive how fast it is on such a weak graphics processor.


Note: you are probably running a distilled version of R1, which is actually LLama or Qwen further trained on the input/output of R1.

The full R1 is huge (~700GB), altough there are still quantized versions, the smallest one is around 150gb (1.58bit)


Oh, that's interesting. I didn't know that the ollama version wasn't the whole thing.


ollama deepseek-r1:671b is


You're most likely running a destilled version. The full model is ~700GB.


The default model on ollama is the 7b distillation.

Its ability to solve basic math problems with reasoning is pretty cool, but other models of that size (qwen 2.5, phi4) have been generally more useful to me.

These tiny models still strike me as toys, not a whole bunch of real-world utility.


Yeah phi4 has been as good or better for me than r1-qwen 32b for general queries




Guidelines | FAQ | Lists | API | Security | Legal | Apply to YC | Contact

Search: