Earlier quoted context omitted.
Can you elaborate which models you are using? I‘m running an R1 distilled Qwen coder with 32B Q4, and while it’s giving useful answers, it‘s quite slow on my M1 Max. Slow enough that I keep reaching for cloud models.
Not on my machine currently, I use the 14b Q4 model I think, which delivers very good answers. I run a 4060 with 16gb memory and performance is quite good. I used the largest model that was recommended with this amount of VRAM, I think it was the 14b one. I do have some applications that process images, text and pdf files and I use smaller models for extracting embeddings. I think my system wouldn't be able to handle…
Open-R1: an open reproduction of DeepSeek-R1
241–246 of 246 posts
Re: Open-R1: an open reproduction of DeepSeek-R1
#242Earlier quoted context omitted.
I can tell you from personal experience that chatgpt is a game changer in universities and schools. Close to 100% of students use chatgpt to study. I know in our university pretty much everyone that attends exams uses chatgpt to study. Chatgpt is arguably more valuable then Wikipedia and Google for studies.
It will be interesting to see if this benefits the learning of students or just helps them pass more easily without retaining any knowledge...
Re: Open-R1: an open reproduction of DeepSeek-R1
#243Earlier quoted context omitted.
I can tell you from personal experience that chatgpt is a game changer in universities and schools. Close to 100% of students use chatgpt to study. I know in our university pretty much everyone that attends exams uses chatgpt to study. Chatgpt is arguably more valuable then Wikipedia and Google for studies.
There was a posting, some time ago, about someone complaining that their young, primary-school-age sister was using ChatGPT to an absurd degree. I'm not sure that's a bad thing. She'll probably be one of the Thought Leaders, of Generation AI. I think that ML will have a really big impact on almost everyone, in every developed (and maybe developing, as well) nation. We need to keep in mind that ML is still very much i…
Re: Open-R1: an open reproduction of DeepSeek-R1
#244Is this what the Web was like in the beginning? Something exciting and fascinating every week?
the web was fascinating every second. You could click on a link without having ANY idea what you would land on. The overall quality was very poor, but it was thrilling. A bit like indie cinema.
There are people around trying to make these sites a bit more findable by creating specialized search engines. I have put the ones I know of at https://brisray.com/web/altsearch.htm
Re: Open-R1: an open reproduction of DeepSeek-R1
#245Earlier quoted context omitted.
the web was fascinating every second. You could click on a link without having ANY idea what you would land on. The overall quality was very poor, but it was thrilling. A bit like indie cinema.
I miss webrings [0], and especially the 'random' link that would take you to a random site within a given webring. [0] it feels weird to have to link to this but there's probably somebody who's never heard of them: https://en.wikipedia.org/wiki/Webring
Re: Open-R1: an open reproduction of DeepSeek-R1
#246Earlier quoted context omitted.
> Chatgpt is arguably more valuable then Wikipedia and Google for studies. But ChatGPT is just a glorified Wikipedia/Google. For the consumers it's an incremental thing (although from the engineering perspective it may seem to be a breakthrough).
Go try to learn a college level mathematics concept from Wikipedia, then try to learn it from ChatGPT. The wiki article may as well be written in a foreign language