Earlier quoted context omitted.
Is that (i.e. GPT) not still a service? What people want is something they can run on their own hardware without sending their queries to some third party service which is doing who knows what with them. This is already possible if you want to mess around with green code that isn't in system repositories yet and buy expensive hardware to make it fast, but you can imagine why some people don't have the time or money f…
I mean you don't need to use GPT, it's just if you wanted to build the product in OP (ChatGPT tuned for your site) you would. Question answering can be tackled by smaller models that run on CPUs: https://huggingface.co/tasks/question-answering And if it's strictly for personal use there's always the chat-tuned stuff being built on top of LLaMA like Alpaca > waiting for Intel or AMD to realize Intel and AMD just got t…
Hence the demand for something else.
> Intel and AMD just got their lunch eaten by Apple Silicon which did exactly that, so I'm sure they're working on it
Apple's GPU doesn't benchmark much different than competing iGPUs for gaming. It may be that the only thing stopping anyone from running this stuff on existing iGPUs is software support.