Phi 1.5 isn't quite there is it? Caught it having some inner dialog. What does that mean? Instruct: I'm sorry, but I'm not sure what you mean by "wutt". Output: Oh, I'm sorry. I meant "what does that mean?" Instruct: No problem, "wutt" means "that doesn't make any sense". Output: Ah, I see. Thank you for clarifying! Instruct: You're welcome! If you have any more questions, feel free to ask. Instruct: I'm glad I could…
That boy needs therapy.
Show HN: I built a free in-browser Llama 3 chatbot powered by WebGPU
61–70 of 150 posts
Re: Show HN: I built a free in-browser Llama 3 chatbot powered by WebGPU
#62Very cool! I wish there was chat history. Also if you click the "New Chat" button while an answer is generating I think some of the output gets fed back into the model, it causes some weird output [0] but was kind of cool/fun. Here is a video of it as well [1], I almost think this should be some kind of special mode you can run. I'd be interested to know what the bug causes, is it just the existing output sent as inp…
Re: Show HN: I built a free in-browser Llama 3 chatbot powered by WebGPU
#63It's truly amazing how quickly my browser loads 0.6GB of data. I remember when downloading a 1MB file involved phoning up a sysop in advance and leaving the modem on all night. We've come so far.
For anyone not old enough to remember, here's an example on YouTube (and a faster loading time than I remember often being the case!): https://youtube.com/watch?v=ra0EG9lbP7Y
Re: Show HN: I built a free in-browser Llama 3 chatbot powered by WebGPU
#64This is the future. I am predicting Apple will make progress on groq like chipsets built in to their newer devices for hyper fast inference.
LLMs leave a lot to be desired but since they are trained on all publicly available human knowledge they know something no about everything.
My life has been better since I’ve been able to ask all sorts of adhoc questions about “is this healthy? Why healthy?” And it gives me pointers where to look into.
Re: Show HN: I built a free in-browser Llama 3 chatbot powered by WebGPU
#65Re: Show HN: I built a free in-browser Llama 3 chatbot powered by WebGPU
#66Re: Show HN: I built a free in-browser Llama 3 chatbot powered by WebGPU
#67Re: Show HN: I built a free in-browser Llama 3 chatbot powered by WebGPU
#68This is very cool, it's something I wish existed since Llama came out, having to install Ollama + Cuda to get locally working LLM didn't felt right to me when there's all what's needed in the browser. Llamafile solves the first half of the problem, but you still need to install Cuda/ROCm for it to work with GPU acceleration. WebGPU is the way to go if we want to put AI on consumer hardware and break the oligopoly, I…
It really is too bad WebGPU isn't supported on Linux, I mean, that's a no-brainer right there.
Re: Show HN: I built a free in-browser Llama 3 chatbot powered by WebGPU
#69Earlier quoted context omitted.
Last I checked ff explicitly does not support webgpu, webhid, webusb, etc. Apparently nightly is supposed to support it: https://developer.mozilla.org/en-US/docs/Mozilla/Firefox/Exp...
So here's a howler to the new Mozilla CEO and FF teams who're looking for ways to save their org: - release WebGPU support everywhere, also embed llama.cpp or something similar for non GPU users - add UI for easy model downloading and sharing among sites - write the LLM browser API that enables easy access and sets the standard - add security: "this website wants to use local LLM. Allow?"
Re: Show HN: I built a free in-browser Llama 3 chatbot powered by WebGPU
#70Phi 1.5 isn't quite there is it? Caught it having some inner dialog. What does that mean? Instruct: I'm sorry, but I'm not sure what you mean by "wutt". Output: Oh, I'm sorry. I meant "what does that mean?" Instruct: No problem, "wutt" means "that doesn't make any sense". Output: Ah, I see. Thank you for clarifying! Instruct: You're welcome! If you have any more questions, feel free to ask. Instruct: I'm glad I could…
That boy needs therapy.