Viewing profile — radq
radq
HN member- Joined
- Sun, Oct 17, 2010, 3:28 PM UTC
- HN karma
- 602
- Public activity
- 147 items
- HN profile
- View on Hacker News ↗
About radq
Recent public activity
-
comment
Comment #48730103
I'm disappointed with the commentary here. "GPU bubble" is an industry standard term, and literally how I would describe this to my colleagues in the industry. Look for example at …
-
comment
Comment #48730060
This is what people in the field call it. I'm sorry you're offended.
-
comment
Comment #48730039
Appreciate you saying the blog was nice. Not sure what you mean by "CODEX fingerprints", but I'll engage with the other points. We work on small models, and our customers want real…
-
comment
Comment #48729090
Thank you for the kind words. We will write and share more of these.
- comment
- story
-
comment
Comment #45392696
The 'point' skill is trained on a ton of UI data; we've heard of a lot of people using it in combination with a bigger driver model for UI automation. We are also planning on post-…
-
comment
Comment #45392661
Thanks! If you could shoot me a note at vik@m87.ai with any examples of the precision/recall issues you saw I'd appreciate it a ton.
-
comment
Comment #44197198
Cool project! The codebase is simple and well documented, a good starting point for anyone interested in how to implement a high-performance inference engine. The prefix sharing is…
-
comment
Comment #42332305
Hello folks, I work on moondream. Posted a demo video on twitter for this release: https://x.com/vikhyatk/status/1864727630093934818 Happy to answer any questions!
-
comment
Comment #42286413
Not true, H100s cost $2-3/GPU/hr on the open market.
-
comment
Comment #40665750
Have you considered sponsoring an open-source project? ;)
-
comment
Comment #39871436
1/3rd "activated parameters", while also requiring 2x the VRAM.
-
comment
Comment #39761525
The training technique used here (fitting something similar to a NeRF to different views of the same image) is pretty similar to this paper which uses a similar technique to denois…
-
story
Show HN: Moondream, a small vision language model that runs on 8GB of RAM
I've been working on training this small vision language model for the last month - excited to release the first prototype today! It is based on SigLIP (image encoder), Phi-1.5 (te…
-
comment
Comment #38518433
I'm confused - you posted in the "who wants to be hired" thread, and then got an email from this company asking if you'd be interested?
-
comment
Comment #36856182
Do outlier features emerge in sub-100M parameter models? I haven't seen any research discuss it below the 124M scale (bert-base). At that scale training a model takes ~4 days on an…
-
story
Show HN: AlignedBot, the safest and most aligned chatbot
I found the mixed reactions to Llama 2 interesting, with some people concerned about the safety implications of open-sourcing a powerful LLM, while others were finding the chat RLH…
- story
-
comment
Comment #36608233
The plugin is supposed to ask for confirmation, according to OpenAI's documentation at least. > When a user asks a relevant question, the model may choose to invoke an API call fro…
-
comment
Comment #36517195
Literally happened to me, though it was during COVID lockdown so it's possible things have gotten better.
-
comment
Comment #36517187
Looks like you're right, there was a recommendation to increase it but I don't see anything about it taking effect.
- comment
-
comment
Comment #36516474
I don't expect a lot of people to take this up. Express Entry was already an option for these folks and you get permanent residency from the start under that program.
-
comment
Comment #36516458
As someone that's used to live in Canada and is currently in the USA, this is not true. Six month wait time to get an x-ray for a hairline fracture, people dying because of long wa…