Viewing profile — agajews
agajews
HN member- Joined
- Wed, Jul 03, 2019, 5:08 PM UTC
- HN karma
- 105
- Public activity
- 23 items
- HN profile
- View on Hacker News ↗
About agajews
No profile information was provided.
Recent public activity
-
comment
Comment #49198521
Hey Hacker News! I'm Alex, one of the founders of Pantograph. A year ago, my cofounder Kelsey and I set out to try to make the most inexpensive robot that's capable enough to do in…
- story
-
comment
Comment #49186214
exciting! you could imagine brain transcriptions being very high signal RLHF data that could help align models more deeply to human preference.
-
comment
Comment #48927002
Pretty different--obviously a lot less consistent today, but capable of a lot more diverse kinds of behavior. As the models get bigger, they'll be easier to prompt, and you could i…
-
comment
Comment #48926971
Awesome! We should be a lot better than ordinary LLMs, especially at tasks that require making a lot of decisions in real-time.
-
comment
Comment #48926956
Probably not the current model, but one of the benefits of doing internet-scale pretraining is that the model has seen a lot of mods already! We think the bigger models will be abl…
-
comment
Comment #48924997
The goal conditioning is just a training objective! You could combine it with tool use, web searches, etc. and still train end-to-end on goal conditioning. (One way of thinking abo…
-
comment
Comment #48924970
you don't need to pick good images manually! it's similar to next-token prediction, many prediction tasks aren't especially interesting, but there are enough hard ones that the mod…
-
comment
Comment #48924583
Not yet, but we'll add language as a modality to the larger models! The models are trained end-to-end on video data, so we'll need datasets that mix video and language, e.g. transc…
-
comment
Comment #48924159
Very much inspired by those papers! One of the things that's interesting about our model is it's goal-conditioned, so it can do any task at inference time without training on it. W…
-
comment
Comment #48924108
Hey everyone! I'm Alex, one of the founders of Pantograph. We've spent the last six months building a pretty smart Minecraft model, coming soon to a server near you! We trained it …
- story
- story
-
comment
Comment #36936770
Ah the hardware isn’t gonna be in SF (not the cheapest datacenter space) But I do think a lot of our customers will be out here —- SF is still probably the best place to do startup…
-
comment
Comment #36936558
Yeah we aren’t going to let anyone book the whole thing for years. If we ever have to make the choice, we’ll choose the startups over the big companies.
-
comment
Comment #36936517
Ah that’s only if you pay for 3 years of compute upfront. Most startups, especially the small ones, really can’t afford that
-
comment
Comment #36935181
Yeah it’s pretty hard to find a big block of GPUs that you can use for a short time, esp if you need infiniband for multinode training. Lambda I think needs a min reservation of 6-…
-
comment
Comment #36934352
[dead]
-
comment
Comment #33552735
Yeah exactly. That's why you can't really do it with a language model like GPT-3, you have to bake into the architecture the concept of a "link" as a first-class object.
-
comment
Comment #33552224
Hey, thanks for posting! We actually have an architecture that lets us expand the index without doing any retraining, so we can add/update pages pretty much for free.
-
comment
Comment #33552199
Hey everyone! Metaphor team here. We launched Metaphor earlier this morning! It's a search engine based on the same sorts of generative modeling ideas behind Stable Diffusion, GPT-…
- job
- job