Live data from Hacker News

Generative AI's Act Two

sequoiacap.com

1–10 of 125 posts

Re: Generative AI's Act Two

#3
Is there room for generative AI in science? I am experimenting with this a lot at https://atomictessellator.com, As a computational chemist I found it difficult just to stay on top of all of the papers that are released, I thought it would be cool to have generative AI attempt to reproduce the experiments using simulation tech.

Here's a few cool insights I have uncovered while working on this:

- Developer tools are popular, but there's an upcoming market for AI tools - I.e. Tooling/API's that are meant to be used primarily by LLMs/AIs. Design the interfaces to be well composable at the right level of abstraction and AIs can design, run and monitor experiments really well.

- Tree of thought and Graph of thought are really, really, really important for this - I think this is because it compensates for the lack of looping mechanisms in LLMs and also adds the ability for recursive problem decomposition, abstraction laddering, composable malleability and so on.

Does it work? Yes, I have a working E2E pipeline now that has already made novel discoveries validated by lab work, I am focusing on scaling this out now to support a broader search space and give the LLMs more freedom to explore.

A shameless plug: I am working on this 4-5 hours a day as well as my day job. If anyone is aware of any grants or investors that I could connect with, I would love to work on this full time. I am neurodivergent so I think running a company is not really my thing, but there must be an alternate way, advice anyone?

Re: Generative AI's Act Two

#4
> Four decades of the internet (accelerated by COVID) has given us trillions of tokens’ worth of training data.

Yeah four decades of stolen intellectual property posted by people on the internet in good will for the world to see, only to have it stolen and monetised by these folks. Wondering, once the good content dies out how will you “train” your “ai”?

Re: Generative AI's Act Two

#7

> Four decades of the internet (accelerated by COVID) has given us trillions of tokens’ worth of training data. Yeah four decades of stolen intellectual property posted by people on the internet in good will for the world to see, only to have it stolen and monetised by these folks. Wondering, once the good content dies out how will you “train” your “ai”?

Are we not allowed to learn from other's work?

Re: Generative AI's Act Two

#8

> Four decades of the internet (accelerated by COVID) has given us trillions of tokens’ worth of training data. Yeah four decades of stolen intellectual property posted by people on the internet in good will for the world to see, only to have it stolen and monetised by these folks. Wondering, once the good content dies out how will you “train” your “ai”?

Are we not allowed to learn from other's work?

You can learn all you want, but ai is a software, not a “we”.

Re: Generative AI's Act Two

#9

> Four decades of the internet (accelerated by COVID) has given us trillions of tokens’ worth of training data. Yeah four decades of stolen intellectual property posted by people on the internet in good will for the world to see, only to have it stolen and monetised by these folks. Wondering, once the good content dies out how will you “train” your “ai”?

[deleted]

Re: Generative AI's Act Two

#10

> Four decades of the internet (accelerated by COVID) has given us trillions of tokens’ worth of training data. Yeah four decades of stolen intellectual property posted by people on the internet in good will for the world to see, only to have it stolen and monetised by these folks. Wondering, once the good content dies out how will you “train” your “ai”?

Modern models are increasingly trained on synthetic data.
Post reply on HN