Live data from Hacker News

Can LLMs invent better ways to train LLMs?

sakana.ai

11–20 of 39 posts

Re: Can LLMs invent better ways to train LLMs?

#11
I'm sure LLMs can optimize the training of other LLMs (either by inventing new ways or fine tuning existing ones). But we can't predict whether this will result in a giant's leap in the field, or just small increments. That's the definition of singularity, isn't it?

Re: Can LLMs invent better ways to train LLMs?

#15
post #5

A better question is "Can LLMs invent anything ?" Don't misunderstand, building systems models using existing system response as a way of analyzing those systems is a useful methodology and it makes some things otherwise tedious things not so tedious. Much like "high level" languages removed the tedium of writing in assembly code. But for the same reason that a compiler won't emit a new, more powerful, CPU instructio…

can you propose a concrete prediction here?

Re: Can LLMs invent better ways to train LLMs?

#16
post #5

A better question is "Can LLMs invent anything ?" Don't misunderstand, building systems models using existing system response as a way of analyzing those systems is a useful methodology and it makes some things otherwise tedious things not so tedious. Much like "high level" languages removed the tedium of writing in assembly code. But for the same reason that a compiler won't emit a new, more powerful, CPU instructio…

I think it's probably unreasonable to expect them to without giving them the ability to experiment and test ideas.

Pretty much no inventions were invented just by thinking, which is the environment most LLMs have.

Re: Can LLMs invent better ways to train LLMs?

#18
post #5

A better question is "Can LLMs invent anything ?" Don't misunderstand, building systems models using existing system response as a way of analyzing those systems is a useful methodology and it makes some things otherwise tedious things not so tedious. Much like "high level" languages removed the tedium of writing in assembly code. But for the same reason that a compiler won't emit a new, more powerful, CPU instructio…

I think it's probably unreasonable to expect them to without giving them the ability to experiment and test ideas. Pretty much no inventions were invented just by thinking, which is the environment most LLMs have.

I think this is the key point. There is no feedback loop right now.

Re: Can LLMs invent better ways to train LLMs?

#19
post #5

A better question is "Can LLMs invent anything ?" Don't misunderstand, building systems models using existing system response as a way of analyzing those systems is a useful methodology and it makes some things otherwise tedious things not so tedious. Much like "high level" languages removed the tedium of writing in assembly code. But for the same reason that a compiler won't emit a new, more powerful, CPU instructio…

[deleted]
Post reply on HN