Outcome-Based Reinforcement Learning to Predict the Future
1–10 of 17 posts
Re: Outcome-Based Reinforcement Learning to Predict the Future
#2Eliminate all agents, all sources of change, all complexity - anything that could introduce unpredictability, and it suddenly becomes far easier to predict the future, no?
Re: Outcome-Based Reinforcement Learning to Predict the Future
#3Do you want paperclips? Because this is how you get paperclips! Eliminate all agents, all sources of change, all complexity - anything that could introduce unpredictability, and it suddenly becomes far easier to predict the future, no?
Don't^W worry, there are many other ways of getting paperclips, and we're doing all of them.
Re: Outcome-Based Reinforcement Learning to Predict the Future
#4Do you want paperclips? Because this is how you get paperclips! Eliminate all agents, all sources of change, all complexity - anything that could introduce unpredictability, and it suddenly becomes far easier to predict the future, no?
Re: Outcome-Based Reinforcement Learning to Predict the Future
#5Do you want paperclips? Because this is how you get paperclips! Eliminate all agents, all sources of change, all complexity - anything that could introduce unpredictability, and it suddenly becomes far easier to predict the future, no?
I don't know. Paperclips are awful useful. Would it be so bad to build more of them?
Re: Outcome-Based Reinforcement Learning to Predict the Future
#6Do you want paperclips? Because this is how you get paperclips! Eliminate all agents, all sources of change, all complexity - anything that could introduce unpredictability, and it suddenly becomes far easier to predict the future, no?
> Do you want paperclips? Because this is how you get paperclips! Don't^W worry, there are many other ways of getting paperclips, and we're doing all of them.
Re: Outcome-Based Reinforcement Learning to Predict the Future
#7Re: Outcome-Based Reinforcement Learning to Predict the Future
#8So instead of next token prediction its next event prediction. At some point this just loops around and we're back to teaching models to predict the next token in the sequence.
Re: Outcome-Based Reinforcement Learning to Predict the Future
#9So instead of next token prediction its next event prediction. At some point this just loops around and we're back to teaching models to predict the next token in the sequence.
Re: Outcome-Based Reinforcement Learning to Predict the Future
#10Do you want paperclips? Because this is how you get paperclips! Eliminate all agents, all sources of change, all complexity - anything that could introduce unpredictability, and it suddenly becomes far easier to predict the future, no?
I don't know. Paperclips are awful useful. Would it be so bad to build more of them?