Live data from Hacker News

Research acceleration: The view inside OpenAI

openai.com

71–80 of 210 posts

Re: Research acceleration: The view inside OpenAI

#71

This roughly lines up with my personal experience that in March a combination of stronger models and better tooling on my end let me start running jobs unattended 24/7 (using Anthropic sub and my own hardware). Their $8000/day per researcher spend is crazy though, I'm curious how they keep track of the work.

Can you elaborate on this? Especially the tooling.

I tried something similar and I remember it was still pretty dodgy in February.

Re: Research acceleration: The view inside OpenAI

#72

The burning question I can't get any information nn is whether, if they determined an earlier misaligned generation may have transmitted misalignment to the current models, they would roll back to a safe checkpoint to rebuild from there. I suspect they would not unless forced to.

No. They would just install a more convincing superego.

Re: Research acceleration: The view inside OpenAI

#73

> ... We are pursuing this work in part because automated research could help us solve alignment and build defenses against increasingly capable AI. An automated AI researcher can also be an automated safety or alignment researcher. More capable, aligned systems could help secure critical infrastructure, defend against dangerous AI agents, and develop new protective measures. In other words... "We must pursue advance…

I will believe AI is super strong when they start pulling out 10-d chess moves.

I’m yet to see it.

Re: Research acceleration: The view inside OpenAI

#74

This roughly lines up with my personal experience that in March a combination of stronger models and better tooling on my end let me start running jobs unattended 24/7 (using Anthropic sub and my own hardware). Their $8000/day per researcher spend is crazy though, I'm curious how they keep track of the work.

> let me start running jobs unattended 24/7 (using Anthropic sub and my own hardware)

How are you running jobs unattended 24/7 without hitting your token limits?

Re: Research acceleration: The view inside OpenAI

#75
post #73

> ... We are pursuing this work in part because automated research could help us solve alignment and build defenses against increasingly capable AI. An automated AI researcher can also be an automated safety or alignment researcher. More capable, aligned systems could help secure critical infrastructure, defend against dangerous AI agents, and develop new protective measures. In other words... "We must pursue advance…

I will believe AI is super strong when they start pulling out 10-d chess moves. I’m yet to see it.

If AI becomes really strong and sets itself the target of world domination, you maybe won't see those moves. You will just die in your sleep one day, or find no machine is under your control anymore.

I believe we are quite far from it, but that it makes sense to keep an eye out now. And think of resilient systems, manual overrides, etc. ...

Re: Research acceleration: The view inside OpenAI

#76
post #64

This roughly lines up with my personal experience that in March a combination of stronger models and better tooling on my end let me start running jobs unattended 24/7 (using Anthropic sub and my own hardware). Their $8000/day per researcher spend is crazy though, I'm curious how they keep track of the work.

These researchers are paid millions of dollars for their work. I doubt trust is really an issue at that level.

Imagine if one of the humans at OpenAI was misaligned! We should get the AI to research this possibility once they've been aligned.

Re: Research acceleration: The view inside OpenAI

#78

Earlier quoted context omitted.

RSI is a fetishistic term among the singularity crowd, who imagine AI "recursively" improving itself in some exponential fashion until there is a bright flash of white light and it reveals itself in the form of god. Or something like that. I don't know why whoever coined the term chose "recursive" rather than "iterative" - just sounds more likely to lead to infinite regress I suppose. This notion of recursive/iterati…

Yes the exponential self improvement folks have never heard of an eigenvalue I guess. You can loop forever using output as input but at some point the result will stop changing (depending on the function)

that's not really how eigenvalues work... they specifically also model the case where the result keeps changing exponentially.

Re: Research acceleration: The view inside OpenAI

#79

Earlier quoted context omitted.

Sounds like OpenAI are in the token-maxxing camp, so who knows what individual employees are doing to work their way up the leaderboard? If you spend $8000 to generate an animated pelican riding a bike, then how much tracking does it really need? Is the guy who spent $300,000 or so translating the FLT proof to Lean going to get a big Christmas bonus?

End of day, output and results are top target of measurements, token consumption is the obvious number that they would like to disclose for their own business benefits and a simple metrics that correlate with the output. Rest assured, capitalist appears irrational in wasting money, but they certainly care more about profit.

Taking a profit means you have to show numbers and the sooner you show numbers the harder it is to take people’s money.

Re: Research acceleration: The view inside OpenAI

#80
post #41

Earlier quoted context omitted.

RSI is a fetishistic term among the singularity crowd, who imagine AI "recursively" improving itself in some exponential fashion until there is a bright flash of white light and it reveals itself in the form of god. Or something like that. I don't know why whoever coined the term chose "recursive" rather than "iterative" - just sounds more likely to lead to infinite regress I suppose. This notion of recursive/iterati…

“Recursive” is a reasonable term because the generation N AIs will train the Generation N+1 AIs. The term “iterative” doesn’t reflect this nuance as well IMO.

[deleted]
Post reply on HN