Live data from Hacker News

Operator research preview

openai.com

411–420 of 448 posts

Re: Operator research preview

#411

From the slide deck on the livestream: "[Operator safety risks and mitigations] Harmful tasks: User is misaligned" Looking forward to seeing some more of the examples for when openai considers their users as "misaligned", whatever that actually even means anymore.

> Looking forward to seeing some more of the examples for when openai considers their users as "misaligned"

All humans with politics not aligned with "The median sentiment of the San Francisco Board of Supervisors"

Re: Operator research preview

#412

Earlier quoted context omitted.

You can also write a python script to achieve the same goals. Except it's not python's responsibility to interpret the intent of your script, just as it's not your phone's responsibility to interpret the contents of your conversation. So our tools are not our morality police. We have a legal system that can operate within the bounds of law and due process. I am well aware of the already applied levels of machine lear…

> You can also write a python script to achieve the same goals. This argument is slightly disingenuous as it requires much higher skill, so it acts as a protective barrier. A person who have never programmed in their life could easily instruct this service to do all sorts of damaging things. (I suppose your counterargument will be that people can ask an LLM to write Python code to do the same. Still, most would strug…

Kids.

Kids are the non-aligned users who will use this to wreak havoc. And they will think it's hillarious.

Re: Operator research preview

#413

Earlier quoted context omitted.

We already have Ignition Interlock Devices which tell you how aligned you are and whether or not you need to improve your alignment score before starting the car. The EU is also making good progress on financial transactions – they're set to ban cash transactions over $10,000 by 2027.

> We already have Ignition Interlock Devices which tell you how aligned you are and whether or not you need to improve your alignment score before starting the car. I didn't know about "Ignition Interlock Devices". Wiki tells me: https://en.wikipedia.org/wiki/Ignition_interlock_device > breath alcohol ignition interlock device (IID or BAIID) is a breathalyzer for an individual's vehicle. It requires the driver to blo…

For buying shit. That you don't want some EU person in an office to know you bought. Which should be a fundamental human right.

Re: Operator research preview

#414

Earlier quoted context omitted.

As the storyline unfolds "AI" seems to be code for "machine learning based censorship". Soon we will have home appliances and vehicles telling you about how aligned you are, and whether you need to improve your alignment score before you can open your fridge. It is only a matter of time before this will apply to your financial transactions as well.

Did you not eat enough already? Come to think of it, do you not think you had enough internet for today Darious? You need to rest so that you can give 110% at . Proper food alignment is very important to a human.

Please align pie with pie-hole.

Re: Operator research preview

#415
post #383

Earlier quoted context omitted.

It's pretty troubling and illiberal to use the same word for a software tool being constrained by its manufacturer's moral framework and for a human user being constrained to that manufacturer's moral framework. While you can see how the word is formally valid and analogous in both cases, the connotation is that the user is being judged by the moral standards of a commercial vendor, which is about as Cyberpunk Dystop…

Being restricted from doing crimes by a vendor of commercial software isn’t a cyberpunk dystopia. Buy or download something else. It’s a typical restriction of software terms of service to prohibit use outside of applicable laws and regulations.

The trouble comes in to effect at the instant that the allowable-behavior boundary shrinks to be smaller than what is the law, to converge toward some other manifold defined by the ideologies of the employees or owners of the commercial software vendor.

Re: Operator research preview

#416
post #383

Earlier quoted context omitted.

Being restricted from doing crimes by a vendor of commercial software isn’t a cyberpunk dystopia. Buy or download something else. It’s a typical restriction of software terms of service to prohibit use outside of applicable laws and regulations.

If "alignment" were just about crimes we wouldn't need a special word for it, we would just say "legal". Alignment is not just about crimes, it's about the AI behaving in a way that is very specifically tailored to the moral framework and practical needs of the creator. Alignment has always gone well beyond the minimum required by law. And I don't think anyone is saying that a piece of software refusing to behave in…

I agree. But I had a devils-advocate moment.

Hacking in Counter Strike to have perfect aim and see your opponents through walls is legal. You aren't violating some anti-computer-misuse statute like DMCA. But Valve has every right to call the users of those scripts assholes who ruin the game for everyone and to ban them.

Re: Operator research preview

#417
post #301

Earlier quoted context omitted.

https://hn.algolia.com/?dateRange=all&page=0&prefix=false&qu... https://hn.algolia.com/?dateRange=all&page=0&prefix=true&que... Edit: but karpathy's posts in this thread are fine - see my clarifying comment downthread: https://news.ycombinator.com/item?id=42816589

That comment by karpathy was not made in bad faith. It's clearly an experiment done right here on HN so we can judge the tool ourselves.

Oh I agree! I thought about saying that but decided it would be confusing, but I guess it was also confusing not to.

What karpathy was doing is obviously in the spirit of the site [1], and HN has always been a spirit-of-the-law place, not a letter-of-the-law place [2].

[1] https://news.ycombinator.com/newsguidelines.html

[2] https://hn.algolia.com/?dateRange=all&page=0&prefix=false&qu....

Re: Operator research preview

#418
post #301

Earlier quoted context omitted.

https://hn.algolia.com/?dateRange=all&page=0&prefix=false&qu... https://hn.algolia.com/?dateRange=all&page=0&prefix=true&que... Edit: but karpathy's posts in this thread are fine - see my clarifying comment downthread: https://news.ycombinator.com/item?id=42816589

Would you consider updating the guidelines to this effect?

Yes, but I sort of feel like it's a corollary of the existing ones.

Re: Operator research preview

#419

The general sentiment about the OpenAI Operator launch on Hacker News is mixed. Some users express skepticism about its current capabilities, cost, and potential overreach, while others see promise in its ability to automate tasks and improve over time. Ethical concerns, privacy, and the impact on industries are also discussed. Overall, there's cautious optimism with acknowledgment of challenges and potential improve…

In the last weeks I experimented with Claude Computer Use to automate some daily tasks (via its Ui.Vision chat integration, see https://forum.ui.vision/t/v9-5-0-brings-computer-use-ai-chat... ) - and the results are mixed. Claude gets things wrong way too often to be useful.

Has anyone done any comparison Claude Computer Use vs OpenAI Operator? Is it signifcantly better?

Re: Operator research preview

#420

Earlier quoted context omitted.

> it's incredibly unsanitary I never thought about this. Does McD's PR team have anything to say about it? I assume that a bunch of people have challenged them about it on Twitter or TikTok. Would you feel better if there was a kind of automatic/robotic window washer that sanitised the screen after each use? The key to me about the kiosks is: (1) initially, replace cashier labour costs with new expensive machines, an…

You opened the door to walk into the McDonalds and never thought twice about it.

At least the outside of the door is bathed in ultraviolet from outer space.
Post reply on HN