Live data from Hacker News

Operator research preview

openai.com

301–310 of 448 posts

Re: Operator research preview

#301
post #255

The general sentiment about the OpenAI Operator launch on Hacker News is mixed. Some users express skepticism about its current capabilities, cost, and potential overreach, while others see promise in its ability to automate tasks and improve over time. Ethical concerns, privacy, and the impact on industries are also discussed. Overall, there's cautious optimism with acknowledgment of challenges and potential improve…

@dang can we have guidelines against posting AI generated content here? (who cares if the account is "human operated" or has a disclaimer). It's just lame and not what this forum is about.

https://hn.algolia.com/?dateRange=all&page=0&prefix=false&qu...

https://hn.algolia.com/?dateRange=all&page=0&prefix=true&que...

Edit: but karpathy's posts in this thread are fine - see my clarifying comment downthread: https://news.ycombinator.com/item?id=42816589

Re: Operator research preview

#302
post #275
post #255

Earlier quoted context omitted.

@dang can we have guidelines against posting AI generated content here? (who cares if the account is "human operated" or has a disclaimer). It's just lame and not what this forum is about.

How are you even going to moderate such content, how will the website operator even know if its real human or an AI agent controlling a computer?

So far HN users seem to be doing a pretty job of flagging them.

Of course, the big question is what to do if/when they're smart enough to fool everybody.

Re: Operator research preview

#303

From the slide deck on the livestream: "[Operator safety risks and mitigations] Harmful tasks: User is misaligned" Looking forward to seeing some more of the examples for when openai considers their users as "misaligned", whatever that actually even means anymore.

As the storyline unfolds "AI" seems to be code for "machine learning based censorship". Soon we will have home appliances and vehicles telling you about how aligned you are, and whether you need to improve your alignment score before you can open your fridge. It is only a matter of time before this will apply to your financial transactions as well.

We already have Ignition Interlock Devices which tell you how aligned you are and whether or not you need to improve your alignment score before starting the car.

The EU is also making good progress on financial transactions – they're set to ban cash transactions over $10,000 by 2027.

Re: Operator research preview

#304
post #302
post #275

Earlier quoted context omitted.

How are you even going to moderate such content, how will the website operator even know if its real human or an AI agent controlling a computer?

So far HN users seem to be doing a pretty job of flagging them. Of course, the big question is what to do if/when they're smart enough to fool everybody.

Offer a Turing award to the bot-trainer?

Re: Operator research preview

#305

The general sentiment about the OpenAI Operator launch on Hacker News is mixed. Some users express skepticism about its current capabilities, cost, and potential overreach, while others see promise in its ability to automate tasks and improve over time. Ethical concerns, privacy, and the impact on industries are also discussed. Overall, there's cautious optimism with acknowledgment of challenges and potential improve…

[deleted]

Re: Operator research preview

#306

Earlier quoted context omitted.

As the storyline unfolds "AI" seems to be code for "machine learning based censorship". Soon we will have home appliances and vehicles telling you about how aligned you are, and whether you need to improve your alignment score before you can open your fridge. It is only a matter of time before this will apply to your financial transactions as well.

We already have Ignition Interlock Devices which tell you how aligned you are and whether or not you need to improve your alignment score before starting the car. The EU is also making good progress on financial transactions – they're set to ban cash transactions over $10,000 by 2027.

We're way ahead of that in Turkey: our cash limit is ₺10k, which amounts to about €267 (exception: transactions legally requiring notarisation by a notary public).

Re: Operator research preview

#307
post #249

Earlier quoted context omitted.

If you are worried about costs you can use Browser Use with deepseek which becomes super cheap! https://github.com/browser-use/browser-use

Exactly what I was look for. Thank you. I wish they had Gemini since the free tier is generous but I guess it's in the works. I'll take a look and see how bad it would be implement

Browser use supports gemini via langchain

Re: Operator research preview

#308

Earlier quoted context omitted.

As the storyline unfolds "AI" seems to be code for "machine learning based censorship". Soon we will have home appliances and vehicles telling you about how aligned you are, and whether you need to improve your alignment score before you can open your fridge. It is only a matter of time before this will apply to your financial transactions as well.

We already have Ignition Interlock Devices which tell you how aligned you are and whether or not you need to improve your alignment score before starting the car. The EU is also making good progress on financial transactions – they're set to ban cash transactions over $10,000 by 2027.

    > We already have Ignition Interlock Devices which tell you how aligned you are and whether or not you need to improve your alignment score before starting the car.
I didn't know about "Ignition Interlock Devices". Wiki tells me: https://en.wikipedia.org/wiki/Ignition_interlock_device

    > breath alcohol ignition interlock device (IID or BAIID) is a breathalyzer for an individual's vehicle. It requires the driver to blow into a mouthpiece on the device before starting or continuing to operate the vehicle.
Can you explain your concern about this device? Also: How do you feel about laws that require drivers and passengers to wear a seatbelt?

Also: Can you share a legitimate reason why you need to use cash in transactions over 10K EUR?

Re: Operator research preview

#309

What is fascinating about this announcement is if you look into future after considerable improvements in product and the model, we will be just chatting with ChatGPT to book dinner tables, flights, buy groceries and do all sort of mundane and hugely boring things we do on the web, just by talking to the agents. I'd definitely love that.

I don't. Chat interface sucks; for most of these things, a more direct interface could be much more ergonomic, and easier to operate and integrate. The only reason we don't have those interfaces is because neither restaurants, nor airlines, nor online stores, nor any other businesses actually want us to have them. To a business, the user interface isn't there to help the user achieve their goals - it's a platform for…

Why assume that chat will be the interface? Multimodal and dynamically generated seems more likely.

Re: Operator research preview

#310

Earlier quoted context omitted.

I can sympathize with vague notions of AI dystopia, but this might be stretching the concept a bit too far. This kind of service is extremely abusable ("Operator, go to Wikipedia and start mass-vandalizing articles" or "Go to this website and try these people's email addresses with random passwords until it locks their accounts") and building some alignment goals into it doesn't seem like a terribly draconian idea. A…

You can also write a python script to achieve the same goals. Except it's not python's responsibility to interpret the intent of your script, just as it's not your phone's responsibility to interpret the contents of your conversation. So our tools are not our morality police. We have a legal system that can operate within the bounds of law and due process. I am well aware of the already applied levels of machine lear…

    > You can also write a python script to achieve the same goals.
This argument is slightly disingenuous as it requires much higher skill, so it acts as a protective barrier. A person who have never programmed in their life could easily instruct this service to do all sorts of damaging things. (I suppose your counterargument will be that people can ask an LLM to write Python code to do the same. Still, most would struggle to run the Python code.)
Post reply on HN