Live data from Hacker News

Ask HN: Is anyone else bearish on OpenAI?

news.ycombinator.com

111–120 of 321 posts

Re: Ask HN: Is anyone else bearish on OpenAI?

#111

OpenAI, at least in my day-day workflow for the last 9+ months has so superseded anything that google ever was to me that I'm having a difficult time comparing the two. I've got a monitor dedicated 100% of the time to ChatGPT, and I interact with it non stop during the flow of technical scenarios and troubleshooting situations that flow into me - working in areas that I have the slimmest of backgrounds in, and shutti…

> OpenAI, at least in my day-day workflow for the last 9+ months has so superseded anything that google ever was to me that I'm having a difficult time comparing the two.

Seconded.

Let me tell you what I used to do.

First, Imagine I have an error executing code or an error running some bit of 3rd party software. I go to Stack Overflow and search. I find posts related to my problem and I spend a great deal of time trying to shoehorn existing answers to my specific issue. Sometimes, my shoehorning works and I fix my problem. Other times, it doesn't work, and then I post on Stack Overflow myself. And I wait ... and wait ... for a response. Sometimes I get a response.

Now, when I have this type of problem, I tell ChatGPT, "Hey, I'm trying to do and I'm getting this error . Help me troubleshoot this. And it almost always helps me fix my problem. And it's ~10× faster than Stack Overflow.

=============

Second, there are times where I have write code to do some relatively 'complex' data manipulation--nothing sophisticated, mind you, but stuff like, "I need these data columns rearranged based on complicated logic. And I need the text in columns A, X, AQ, and F merged, but only if . Otherwise, Just merge text in Columns A and AQ, except if the date in column ZZ is after January 1, 2019.". I can do this stuff on my own, but: a] it's cognitively draining, b] it takes time, c] I often make silly errors due to the complexity.

ChatGPT is, again, an order of magnitude faster than I am. And it makes _fewer_ errors.

It still makes errors. And I still have to know what to look for to catch those errors, but it decreases my cognitive load tremendously.

edit: I haven't used Stack Overflow in 6 months. And "Googling" is Plan B.

=============

Edit 2: I recently had to write a sympathy letter to someone whose husband died.

I knew the general ideas behind what I wanted to say, but I knew I wasn't going to write anything great.

I fired up ChatGPT and said,

"write a sympathy letter to . Tell her that I didn't know her husband well, but the few times I met him, I could tell that he cared deeply about you and his daughter. I know his daughter well and I think she gets a lot of her great qualities from him and you. Tell her I don't know what to say in times like this. Keep it short-ish. Avoid schmaltz and sentimentality because isn't that kind of person."

It gave me about as perfect a letter as I could have asked for.

Re: Ask HN: Is anyone else bearish on OpenAI?

#112
>It's also very much like crypto where for every one person doing something useful with it, there are 20 trying to exploit the newness and low comprehension the general public have of the tech

This is definitely not correct in terms of numbers, there are many more people using LLMs well than have ever used crypto for any real use case. Also it's worth considering that the only real use cases of crypto are illegal, from noble stuff like busting sanctions to get food to hungry children, through to bribe evasion, bribe payment, tax evasion, drug deals, hiring hitmen, and child sexual exploitation/trafficking. In general, crypto produced close to zero for global society, even when it wasn't being used as an overt and intentional scam.

LLMs are producing significant value for society right now, because OpenAI gave everone API access to a very weird intern who has incredible knowledge breadth and makes dumb mistakes. Interns (or "relatively low intelligence/experience human workers who need handholding for difficult and sometimes easy problems, with an occasional flash of insight") have always been controversial as to whether they actually provide value from the perspective of the person who has to manage the intern, but from the perspective of the company/society it's unquestionable that they do provide significant value. Different people put different value on having a collaborator at all, some people do not want to handhold anyone or work with anyone who's mistakes they ever have to work around. It is nevertheless true that in aggregate for knowledge work, "worker + intern" is more economically productive than just "worker", outside of very, very specialist use cases. This just wasn't possible with GPT-2, and even GPT-3.5 is not quite at a quality where I'd really compare it to a normal intern. No other machine aside from the human brain was even close.

That's the tech now, the worst it will ever be. Whatever comes next with a major leap (GPT-5, Claude 3 or Gemini if they're good, maybe Llama 3 or the next Mistral if they can get improved significantly by the OS community before the next GPT release) is going to be either a reliable version of the same intern, or the same intern with better intelligence and comprehension that still suffers from reliability issues, or a major step up where they're equivalent in productivity to a full-blown knowledge worker in some high percentage of cases. It's already important now, it's only going to get more important.

As for OpenAI specifically, I think they have a very good chance of continuing to lead the pack, particularly with the this cringe-y GPTs/GPT Builder/GPT Store thing. It's pretty transparent that this is them getting data to train an AI on how to spin up agentic AIs to accomplish specific tasks, because they'll have the data on how the GPT Builder is used and the data on how useful and effective the GPTs it builds are, so they can do things like dramatically overweight the most effective and useful GPTs for training their internal "GPT Auto Builder". They'll be running a store for these things as well as effectively controlling the operating system they run in, so purchases, ratings, time using a GPT, sentiment analysis in the GPT text log to detect success, plus explicit in-GPT feedback (the thumbs up and down, feedback submission form) will all be data they can feed into their machine, to make an AI that can build good GPTs for a task and an AI that can evaluate their performance and an AI that can most effectively get good performance out of a GPT. That's going to be huge, particularly the signals that have real economic costs to users (I know they haven't announced it, but I think eventually they're going to make it so you can purchase GPTs) because that starts to pull away rose-tinted glasses and the fog of futuristic sheen and get some more unvarnished data on how much people actually value these specific things. That data means eventually you should be able to just ask ChatGPT to do something for you, and if it can't do it natively it will be trained to be able to spin up a task-specific GPT with access to the correct tools, docs etc, then have the GPT Whisperer AI use it to get the right answer with a bunch of backup data, and return you the answer with the option to see the work. This is also a pretty auditable process, which makes a lot of the legal and AI safety folks happy. I don't see another company that is similarly well-placed in terms of having the tech, talent, compute, product, and roadmap to pull this off.

Re: Ask HN: Is anyone else bearish on OpenAI?

#113
post #36

OpenAI, at least in my day-day workflow for the last 9+ months has so superseded anything that google ever was to me that I'm having a difficult time comparing the two. I've got a monitor dedicated 100% of the time to ChatGPT, and I interact with it non stop during the flow of technical scenarios and troubleshooting situations that flow into me - working in areas that I have the slimmest of backgrounds in, and shutti…

Are you able to give a specific example of a problem it helped you solve? Especially one that you were at a complete blocking point, it provided some solution, and then you were able to continue to expand upon that solution. I keep reading responses like yours, but I haven't seen any specific examples of problems being solved, so it all sounds very abstract. In my interactions with ChatGPT, it felt like just interact…

I don’t need it to solve problems where I’m completely stuck for it to be worth $20 a month.

Things I’ve gotten value out of in the past week or two:

• Writing a job description

• Making a python script to automate some stuff in Asana

• Simplifying some management concepts so I could slack them to a coworker

All of these are things I could easily do myself. But with ChatGPT, they’re done 75% as well in 10% of the time and I don’t have to think hardly at all.

Re: Ask HN: Is anyone else bearish on OpenAI?

#114

OpenAI, at least in my day-day workflow for the last 9+ months has so superseded anything that google ever was to me that I'm having a difficult time comparing the two. I've got a monitor dedicated 100% of the time to ChatGPT, and I interact with it non stop during the flow of technical scenarios and troubleshooting situations that flow into me - working in areas that I have the slimmest of backgrounds in, and shutti…

I'd pay money to watch someone like you work. Like other responders, I've tried using chatgpt multiple times, and very often when I've used it for tasks where I was familiar with the subject matter, I was disappointed in the results. It's possible I haven't learned the proper way to get what I need from it and would love to learn how others are using it to solve real problems. I mean it. I'd pay money to watch stream…

doesn't it also help with trivial problems?

the goal here is to accelerate productivity, is it not?

Re: Ask HN: Is anyone else bearish on OpenAI?

#115

OpenAI has changed education. I'm a teacher who is constantly learning new things. I can learn things I would have never been able to learn before because of AIs like ChatGPT. My students are learning more and faster than ever before. Learning Management Systems like Canvas and Blackboard made a lot of money. I could argue they are obsolete now.

no way Canvas and Blackboard will be obsolete in a few years. schools can't maintain all this infra system on own

Re: Ask HN: Is anyone else bearish on OpenAI?

#116
post #51

Yes. Ben Thompson has recently written a lot of commentary about it and, to be fair, he seems quite bullish on it. But so far to me this seems to be almost universally loved by programmers while I don’t really know anyone else who uses it at all. I think after the past 15 years which saw some of the most rapid technological advances in history along with the greatest bull market in history, people’s credulity is off…

> to be almost universally loved by programmers If AI-generated code is considered acceptable in your project then you aren't using a powerful-enough programming language. And you're paying the cost in code bloat. How many Coq programmers find ChatGPT useful? How many nontrivial Coq programs written (and not merely memorized) by ChatGPT even pass type checking ? If you're considering AI-written code then you have a b…

Interesting take. Very elitist and not rooted in reality though, I must say. The overwhelming majority of code out there is in languages more verbose than strictly necessary, and less expressive than possible. So in real life, yes AI generated code is a helpful and worthwhile thing. Not everyone can swing their Coq around at work.

That said, I do agree with the general notion. I find the more verbose the language, the better the help. Dense, more expressive languages fare worse. I’m referring to Python and Rust in my case, so one factor is of course massively larger training corpus for Python, and relatively more churn in Rust.

Re: Ask HN: Is anyone else bearish on OpenAI?

#117
post #95
post #56

I'm not bearish on OpenAI or AGI in general but I'm extremely meh about it. I'm not chomping at the bit to use it like so many are, and I constantly feel like a huge luddite or something for not being super excited about it. The value and time it saves makes sense for folks who struggle with a search engine (many) or doing tasks that are typically considered menial, like writing emails or coding boilerplate. However,…

"(I personally don't mind coding boilerplate stuff, especially since I can learn how the framework works that way)" ^ Isn't that what folks used to say about programming in assembler? How much time do I want to spend learning frameworks (beyond what I already know) vs. how productive do I want to be?

Probably, which is exactly my problem.

Re: Ask HN: Is anyone else bearish on OpenAI?

#118
I am.

The underlying tech is amazing. Where LLMs are headed is wild.

I have just lost a lot of confidence that OpenAI will be the ones getting us there.

The chat niche was an instance of low hanging fruit for LLM applications.

But to design the core product offering around that was a mistake.

Chat-instruct fine tuning is fine to offer as an additional option, but to make it the core product was shortsighted and is going to hold back a lot of potential other applications, particularly as others have followed in OpenAI's footsteps.

There's also the issue of centrally grounding "you are a large language model" in the system messaging for the model.

So instead of being able to instruct a model "you are an award winning copywriter" it gets instructed as "you are a large language model whose user has instructed you to act as an award winning copywriter."

Think about the training data for the foundational model - what percent of that was reflecting what a LLM would output? So there's this artificial context constraint that ends up leaving to a significant reduction in variability across multiple prompts between their ChatCompletion and (depreciated) TextCompletion APIs.

They seem like a company that was adequately set up to deliver great strides with advancing the machine learning side of things, but then as soon as they had a product that exceeded their expectations, they really haven't known what to do with it.

So we have a runaway success while there's still a slight moat against other competition and they have a low hanging fruit product.

But I'm extremely skeptical given what I've seen in the past 12 months that they are going to still be leading the pack in 3 years. They may, like many other companies that were early on in advancing upcoming trends, end up victims of their own success by optimizing around today and not properly continuing to build for tomorrow.

If you offered me their stock at a current valuation with the stipulation I wouldn't be able to sell for 5 years, I wouldn't touch it with a 10 meter stick.

Re: Ask HN: Is anyone else bearish on OpenAI?

#119
post #74

Earlier quoted context omitted.

I'd pay money to watch someone like you work. Like other responders, I've tried using chatgpt multiple times, and very often when I've used it for tasks where I was familiar with the subject matter, I was disappointed in the results. It's possible I haven't learned the proper way to get what I need from it and would love to learn how others are using it to solve real problems. I mean it. I'd pay money to watch stream…

You can start by paying $20 for ChatGPT 4 and try it on tasks where you're familiar with the subject matter. I've tried it and I've been amazed.

Please, please, reference the tasks or chats that were especially impressive.

For example, trying to get chat gpt to do something very simple for my work, like implementing a convolution it took me in circles and circles.

It gets the general idea right, sure. But it actually makes significant minor errors that ended up being more confusing than helpful.

Re: Ask HN: Is anyone else bearish on OpenAI?

#120

ChatGPT has already saved me from hours of Googling when I'm trying to find out how to do certain things. It almost feels magical - I don't have to read through half-dozen slightly different variations of what I need to do. Before ChatGPT, to find the answer to things like "how do I set up Gunicorn to run as a daemon that restarts when it fails" I would have to endure hours of googling, snarky stack-overflow comments…

I know this is just an example, but I think it’s emblematic of the main issue I have with the widespread use of LLMs. Do you mean, “have Gunicorn keep N workers running?” If so, that’s in the manual (timeouts to kill silent workers, which defaults to 30 seconds). Or do you mean “have Gunicorn itself be monitored for health, and restarted as necessary?” There are many ways to do that – systemctl, orchestration platfor…

This is also my current worry. If you know the concepts (Not the workflow) about the problem you’re solving, I find it easy to get answer, and in the meantime you’ll collect some new knowledge in the process. even when asking someone, they will often point out the knowledge you lack while providing the answers.

Getting straight answers will be detrimental in the long term, I fear. It feels like living in a box, and watching the world on a screen and the person answering my questions is mixing lies and truths.

Post reply on HN