Earlier quoted context omitted.
here's an extension I created for this https://chrome.google.com/webstore/detail/hacker-news-filter... firefox : https://addons.mozilla.org/en-US/firefox/addon/hacker-news-f... . no llm filter yet but I can add that too.
Nice! We're wondering if we should build a browser extension based on the "similar stories" model there, if you would be interested in doing something like that happy to have a chat!
It's not only you, there's an explosion of ChatGPT on HN
121–130 of 148 posts
Re: It's not only you, there's an explosion of ChatGPT on HN
#122Earlier quoted context omitted.
In case you don't click the links above: dang recognizes the tsunami of AI content and it's being downweighted. Some people complain about all the AI content, and those complaints are tiresome too. Thanks dang. It's good to know that 1) things are being done about it (even if it's not enough for some people), and 2) we'd be better off resisting our urge to complain.
Sorry - I meant to add an explanation and got distracted. Thanks for doing it for me :) If anyone reads those links and still has a question I haven't answered there, I'd be happy to take a crack at it. It's a tricky situation because all of these things are true: (1) it's the biggest technological development in a long time; (2) there's way too much material about it, a lot of which is mediocre; (3) every user has a…
Some of the new research coming out is very interesting.
Re: It's not only you, there's an explosion of ChatGPT on HN
#123Earlier quoted context omitted.
In case you don't click the links above: dang recognizes the tsunami of AI content and it's being downweighted. Some people complain about all the AI content, and those complaints are tiresome too. Thanks dang. It's good to know that 1) things are being done about it (even if it's not enough for some people), and 2) we'd be better off resisting our urge to complain.
Sorry - I meant to add an explanation and got distracted. Thanks for doing it for me :) If anyone reads those links and still has a question I haven't answered there, I'd be happy to take a crack at it. It's a tricky situation because all of these things are true: (1) it's the biggest technological development in a long time; (2) there's way too much material about it, a lot of which is mediocre; (3) every user has a…
What I was surprised by putting this together was just how much similar content is posted here every day - just click on the "Similar stories" under each post on our site to see how almost any given topic has quite a lot of content! Makes one give value to the curation being done by users and mods here that avoid making this repetition too visible (and annoying) on the site!
Re: It's not only you, there's an explosion of ChatGPT on HN
#124Earlier quoted context omitted.
after having studied this more and listened to a talk by one of the OpenAI team I think this will actually go away as a problem in the semi-near term future. Is that talk available online? I’m skeptical that they will ever solve the problem of factuality. I’d love to hear their arguments for why I’m wrong.
Sure, here you go: https://www.youtube.com/watch?v=hhiLw5Q_UFg Summarizing: 1. Obviously you can't completely "solve" truthfulness because people disagree on what is and is not true. But you can go a long way. 2. The models do know what they don't know. Their level of uncertainty is not only expressed in the final token logprobs but also seems to be reified somehow, such that they can express their own level of certa…
I've now taken the time to watch John's talk and I have some thoughts. It's not only difficult to solve truthfulness due to disagreement (and subjectivity), but it's also very difficult because different contexts have different standards of evidence.
In some contexts, like programming, we'd rather have the model output its best guess for what the program should be, no matter how low the confidence, because we would like a starting point and we can debug the program from there. The answer "I don't know how to write that program" is not a useful starting point and it may even be an example of the model withholding information it does have due to low confidence.
In other contexts, such as scientific or historical questions, we want a high standard of evidence. Asking the question "what year did Neil Armstrong land on Mars?" should not produce a hallucinated response with fully unhedged language complete with fictitious date of landing. This problem may be solvable by training the model to hedge or even to question the premise when the confidence is low. Of course, this also suffers from the garbage-in-garbage-out problem of having falsehoods buried in the training set.
A more subtle and difficult problem with scientific/historical questions is with long-form answers. Currently, models tend to produce long-form answers that fairly consistently contain a mixture of true facts and falsehoods, and it can be quite difficult for even expert readers to spot all of the mistakes every time. Furthermore, the human labellers were given very sophisticated tools for highlighting sentences in long-form output but the information this produced had to be reduced down to a single bit per example since the detailed information did not improve training very much.
Personally, I think it's going to be very difficult to teach the model how to recognize the appropriate contexts and associated standards. This is a very subtle problem and one of the issues is that it relies on information the model does not have access to, for example: the identity of the question-asker. If a child asks an astrophysicist about black holes they're going to get a different answer than if an undergraduate student asks the same question in class. Yes, this additional context can be included in the prompt, but at some point it becomes a pain to have to copy-and-paste the context for every prompt.
Perhaps people will create a tool to save this additional context in the form of presets but this imposes additional effort on humans. At some point I think the amount of human curating and feedback that goes into these models will cause a collapse and backlash. We saw the same thing happen in the early days of search engines, when Google (fully automated) trounced Yahoo (human curated), leading to Yahoo's abandonment of human curation. We also see the same problem manifest itself at the Patent Office, where human review is policy. The entire patent system has become grossly dysfunctional at least partly due to the overwhelming complexity of this problem.
One thing I really liked was the "inner monologue" of the model performing a sequence of steps to answer a question by doing a search. If this could be generalized to other tasks it could be a home run for automated assistants (Google/Alex/Siri).
Re: It's not only you, there's an explosion of ChatGPT on HN
#125Earlier quoted context omitted.
You'd have to substantiate that claim which would probably take significant work.
From 2015 to today it appreciated 8000%. Do I need to elaborate further?
You can't find a single other investment that would have returned higher? Very different, but maybe options plays on GME? Anything like that?
I'm sure there is one! You said it's the HIGHEST - very specific claim!
You're also picking a date period that works for you - certainly I can pick other dates that don't support your argument!
Re: It's not only you, there's an explosion of ChatGPT on HN
#126Earlier quoted context omitted.
Partially, it is hard to "prove the negative" in a way. The real issue and logical fallacy is people are frontrunning the argument: they are making extraordinary claims of either the current capabilities of the tech or future capabilities, without offering evidence beyond "it's obvious!" or "have you tried it?" or "use the latest model, it's much better!" etc.. and maybe they don't realize it, but these are not good-…
I know exactly where you're coming from. I've tried to avoid the hype train but I still feel the hot wind from it because I literally had my manager ask me if I could use ChatGPT to do my last project faster. I was astounded. It's tough when you know a dumb idea is not going to work, but now you have to articulate exactly why to your manager, without insulting them. Also without wasting a lot of valuable time proving…
Re: It's not only you, there's an explosion of ChatGPT on HN
#127Earlier quoted context omitted.
I can't really argue with investing in BTC or Eth, I probably should have for some portion of my portfolio!
It’s never too late :]!
Re: It's not only you, there's an explosion of ChatGPT on HN
#128Earlier quoted context omitted.
Good question. GP is claiming that because ycombinator (YC) investors have put money into AI, we're going to see a lot of AI posts here for the foreseeable future. That's because YC companies use HN to promote their products.
There are hundreds of posts in HN every day and only a few dozens are upvoted to the first page. Assuming you trust YC not to game the upvotes, the amount of AI on the front page should primarily be explained by the interest of the community.
Re: It's not only you, there's an explosion of ChatGPT on HN
#129Earlier quoted context omitted.
GPT can tell you what's wrong with Bitcoin: You throw 25 grand on a Bitcoin bet, The dough lands in a chest, smoky and wet. You've got your Bitcoin, feelin' alive, But the chest has plans, man, it's a damn beehive. It pays the skeptics, those withdrawing in fright, The electric bill's covered, in the dark of the night, A Lambo, a threesome, a few folks living in sin, As the value hits 27 grand, you think it's your tu…
We need a “!RemindMe 10 years” feature like they have on Reddit.
Put the permalink on your calendar.
Re: It's not only you, there's an explosion of ChatGPT on HN
#130Earlier quoted context omitted.
> like Bitcoin did You think it's over for Bitcoin?
It surely looks like it is getting there https://hn.curiosity.ai/#/trends?terms=Bitcoin;Web3
Careful not to overfit ;)
Neat tool, thanks for sharing.