It wields reddit as meaningful source of openAI's performance, and then goes down a rabbit hole from there.
OpenAI's moat is the thousands of 'human in the loop' contractors that they hired in south america... for years... Not reddit...
311–319 of 319 posts
It wields reddit as meaningful source of openAI's performance, and then goes down a rabbit hole from there.
OpenAI's moat is the thousands of 'human in the loop' contractors that they hired in south america... for years... Not reddit...
Earlier quoted context omitted.
I have literally never heard this word before the infamous Google post, since then it's everywhere.
What google post? I think I missed that one
I think from Reddit's perspective, they are extremely upset with OpenAI, in the same way that I'm sure StackOverflow is upset -- OpenAI took: - The entire corpus of data the community had curated over the last XX years - The "goodwill" that these platforms had developed towards third party developers in allowing developers to work with their data - Potentially large amounts of traffic that would normally come to thei…
Given that sama is a board member of Reddit Inc, and that this is happening after GPT-4 was trained on Reddit data, I wouldn't jump to conclude they're upset at OpenAI. SO had publicly available, no-auth-required data dumps. This makes it difficult for them to know who is using their data. However, this surely isn't the case for Reddit who offered only API endpoints for this content, and I'm guessing you couldn't use…
I think from Reddit's perspective, they are extremely upset with OpenAI, in the same way that I'm sure StackOverflow is upset -- OpenAI took: - The entire corpus of data the community had curated over the last XX years - The "goodwill" that these platforms had developed towards third party developers in allowing developers to work with their data - Potentially large amounts of traffic that would normally come to thei…
Google or Reddit don't matter, if the Info Explosion problem that the Internet produced gets solved some other way. Main reason such sites came into existence was there was too much info on the Internet. These sites where attempts at simplifying the numerous websites and blogs ppl had to manually discover and track. They have meandered around that problem, and got totally distracted by all kinds of other problems (ma…
Earlier quoted context omitted.
I appreciate everything you and the mods do here. Enough so that I feel my past behavior on HN has been less than ideal from time to time and I am incentivized to improve it because I respect what you're trying to do here. However, the staff at HN are doing themselves a disservice and legitimizing criticism with this sort of action. If you feel this content is grand and speculative then downvote it or flag it like an…
I appreciate that you're working to improve your contributions to HN! You might be underestimating how much curation/moderation we do of HN's front page. We do a lot. The front page is not generated by votes and flags alone ( https://hn.algolia.com/?dateRange=all&page=0&prefix=false&so... ). Upvotes and flags are important but the front page is a complex combination of those plus software action plus moderator action…
Earlier quoted context omitted.
Given that sama is a board member of Reddit Inc, and that this is happening after GPT-4 was trained on Reddit data, I wouldn't jump to conclude they're upset at OpenAI. SO had publicly available, no-auth-required data dumps. This makes it difficult for them to know who is using their data. However, this surely isn't the case for Reddit who offered only API endpoints for this content, and I'm guessing you couldn't use…
> I think this is more important to Reddit than ad revenue, as they could've simply built an SDK for probably less than this PR nightmare will cost us. Who is "us"?
Earlier quoted context omitted.
It lies like crazy on surprising things. Database parameters for an enterprise provider, for example, I've seen hallucination in 5% of cases. That's _bad_ when its taken as authoritative.
I like that you can get results quickly and it solves the "I don't know what I'm looking for" problem. I'm learning typescript and ran into a weird typescript construct the other day, I threw it into chatgpt and asked "what is this" and it explained it to me. I'm not entirely sure if pasting in a bunch of braces and parentheses into google would find me the same result.
Earlier quoted context omitted.
> I genuinely don't understand the appeal of AI for search. The provenance of information is just as important as the resulting information for pretty much anything I search for. Search requires work on the part of the user to distinguish between good links and bad. AI is an oracle just tells you what you're looking for. Now you and I might think this is a terrible way to evaluate the veracity of information. But thi…
It sounds like we are dooming an entire generation by catering to a maladaption of technology-addiction induced ADD.
Earlier quoted context omitted.
It lies like crazy on surprising things. Database parameters for an enterprise provider, for example, I've seen hallucination in 5% of cases. That's _bad_ when its taken as authoritative.
None of the results are surprising if you simply read what OpenAI writes about their own machine. It's an autocomplete engine.