Live data from Hacker News

ChatGPT is now finding bugs in databases

celerdata.com

81–90 of 204 posts

Re: ChatGPT is now finding bugs in databases

#81

Earlier quoted context omitted.

The current AI revolution is a devolution of the web. The thesis of the web revolution was “the world is too big to be understood by any one person”. The response was “let us all contribute to a shared map to compensate” The thesis of the AI revolution is “the world seems compressible”. The response has been “let us condense and contain all the meaningful parts in a single generative pattern that can understand every…

Isn't that solved with supervised learning? The way I see it is that internet content is used to bootstrap the models, then supervision is used to train the models without the risk of a feedback loop causing quality loss. I'm pretty new to ML so I may be missing something.

Supervision requires knowing the results you want.

Most of these ML projects are essentially just creating a best fit line/shape connecting a huge number of points in multiple dimensions and giving you output coordinates on that line/shape based on your input coordinates (as I understand it). The more supervision, the more you’re negating the value, as you’re basically telling it to make a shape more like something you already understand (instead of something new/actually generative, which requires interesting/novel human input)

I’m not an ML expert either, and if one wants to chime in about how this picture I’m painting is wrong or what else is going on that would be welcome. I’m not trying to belittle how impressive progress has been (I have no idea how the parameters are determined and have a huge amount of respect for people able to handle a hyper-dimensional best fit optimization problem). But I don’t see how all the value isn’t inevitably downstream of high quality human generated digital content, which seems likely to decrease rapidly as more automated content floods the internet and lowers incentives for creators.

Re: ChatGPT is now finding bugs in databases

#82

Earlier quoted context omitted.

That's almost every reddit thread I've been involved in. Oratory prowess is way more important than accuracy. If you ever want to "lose a debate" on reddit, pick something that's counter-intuitive or widely misunderstood and use nuance and citations in your defense. You'll get cyberbullied every time - sometimes even banned from the sub.

From those examples, I think you’re maybe also running up against people generally having a low threshold for giving a stranger on the internet the benefit of the doubt. It’s hard to be surprised that “eugenics, but the good kind, just hear me out!” and “state’s rights, but not the disingenuous lost cause kind, just hear me out!” are a tough sell with complete strangers who have no prior reason to trust you or be int…

correct!

Btw, I removed them from my examples, didn't want to "discuss" that here.

I'm kind of a history nerd and it's like a fishing hook with delicious bait when I see stuff like that. "Oh wait! You think it was just the south that argued states rights?! Actually (brings up fugitive slave clause jurisprudence)"

Or how manner and conduct books of the 1800s, eugenics texts, and jordan peterson's 12 rules for life are basically dots in a straight line. That stuff is just catnip.

I just need to stop being distracted and go write actual books or something.

Re: ChatGPT is now finding bugs in databases

#83
post #78

Earlier quoted context omitted.

Creating front ends that integrate different AI APIs is kind of over saturated, that's probably why you're not getting much attention. It's also somewhat hard to understand what's the value proposition of the website, you should focus on putting as much of that kind of information in everyone's face in the landing page, otherwise they'll just bounce. Take everything from your "Learn more" page and put it in your land…

How do I get it to an incubator? I believe in it, because its 100% useful in incredibly obvious ways. And you are basically saying that there is no such thing as bootstrapping. And honestly I feel that this is a failure of HN not of my website. But you still give good advice to help people by not expecting them to click a single link to get to the FAQ page.

Bootstrapping is cheap and easy, but the chance of failure is much larger, it also comes with the consequence of giving you a lot of unnecessary stress since you're playing with your own assets. Regarding applying to an incubator: https://www.ycombinator.com/apply

All the best.

Re: ChatGPT is now finding bugs in databases

#85

Earlier quoted context omitted.

This is why I've said Trump is right about fake news. When I say that people think I'm a weirdo Trumper.

One nuance might be that those are two difference senses of fake news. In one case, we have news that is wrong. In another case, we have an insult aimed at reporting that we don't like, regardless of accuracy.

I haven't trusted the news since the time I participated in manipulating local news stations to create stories about jenkem. More modern equivalent is Momo. It's all fake

Re: ChatGPT is now finding bugs in databases

#86

Earlier quoted context omitted.

Before the had the internet, back when we still had newspapers, we had a similar phenomenon referred to as Gell-Mann Amnesia. > Briefly stated, the Gell-Mann Amnesia effect is as follows. You open the newspaper to an article on some subject you know well. In Murray’s case, physics. In mine, show business. You read the article and see the journalist has absolutely no understanding of either the facts or the issues. >…

I think the effect is a bit overblown. Yes, I have noticed coverage of topics I know about that are oversimplified or straight up mistaken. But these are technical topics that I don’t expect journalists to get right (and in many cases I’d prefer they didn’t try to cover them at all). And I also won’t implicitly trust coverage about other technical subjects that I’m unfamiliar with. I would argue that national politic…

Check out this excerpt from a local news story I saw yesterday. You know what else acetic anhydride is used in? Production of aspirin. This is basic scare tactics to manipulate people and I see it everyday. A combination of ignorance, laziness, and being the voice of those in power. Fake news and manipulation of the public.

https://www.valleynewslive.com/2023/03/09/fargo-cbd-store-ow...

"“Semi-synthetic derivatives, or chemically derived cannabinoids, refer to certain types of substances that are produced by converting a cannabis extract into a different substance through chemical reaction,” said forensic scientist with the North Dakota State Crime Laboratory Charlene Rittenbach at the March 3 committee hearing of the bill.

For example, according to the Addiction Prevention Coalition, to make THC-O: you need to extract CBD, extract Delta 8 THC from the CBD and then add acetic anhydride.

The National Library of Medicine says acetic anhydride is corrosive to metals and tissue, and it’s used to make fibers, plastics, dyes and explosives."

Re: ChatGPT is now finding bugs in databases

#87
post #3

ChatGPT can't access the external internet. The example ChatGPT instruction where they "teach it" just has the user giving it a URL to their docs and ChatGPT is like "oh ok got it" (Figure 3). I'm not sure if the blog post writers are aware it's hallucinating what the JSON implementation is and hasn't actually read the documentation at all.

Is it possible that it "remembers" that url from it's training and the content tied to it?

I don't think they prefix each document with the URL in the training set, but there might be URLs in the clear text.

Re: ChatGPT is now finding bugs in databases

#88

Earlier quoted context omitted.

From those examples, I think you’re maybe also running up against people generally having a low threshold for giving a stranger on the internet the benefit of the doubt. It’s hard to be surprised that “eugenics, but the good kind, just hear me out!” and “state’s rights, but not the disingenuous lost cause kind, just hear me out!” are a tough sell with complete strangers who have no prior reason to trust you or be int…

correct! Btw, I removed them from my examples, didn't want to "discuss" that here. I'm kind of a history nerd and it's like a fishing hook with delicious bait when I see stuff like that. "Oh wait! You think it was just the south that argued states rights?! Actually (brings up fugitive slave clause jurisprudence)" Or how manner and conduct books of the 1800s, eugenics texts, and jordan peterson's 12 rules for life are…

I’m no fan of Peterson, but I’m curious— how are intentionally altering the gene pool of a group and cleaning your room colinear?

Re: ChatGPT is now finding bugs in databases

#89

Earlier quoted context omitted.

One nuance might be that those are two difference senses of fake news. In one case, we have news that is wrong. In another case, we have an insult aimed at reporting that we don't like, regardless of accuracy.

I haven't trusted the news since the time I participated in manipulating local news stations to create stories about jenkem. More modern equivalent is Momo. It's all fake

This sounds interesting! What did you do?

Re: ChatGPT is now finding bugs in databases

#90
post #19
post #3

ChatGPT can't access the external internet. The example ChatGPT instruction where they "teach it" just has the user giving it a URL to their docs and ChatGPT is like "oh ok got it" (Figure 3). I'm not sure if the blog post writers are aware it's hallucinating what the JSON implementation is and hasn't actually read the documentation at all.

There's this internet phenomenon (common on Reddit and other such forums) where you find high-quality discussions on a variety of topics and everyone seems so clever and informative, until you come across a thread on a topic that you are actually an expert on and find that every top-voted comment is objectively incorrect. ChatGPT is essentially that but on steroids. Its primary ability is being able to sound extremel…

That was also common on TV. Everyone seems credible, until we spot errors when someone in our field starts talking.

edit: minor typo

Post reply on HN