Live data from Hacker News

Ask HN: 6 months later. How is Bard doing?

news.ycombinator.com

171–180 of 218 posts

Re: Ask HN: 6 months later. How is Bard doing?

#171
I've been doing a lot of coding using google apps script lately for personal projects. ChatGPT still runs circles around Bard when it comes to providing workable code suggestions and fixes when something doesn't work. I test against Bard regularly and never fail to be surprised how bad Google's own "AI" is at even helping develop on its own platform.

I use chatall (find it on github) which searches all the freely available AIs and delivers answers from all of them. That's been a great way to check the pulse on accuracy

Re: Ask HN: 6 months later. How is Bard doing?

#172

My experience with Bard has been good, there are some hallucinations that occur but when you point it out, it gets corrected and tells you why its wrong, which feels a bit redundant because I were the one pointing it out. Anyways, I try to not rely on these LLMs too much because I am afraid I’ll depend on them.

> Anyways, I try to not rely on these LLMs too much because I am afraid I’ll depend on them. You will, we will. We depend on syntax highlighting as well. LLMs are here to stay, so I am not worried.

Is syntax highlighting really comparable to LLMs? Like I get your point, but are they really that close to each other?

Re: Ask HN: 6 months later. How is Bard doing?

#173
post #27

Look at Gemini, it’s their new model, currently in closed beta. Hearsay says that it’s multimodal (can describe images), GPT-4 like param count, and apparently has search built in so no model knowledge cutoff. Basically they realized Bard couldn’t cut it and merged DeepMind into Google Brain, and got the combined team to work on a better LLM using the stuff OpenAI has figured out since Bard was designed. Takes months…

I think the DeepMind / Brain reorg happened way before all this, didn't it? Might be misremembering history...

The merger was April ‘23, GPT-4 was March ‘23 which is too close to be the trigger for the re-org, so it was in response to ChatGPT (GPT-3.5). There was talk of a “code red” around that time (December ‘22).

https://www.theverge.com/2023/4/20/23691468/google-ai-deepmi...

https://www.aiwithvibes.com/p/google-issues-code-red-respons...

Re: Ask HN: 6 months later. How is Bard doing?

#174

Really wishing benchmarks for AI included evaluating how well they come up with plans for peaceful anticapitalist revolution. This is not a joke.

Be the change you want to see in the world.

Yes, I'm using AI to design systems for revolution. I can still lament the absence of benchmarks for it. I care for a 5-year old 24/7, so my time for this kind of thing is a bit limited. I see u'all are into graph theory. There's a network structure called a "selection reactor" from evolutionary graph theory I'm trying in a fractal pattern of human irl interaction to see if it rapidly evolves human culture in advantageous ways. Want to help me be the change I want to see in the world & explore this idea further?

Re: Ask HN: 6 months later. How is Bard doing?

#175

Earlier quoted context omitted.

I had a similar issue so I made https://TLDWai.com to summarize YouTube videos

I could use this for my project but most of my videos don't have any dialogue or voice overs. It would be perfect if it described the actual (visual) video content.

For now it transcribes the audio of the video using Whisper.cpp; but what you say is a good feature that I will be reviewing.

Re: Ask HN: 6 months later. How is Bard doing?

#176

Earlier quoted context omitted.

I had a similar issue so I made https://TLDWai.com to summarize YouTube videos

how is that site fairing for traffic and conversions?

It was only launched a couple months ago, so low traffic, and no conversions

Re: Ask HN: 6 months later. How is Bard doing?

#177
Bard is pretty terrible. Spent a few hours testing it out. Beyond just giving an incorrect or incomplete answer, it has repeatedly lied about knowing my location and how it knows it. It has also claimed a friend was dead and his son was selling his home through a trust.

Re: Ask HN: 6 months later. How is Bard doing?

#178

Earlier quoted context omitted.

What an absolute slurry this is: Jumping from defining sentience in terms of what upsets people when subjected to animal cruelty... to arbitrarily selecting chickens as a lynchpin based on that. Then diving on deeper still on a rain puddle deep thought. Fruit flies are also sentient, while you're out here inventing thresholds why aim so high? You could have even gone with a shrimp and let Weizenbaum know ELIZA was se…

I defined sentience as experiencing qualia, then decided to back up my assertion that most people consider animals to be sentient with an example. Pain is the one animal sensation humans care about, so I picked animal cruelty. I chose chickens because they're the dumbest animal that humans worry about hurting. I'm sorry that you've taken umbrage with my example. I didn't select fruit flies because I don't think a maj…

You literally redefined sentience again here: "don't consider them sentient" and your flag pole is "or I guess not enough because they don't feel bad about killing them".

You're drawing arbitrary goalposts for goals that aren't even relevant: At the end of the day we don't need philosophy to prove Lemoine was a schmuck.

Millions got got access to RLHF chat. We can see how they would have made his initial mistake. But following up with months of badgering and protest after being guided with kiddie gloves until he gets fired was the height of delusion.

The fact he now works on optimizing for the thing he rang the alarm on says it all.

Also your comment shows why guidelines aren't perfect: From your first comment you've taken the most aggrandizing borderline inflammatory tone possible without technically being un-nice.

Are little pot shots like: "So how can a subject-matter expert confidently prove that a language model isn't sentient? And please let David Chalmers know while you're at it, I hear he's keen to settle the matter."

really justified after dropping what, quite frankly, was not a well formed or even self-consistent argument?

Not everyone plays the HN backhanded-niceness game: I assume most people here are adults and can handle some directness.

Re: Ask HN: 6 months later. How is Bard doing?

#179

Earlier quoted context omitted.

Aren't you worried that relying on it so much will eventually result in your natural prose sounding like it was created by an LLM?

LLMs can write as naturally as you want. You can even paste some of your writing to emulate. The default style is just the default.

Hmm, maybe it got better at this but my experiments a few months ago were pretty underwhelming. It did ok at emulating really distinct styles with lots of examples (Shakespeare) but surprisingly badly at more subtle styles with less examples (Tupac). In the latter case it would revert to default-speak after a while, with the odd bit of vocabulary thrown in. Tupac's entire ouevre is online, so it should be able to emulate him flawlessly. How much of my text will I have to feed it so it sounds like me?

My point is that we learn to write by reading. If someone is constantly looking at chatGPT output as exemplar that's going to change the way they write. The comment I was replying to is classic default GPT style, especially that last paragraph, even if it was written by a human.

Re: Ask HN: 6 months later. How is Bard doing?

#180

Earlier quoted context omitted.

This makes sense. In the "token window" of a human being, the same strategy would also work, e.g., p1: "What do you think of my story? Be honest." p2: "I'd rather not say." p1: "Seriously, tell me what you think, it's fine if you hate it. I need the feedback." When you think about it from that perspective, it's no dumber than people are.

That's what I find so interesting about LLMs. I have yet to see a single criticism of them that doesn't apply to humans. "Well, it's just a stochastic parrot." And most people aren't? "Meh, it just makes stuff up." And people don't do that? "It doesn't know when it's wrong." Most people not only don't know when they're wrong, they don't care . "It sucks at math." Yeah, let's not go there. "It doesn't know anything th…

How about “it cannot tell if you if it made something up / guessed with intuitive levels of confidence”.
Post reply on HN