Live data from Hacker News

Phind-405B and faster, high quality AI answers for everyone

phind.com

131–140 of 163 posts

Re: Phind-405B and faster, high quality AI answers for everyone

#131

Earlier quoted context omitted.

That's all they can do. They seem impressive at first because they're basically trained as an adversarial attack on the ways we express our own intelligence. But they fall apart quickly because they don't have actually have any of the internal state that allows our words to mean anything. They're a mask with nothing behind it. Ctrl+F for "Central nervous system": https://en.wikipedia.org/wiki/List_of_human_cell_types…

Wait for the first large scale LLM using source-aware training: https://github.com/mukhal/intrinsic-source-citation This is not something that can be LoRa finetuned after the pretraining step. What we need is a human curated benchmark for different types of source-aware training, to allow competition, and an extra column in the most popular leaderboards, including it in the Average column, to incentivice AI companies…

Yeah. Treating these things as advanced, semantically aware search engines would actually be really cool.

But I find the anthropomorphization and "AGI" narrative really creepy and grifty. Such a waste that that's the direction it's going.

Re: Phind-405B and faster, high quality AI answers for everyone

#132
post #54

Phind continues to be my favorite AI-enhanced search engine. They do a really nice job giving answers to technical questions with links to references where I can verify the answer or learn more detail. Some recent examples from my history: what video formats does mastodon support? https://www.phind.com/search?cache=jpa8gv7lv54orvpu2c7j1b5j compare xfs and ext4fs https://www.phind.com/search?cache=h9rmhe6ddav1bnb2odtc…

In my tests, it does hallucinate answers, even with Phind 70B. For example, I asked for bluetooth earplugs that have easy battery replacements. It always kept giving me answers for earplugs with I know have their battery soldered into the casing. Tbf, perplexity also fails at this question.

I just asked your question in French and the right answer (the Fairphone Fairbuds) is in third position on the right: https://imgur.com/a/dmoaB5r

Google seems to be better at this, giving me the Fairbuds directly: https://imgur.com/a/7En4e9u

Re: Phind-405B and faster, high quality AI answers for everyone

#133

Earlier quoted context omitted.

Absolutely. Behaviour that in normal life in clean societies would be "eliciting violence": automated hypocritical lying, apologizing in form and not in substance, making statements based on fictional value instead of truthfulness...

That's all they can do. They seem impressive at first because they're basically trained as an adversarial attack on the ways we express our own intelligence. But they fall apart quickly because they don't have actually have any of the internal state that allows our words to mean anything. They're a mask with nothing behind it. Ctrl+F for "Central nervous system": https://en.wikipedia.org/wiki/List_of_human_cell_types…

This is just the start. Imagine giving up on progressing these models because they're not yet perfect (and probably never will be). Humans wouldn't accomplish anything at all this way, aha.

And I wouldn't say lazy at _all_. I would say efficient. Even evolutionary features that look "bad" on the surface can still make sense if you look at the wider system they're a part of. If our tailbone caused us problems, then we'd evolve it away, but instead we have a vestigial part that remains because there are no forces driving its removal.

Re: Phind-405B and faster, high quality AI answers for everyone

#134

Recently I asked an AI following question const MyClass& getMyClass(){....} auto obj = getMyClass(); this makes a copy right? And it was very confident about it not making a copy. It thinks auto will deduce the type as a const ref and not make a copy. Which is wrong, you need auto& or const auto& for that. I asked it if it is sure and it was even more confident. Here is the godbolt output https://godbolt.org/z/Mz8x74…

You prove the point that these are just token generation machines whose output is psuedo-intelligent. It’s probably not there yet to be blindly trusted.

More to the point; I wouldn't blindly trust 99% of humans, let alone a machine.

Though to be fair we will hopefully quickly approach a point where a machine can be much more trusted than a human being, which will be fun. Don't cry about it, it's our collectives faults for proving that meat bags can develop ulterior motives.

Re: Phind-405B and faster, high quality AI answers for everyone

#135

Earlier quoted context omitted.

‘Easy battery replacements’ is pretty subjective. This feels like one of those error that demonstrate how good the tech is, because it’s being used for very specific and subjective requests

It should be able to figure out what an average joe would understand from such questions. I think any human would interpret "easy battery replacement" as "you can just remove the old battery out and put the new one in". If a random person asked you such a question, would you assume he has the tools and the skill needed to solder new batteries and considers that easy?

this is why search sucks the average person(if they didn't really know what you meant would ask...what are you talking about?

Re: Phind-405B and faster, high quality AI answers for everyone

#137

Earlier quoted context omitted.

That's all they can do. They seem impressive at first because they're basically trained as an adversarial attack on the ways we express our own intelligence. But they fall apart quickly because they don't have actually have any of the internal state that allows our words to mean anything. They're a mask with nothing behind it. Ctrl+F for "Central nervous system": https://en.wikipedia.org/wiki/List_of_human_cell_types…

This is just the start. Imagine giving up on progressing these models because they're not yet perfect (and probably never will be). Humans wouldn't accomplish anything at all this way, aha. And I wouldn't say lazy at _all_. I would say efficient. Even evolutionary features that look "bad" on the surface can still make sense if you look at the wider system they're a part of. If our tailbone caused us problems, then we…

> This is just the start

But the issue is with calling finished products what are laboratory partials. "Oh look, they invented a puppet" // "Oh, nice!" // "It's alive..."

Re: Phind-405B and faster, high quality AI answers for everyone

#138

Earlier quoted context omitted.

> They can only lie That is definitely not true.

Lying is a state of mind. LLMs can output true statements, and they can even do so consistently for a range of inputs, but unlike a human there isn't a clear distinction in an LLM's internal state based on whether its statements are true or not. The output's truthfulness is incidental to its mode of operation, which is always the same, and certainly not itself truthful. In the context of the comment chain I replied t…

I am not sure that lying is structural to the whole system though: it seems that some parts may encode a world model, and that «the sensory, object permanence, and memory faculties» may not be crucial - surely we need a system that encodes a world model and that refines it, that reasons on it and assesses its details to develop it (I have been insisting on this for the past years also as the "look, there's something wrong here" reaction).

Some parts seemingly stopped at "output something plausible", but it does not seem theoretically impossible to direct the output towards "adhere to the truth", if a world model is there.

We would still need to implement the "reason on your world model and refine it" part, for the purpose of AGI - meanwhile, fixing the "impersonation" fumble ("probabilistic calculus say your interlocutor should offer stochastic condolences") would be a decent move. After a while with present chatbots it seems clear that "this is writing a fiction, not answering questions".

Re: Phind-405B and faster, high quality AI answers for everyone

#140

Does anybody use Phind? What do you use it for?

I was subscribed for about 6 months between the end of last year and beginning of this, but canceled and haven't looked back. The web interface was constantly buggy for me, and they seemed to be very focused on the VSCode extension without integrations for other editors, so I ended up canceling.

what are you using now?
Post reply on HN