Live data from Hacker News

Phind-405B and faster, high quality AI answers for everyone

phind.com

81–90 of 163 posts

Re: Phind-405B and faster, high quality AI answers for everyone

#81
post #58

Earlier quoted context omitted.

I payed and used 6 months for Phind. I am more satisfied with the Kagi Assistant currently. It does not give that many links but overall results are as good or even better, and you can use lenses. You get general search engine too. There was one UI related annoyance with Phind; scroll bar sometimes jumped randomly, maybe even after each input or during token generation (on Firefox). You start wasting a lot of time if…

Thanks for the feedback. We've fixed the UI jumping issue. The new Phind update today should also work as a general search engine.

The icons cover up the input area on this crappy android work phone.

I stubornly continued to type my complaint about my json getting to large for phones with slow cpus or slow connections and got 100 solutions to explore. I couldnt help but think this is the worse case robot overlord, it gave me a year worth of materials to study complete with the urge to go do the work. That future we use to joke about is here!

Some of the suggestions are familiar but i dont have the time to read books about with little titbits of semi practical information smeard out over countless pages in the wrong order for my use case.

Im having flashbacks reading for days, digging though a humongous library only to end up with 5 lines of curl. I still cant tell if im a genius or just that dumb.

This long response unexpectedly makes me want to code all day long. One can chose where to go next which is much more exciting than the short linear answers... apparently.

Well done

Re: Phind-405B and faster, high quality AI answers for everyone

#82
post #79

Earlier quoted context omitted.

Our fast model, Phind Instant, is completely free

> The model, based on Meta Llama 3.1 8B, runs on a Phind-customized NVIDIA TensorRT-LLM inference server that offers extremely fast speeds on H100 GPUs. We start by running the model in FP8, and also enable flash decoding and fused CUDA kernels for MLP. as far as i know you are running your own GPUs - what do you do in overload? have a queue system? what do you do in underload? just eat the costs? is there a "serverl…

We run the nodes "hot" and close to overload for peak throughput. That's why NVIDIA's XQA innovation was so interesting, because it allows for much higher throughput for a given latency budget: https://github.com/NVIDIA/TensorRT-LLM/blob/main/docs/source....

Serverless would make more sense if we had a significant underutilization problem.

Re: Phind-405B and faster, high quality AI answers for everyone

#84
post #44
post #36

"Phind-405B scores 92% on HumanEval (0-shot), matching Claude 3.5 Sonnet". I'd love to see examples of actual code modifications created by Phind and Sonnet back-to-back. This level of transparency would give me the confidence to try to pro. As it is, I'm skeptical by the claim and actual performance as I've yet to see a finetuned model from Llama3.1 that performed notably better in an area without suffering problems…

I’ve been a customer of Phind for a number of months now, so I’m familiar with the capabilities of all the models they offer. I found even Phind-70B to often be preferable to Claude Sonnet and would commonly opt for it. I’ve been using the 405B today and it seems to be even better at answering. I’ve found it does depend on the task. For instance, for formatting JSON in the past, GPT-4 was actually the best. Because y…

Tbh formatting JSON... should be a solved problem already for the last decade, why consume AI resources for that ??

Re: Phind-405B and faster, high quality AI answers for everyone

#85
post #54

Phind continues to be my favorite AI-enhanced search engine. They do a really nice job giving answers to technical questions with links to references where I can verify the answer or learn more detail. Some recent examples from my history: what video formats does mastodon support? https://www.phind.com/search?cache=jpa8gv7lv54orvpu2c7j1b5j compare xfs and ext4fs https://www.phind.com/search?cache=h9rmhe6ddav1bnb2odtc…

In my tests, it does hallucinate answers, even with Phind 70B. For example, I asked for bluetooth earplugs that have easy battery replacements. It always kept giving me answers for earplugs with I know have their battery soldered into the casing. Tbf, perplexity also fails at this question.

‘Easy battery replacements’ is pretty subjective. This feels like one of those error that demonstrate how good the tech is, because it’s being used for very specific and subjective requests

Re: Phind-405B and faster, high quality AI answers for everyone

#87

Hmm this versus Kagi Assistant? Plan page says $20/mo Unlimited powerful Phind-405B and Phind-70B searches; Daily GPT-4o (500+) , Claude 3.5 Sonnet (500+), Claude Opus (10) uses > Phind-405B scores 92% on HumanEval (0-shot), matching Claude 3.5 Sonnet. Any other benchmarks?

92% suggests a harder benchmark is needed, so it's difficult judge. Especially when a lot of "high scoring" models produce cogent results with a high level of hallucination (eg Llama 3 is chatty, confident and quite often wrong for me).

At that level of performance you're probably in the realm of hard edge cases with ambiguous ground truth.

Re: Phind-405B and faster, high quality AI answers for everyone

#89

I just tried. Asked a question on a research topic I'm digging into. It gave me some answers but no references. Then I copy the answers it gave me and specifically ask for references. Then I got: I sincerely apologize for my earlier response. Upon reviewing the search results provided, I realize I made an error in referencing those specific studies. The search results don't contain any relevant information for the cl…

> As an AI assistant, I should be more careful

I hate this kind of thing so much.

Re: Phind-405B and faster, high quality AI answers for everyone

#90

Earlier quoted context omitted.

Our fast model, Phind Instant, is completely free

Maybe OP was referring to Phind-405B (the model from the article). I certainly wonder how good the 405B model really is.

It's just an innovated (enshittified) version of Facebook's free 405b model.
Post reply on HN