Phi-2 was useless for practical purposes except if you want to show your friends that it can write a poem, llama3 8b was slightly better but is still same category, it’s complete trash with coding vs gpt4. Llama3 400b “iS OPen SoURce!” But no you will need to pay to access because most one can not practically afford an A100 and set it up properly. What I’m trying to say is that user experience is now as key as the mo…
Phi-3 Technical Report
111–120 of 132 posts
Re: Phi-3 Technical Report
#112If I was Apple I'd be quaking in my boots. They are getting too far behind to ever catch up. Nokia in 2010 vibes.
Apple's advantage is that their devices are safeguarding people from the dangers of AI
Are you next going to tell us that the CIA's access to iCloud data protects their users from terrorism too?
Re: Phi-3 Technical Report
#113Earlier quoted context omitted.
Even llama3 has its issues. Ive been quite impressed so far but if the context gets a little long it freaks out, gets stuck repeating the same token or just fails to finish an answer. This is for the full f16 8B model, so it cant be put down to quantization. It also doesnt quite handle complex instructions as well as the benchmarks would imply should.
Supposedly LLMs (especially smaller ones) are best suited to tasks where the answer is in the text, i.e. summarization, translation, and answering questions. Asking it to answer questions on its own is much more prone to hallucination. To that end I've been using Llama 3 for summarizing transcripts of YouTube videos. It does a decent job, but... every single time (literally 100% of the time), it will hallucinate a ra…
Re: Phi-3 Technical Report
#114Earlier quoted context omitted.
5 years? 5 years is a millennia these days. We’ll have small local models beating gpt-4/Claude opus in 2024. We already have sub 100b models trading blows with former gpt-4 models, and the future is racing toward us. All these little breakthroughs are piling up.
Absolutely not on the first one. Not even close.
Re: Phi-3 Technical Report
#115Hm, roundabout 84 authors of one "scientific" paper. I wonder if this says something about (a) the quality of its content, (b) the path were academic (?) paper publishing goes to, (c) nothing at all, or (d), something entirely else.
It costs so little to share the credit if someone was an asset.
Re: Phi-3 Technical Report
#116Earlier quoted context omitted.
Even llama3 has its issues. Ive been quite impressed so far but if the context gets a little long it freaks out, gets stuck repeating the same token or just fails to finish an answer. This is for the full f16 8B model, so it cant be put down to quantization. It also doesnt quite handle complex instructions as well as the benchmarks would imply should.
Supposedly LLMs (especially smaller ones) are best suited to tasks where the answer is in the text, i.e. summarization, translation, and answering questions. Asking it to answer questions on its own is much more prone to hallucination. To that end I've been using Llama 3 for summarizing transcripts of YouTube videos. It does a decent job, but... every single time (literally 100% of the time), it will hallucinate a ra…
"{text}\n\n###\n\nPlease summarize the text above. The text is a video transcript. It may not have the names of the speakers in it. If you need to refer to an unnamed speaker, call them Speaker_1, Speaker_2 and so on."
Re: Phi-3 Technical Report
#117Both precious phi have been epic letdowns when I actually tried them myself so quite low confidence in this being reflective of real world. Will try it anyway though
Re: Phi-3 Technical Report
#118They have started putting some models in huggingface: https://huggingface.co/collections/microsoft/phi-3-6626e15e9...
Re: Phi-3 Technical Report
#119Re: Phi-3 Technical Report
#120Earlier quoted context omitted.
Absolutely not on the first one. Not even close.
Why not? There's still 7 months left for breakthroughs.