Live data from Hacker News

Llama 3.2: Revolutionizing edge AI and vision with open, customizable models

ai.meta.com

281–290 of 343 posts

Re: Llama 3.2: Revolutionizing edge AI and vision with open, customizable models

#281

Llama3.2 3B feels a lot better than other models with same size (e.g. Gemma2, Phi3.5-mini models). For anyone looking for a simple way to test Llama3.2 3B locally with UI, Install nexa-sdk( https://github.com/NexaAI/nexa-sdk ) and type in terminal: nexa run llama3.2 --streamlit Disclaimer: I am from Nexa AI and nexa-sdk is an open-sourced. We'd love your feedback.

It's a great tool. Thanks! I had to test it with Llama3.1 and was really easy. At a first glance Llama3.2 didn't seem available. The command you provided did not work, raising "An error occurred while pulling the model: not enough values to unpack (expected 2, got 1)".

Thanks for reporting. We are investigating this issue. Could you help submit an issue to our GitHub and provide a screenshot of the terminal (with pip show nexaai)? This could help us reproduce this issue faster. Much appreciated!

Re: Llama 3.2: Revolutionizing edge AI and vision with open, customizable models

#282

Zuckerberg has never liked having Android/iOs as gatekeepers i.e. "platforms" for his apps. He's hoping to control AI as the next platform through which users interact with apps. Free AI is then fine if the surplus value created by not having a gatekeeper to his apps exceeds the cost of the free AI. That's the strategy. No values here - just strategy folks.

[deleted]

Re: Llama 3.2: Revolutionizing edge AI and vision with open, customizable models

#283

Llama3.2 3B feels a lot better than other models with same size (e.g. Gemma2, Phi3.5-mini models). For anyone looking for a simple way to test Llama3.2 3B locally with UI, Install nexa-sdk( https://github.com/NexaAI/nexa-sdk ) and type in terminal: nexa run llama3.2 --streamlit Disclaimer: I am from Nexa AI and nexa-sdk is an open-sourced. We'd love your feedback.

or grab lmstudio

Re: Llama 3.2: Revolutionizing edge AI and vision with open, customizable models

#284

"The Llama jumped over the ______!" (Fence? River? Wall? Synagogue?) With 1-hot encoding, the answer is "wall", with 100% probability. Oh, you gave plausibility to "fence" too? WRONG! ENJOY MORE PENALTY, SCRUB! I believe this unforgiving dynamic is why model distillation works well. The original teacher model had to learn via the "hot or cold" game on text answers. But when the child instead imitates the teacher's pr…

I would have went with “moon”

Re: Llama 3.2: Revolutionizing edge AI and vision with open, customizable models

#285
post #94

Newbie question, what size model would be needed to have a 10x software engineer skills and no knowledge of the human kind (ie, no need to know how to make a pizza or sequence your DNA). Is there such a model?

Try codegemma.

Or Gemini Flash for code completion and generation.

Re: Llama 3.2: Revolutionizing edge AI and vision with open, customizable models

#286

Earlier quoted context omitted.

Can it speak foreign languages like German, Spanish, Ancient Greek?

Yes. It can converse perfectly normal in German. However when quizzed about German idioms it hallucinates them (in fluent German). Though that's the kind of stuff even larger models often have trouble with. For example if you ask GPT 4 about jokes in German it will give you jokes that depend on word play that only works when translated to English. In normal conversation Llama seems to speak fluent German For Ancient…

“However Llama did add a cheeky "Πηγή: Google Translate" at the end”

That’s interesting; could this be an indicator that someone is running content through GT and training on the results?

Re: Llama 3.2: Revolutionizing edge AI and vision with open, customizable models

#287
post #186

Tried the 1B model with the "think step by step" prompt. It gets "which is larger: 9.11 or 9.9?" right if it manages to mention that decimals need to be compared first in its step-by-step thinking. If it skips mentioning decimals, then it says 9.11 is larger. It gets the strawberry question wrong even after enumerating all the letters correctly, probably because it can't properly count.

Does anyone know of a CoT dataset somewhere for finetuning? I would think exposing it to that type of modality during a finetune/lora would help.

Re: Llama 3.2: Revolutionizing edge AI and vision with open, customizable models

#288
post #190

It's hallucinating so badly, it's kinda hilarious Literally everything about the quote below is wrong. The problem with these "amazing ground-breaking models" is that you never know which knowledge they actually possess. --- start quote --- >>> what is the last version of Ash Framework for Elixir you know My training data goes up to December 2023, and according to the Ash Framework releases, the last stable version t…

Wildly incorrect

Re: Llama 3.2: Revolutionizing edge AI and vision with open, customizable models

#289

"The Llama jumped over the ______!" (Fence? River? Wall? Synagogue?) With 1-hot encoding, the answer is "wall", with 100% probability. Oh, you gave plausibility to "fence" too? WRONG! ENJOY MORE PENALTY, SCRUB! I believe this unforgiving dynamic is why model distillation works well. The original teacher model had to learn via the "hot or cold" game on text answers. But when the child instead imitates the teacher's pr…

Moat

Re: Llama 3.2: Revolutionizing edge AI and vision with open, customizable models

#290
post #229

Earlier quoted context omitted.

You're objectively correct but judging from your downvotes there seems to be some denial here about that! The atheism alone means it's different from a big chunk of the world's population, possibly the majority. Supposedly around 80% of the world's population identify with a religion though I guess you can debate how many people are truly devout. The good news is that the big AI labs seem to be slowly getting a grip…

You want the computer to believe in God?

No I'm just agreeing that it's not 'aligned' with the bulk of humanity if it doesn't believe in some god. I'm happy for it to be agnostic on the issue, personally. So you have to be careful what alignment means.
Post reply on HN