> A 9M model can't conditionally follow instructions How many parameters would you need for that?
Show HN: I built a tiny LLM to demystify how language models work
81–90 of 174 posts
My initial idea was to train a navigation decision model with 25M parameters for a Raspberry Pi, which, in testing, was getting about 60% of tool calls correct. IMO, it seems like around 20M parameters would be a good size for following some narrow & basic language instructions.
Re: Show HN: I built a tiny LLM to demystify how language models work
#82Re: Show HN: I built a tiny LLM to demystify how language models work
#83[dead]
Re: Show HN: I built a tiny LLM to demystify how language models work
#84I love this! Seems like it can't understand uppercase letters though
Uppercase letters were intentionally ignored.
Re: Show HN: I built a tiny LLM to demystify how language models work
#85https://bbycroft.net/llm has 3d Visualization of tiny example LLM layers that do a very good job at showing what is going on ( https://news.ycombinator.com/item?id=38505211 )
Pretty neat! I'll definitely take a deeper look into this.
Re: Show HN: I built a tiny LLM to demystify how language models work
#86[flagged]
Meaning/goal of life is to reproduce. Food (and everything else) is only a means to it. Reproduction is the only root goal given by nature to any life form. All resources and qualities are provided are only to help mating.
Re: Show HN: I built a tiny LLM to demystify how language models work
#87This really makes me think if it would be feasible to make an llm trained exclusively on toki pona (https://en.wikipedia.org/wiki/Toki_Pona)
Re: Show HN: I built a tiny LLM to demystify how language models work
#88Why are there so many dead comments from new accounts?
Because despite what HN users seem to think, HN is a LLM-infested hellscape to the same degree as Reddit, if not more.
But what should we do? The parent company isn't transparent about communicating the seriousness of this problem
Re: Show HN: I built a tiny LLM to demystify how language models work
#89Adorable! Maybe a personality that speaks in emojis?
OMG! You just gave me the next idea..
Re: Show HN: I built a tiny LLM to demystify how language models work
#90Would have been funny if it were called "DORY" due to memory recall issues of the fish vs LLMs similar recall issues :)
OMG! Why didn't I thought fo this first :P