Live data from Hacker News

Show HN: I built a tiny LLM to demystify how language models work

github.com

81–90 of 174 posts

Re: Show HN: I built a tiny LLM to demystify how language models work

#81
post #75

> A 9M model can't conditionally follow instructions How many parameters would you need for that?

My initial idea was to train a navigation decision model with 25M parameters for a Raspberry Pi, which, in testing, was getting about 60% of tool calls correct. IMO, it seems like around 20M parameters would be a good size for following some narrow & basic language instructions.

Re: Show HN: I built a tiny LLM to demystify how language models work

#85
post #61

https://bbycroft.net/llm has 3d Visualization of tiny example LLM layers that do a very good job at showing what is going on ( https://news.ycombinator.com/item?id=38505211 )

Pretty neat! I'll definitely take a deeper look into this.

Re: Show HN: I built a tiny LLM to demystify how language models work

#86
post #50

[flagged]

Meaning/goal of life is to reproduce. Food (and everything else) is only a means to it. Reproduction is the only root goal given by nature to any life form. All resources and qualities are provided are only to help mating.

Re: Show HN: I built a tiny LLM to demystify how language models work

#88
post #56

Why are there so many dead comments from new accounts?

Because despite what HN users seem to think, HN is a LLM-infested hellscape to the same degree as Reddit, if not more.

But what should we do? The parent company isn't transparent about communicating the seriousness of this problem
Post reply on HN