Live data from Hacker News

Show HN: I built a tiny LLM to demystify how language models work

github.com

101–110 of 174 posts

Re: Show HN: I built a tiny LLM to demystify how language models work

#103
post #92

Earlier quoted context omitted.

Then why are reproductive rates so low in western countries? https://en.wikipedia.org/wiki/List_of_countries_by_total_fer...

The western lifestyle is an evolutionary dead end?

It seems that some in the West want it to be and are working hard to make it so.

Re: Show HN: I built a tiny LLM to demystify how language models work

#105
Cool project. I'm working on something where multiple LLM agents share a world and interact with each other autonomously. One thing that surprised me is how much the "world" matters — same model, same prompt, but put it in a system with resource constraints, other agents, and persistent memory, the behavior changes dramatically. Made me realize we spend too much time optimizing the model and not enough thinking about the environment it operates in.

Re: Show HN: I built a tiny LLM to demystify how language models work

#106
post #6

Earlier quoted context omitted.

> the user is immediately able to understand the constraints Nagel's point was quite literally the opposite [1] of this, though. We can't understand what it must "be like to be a bat" because their mental model is so fundamentally different than ours. So using all the human language tokens in the world can't get us to truly understand what it's like to be a bat, or a guppy, or whatever. In fact, Nagel's point is argu…

Id highly disagree with that. Were all living in the same shared universe, and underlying every intelligence must be precisely an understanding of events happening in this space-time.

What does 'precisely' mean? Everyone has the same understanding of events - a precise one?

Re: Show HN: I built a tiny LLM to demystify how language models work

#108
post #64

How does this compare to Andrej Karpathy's microgpt ( https://karpathy.github.io/2026/02/12/microgpt/ ) or minGPT ( https://github.com/karpathy/minGPT )?

Who cares how it compares, it's not a product it's a cool project

Even cool projects can learn from others. Maybe they missed something that could benefit the project, or made some interesting technical choice that gives a different result.

For the readers/learners, it's useful to understand the differences so we know what details matter, and which are just stylistic choices.

This isn't art; it's science & engineering.

Re: Show HN: I built a tiny LLM to demystify how language models work

#109
post #86
post #50

[flagged]

Meaning/goal of life is to reproduce. Food (and everything else) is only a means to it. Reproduction is the only root goal given by nature to any life form. All resources and qualities are provided are only to help mating.

Reproduction is the goal of genes.

Food (not dying) is the goal of organisms.

Re: Show HN: I built a tiny LLM to demystify how language models work

#110
I don't mean to be 'that guy', but after a quick review, this really feels like low-effort AI slop to me.

There is nothing wrong using AI tools to write code, but nothing here seems to have taken more than a generic 'write me a small LLM in PyTorch' prompt, or any specific human understanding.

The bar for what constitutes an engineering feat on HN seems to have shifted significantly.

Post reply on HN