Viewing profile — michaelhartm
michaelhartm
HN member- Joined
- Fri, Jul 04, 2014, 5:40 AM UTC
- HN karma
- 27
- Public activity
- 9 items
- HN profile
- View on Hacker News ↗
About michaelhartm
No profile information was provided.
Recent public activity
-
comment
Comment #36316100
Btw, semantics and syntax is separated in the LLMs (the author is wrong). The embedding function (matmul) can map syntax and the proximity in the embedding (e.g. cosine similarity)…
-
comment
Comment #36315347
Nobody knows how the LLMs work under the hood. It's just lots of stacked transformers that encode various concepts. Nothing in this book refutes whether Chomsky's concepts are actu…
-
comment
Comment #35539389
I guess Databricks is now going after OpenAI?
-
comment
Comment #35294087
Btw, it's kinda crazy how bad the GPT4-J results in the blog are compared to the Dolly one, which seem pretty good. Do we know why it works so well to use this 50k dataset?
-
comment
Comment #35294040
They used the 6b GPT4-J, not 20B. That's what's interesting, it's a smallish large language model :).
-
comment
Comment #29232654
Data Wars: Snowflake vs Databricks (0 - 2)?
-
comment
Comment #29208014
* Databricks is unethical * Nobody should benchmark anymore, just focus on customers instead * But hey, we just did some benchmarks and we look better than what Databricks claims *…
- comment
- story