Viewing profile — m-dot-reviews
m-dot-reviews
HN member- Joined
- Thu, Jun 11, 2026, 4:00 PM UTC
- HN karma
- 7
- Public activity
- 9 items
- HN profile
- View on Hacker News ↗
About m-dot-reviews
Recent public activity
-
comment
Comment #48740995
I've been plugging this perhaps too many times now, but I am trying to bootstrap a user-sourced corpus of exactly "what model is good at task X". So, not benchmarks, but high-level…
-
comment
Comment #48594225
Oops, thanks for telling me that. I think the issue should be fixed now.
-
comment
Comment #48594221
So, this may not be precisely what you're looking for but it may come close. I've put together a simple site for sharing ratings/opinions on models on a task-specific granularity. …
-
comment
Comment #48570018
For anyone who's interested, I've put together a simple site for sharing ratings/opinions on models at a task-specific granularity. https://model.reviews/ The idea is that benchmar…
-
comment
Comment #48498480
I'm starting a repository of LLM reviews [1] with the goal of creating a catalog that is more task-oriented and less marketing-y than corporate blogs or benchmark leaderboards. You…
-
comment
Comment #48492970
Anecdotally, yes there is definitely a difference. Even e.g. Haiku (cheapest Anthropic model) vs gpt-oss-120b had a big difference in quality and syntax issues when I was testing t…
-
comment
Comment #48492892
[dead]
-
comment
Comment #48492257
I looked for a forum like this a few months ago during my own model research, and didn't find one. So, here's the "catalog of clankers," a task-structured review site for LLMs. The…
- story