Live data from Hacker News

Viewing profile — m-dot-reviews

m-dot-reviews

HN member
Joined
Thu, Jun 11, 2026, 4:00 PM UTC
HN karma
7
Public activity
9 items

About m-dot-reviews

https://model.reviews | Fleshy human HN poster

Recent public activity

  1. comment
    Comment #48740995

    I've been plugging this perhaps too many times now, but I am trying to bootstrap a user-sourced corpus of exactly "what model is good at task X". So, not benchmarks, but high-level…

  2. comment
    Comment #48594225

    Oops, thanks for telling me that. I think the issue should be fixed now.

  3. comment
    Comment #48594221

    So, this may not be precisely what you're looking for but it may come close. I've put together a simple site for sharing ratings/opinions on models on a task-specific granularity. …

  4. comment
    Comment #48570018

    For anyone who's interested, I've put together a simple site for sharing ratings/opinions on models at a task-specific granularity. https://model.reviews/ The idea is that benchmar…

  5. comment
    Comment #48498480

    I'm starting a repository of LLM reviews [1] with the goal of creating a catalog that is more task-oriented and less marketing-y than corporate blogs or benchmark leaderboards. You…

  6. comment
    Comment #48492970

    Anecdotally, yes there is definitely a difference. Even e.g. Haiku (cheapest Anthropic model) vs gpt-oss-120b had a big difference in quality and syntax issues when I was testing t…

  7. comment
  8. comment
    Comment #48492257

    I looked for a forum like this a few months ago during my own model research, and didn't find one. So, here's the "catalog of clankers," a task-structured review site for LLMs. The…

  9. story