Live data from Hacker News

Viewing profile — mesmertech

mesmertech

HN member
Joined
Thu, Dec 15, 2022, 12:28 PM UTC
HN karma
208
Public activity
110 items

About mesmertech

No profile information was provided.

Recent public activity

  1. comment
    Comment #49156114

    Announcement for the appearance of a commit, thats where we're at

  2. comment
    Comment #49082020

    Why are you guys not on Openrouter? I assume you'd get way more volume that way no? Or does openrouter have like a specific contract you have to sign with them and requirements or …

  3. comment
    Comment #48949263

    And on creativity at least visually, Gemini 3.1 pro is somehow still up there. But its really hindered by its inability to use tool calls effectively or make a long term plan.

  4. comment
    Comment #48949239

    My personal benchmark for new models has been to compare video making skills with something like remotion. Usually reveals if they have any "taste" or outside the box thinking. I'm…

  5. comment
    Comment #48946735

    Hadn't tested out any of the image models for training loras since Flux 1 dev, so I was curious which one is the current best. Results were pretty interesting Models tested: Ideogr…

  6. story
  7. comment
    Comment #48872906

    I'm on max $200, I had this idea for around a month since Fable got banned. For the usage, not exactly sure since I had like 4 agent heavy tasks running at the same time. I think p…

  8. comment
    Comment #48871804

    Had some fable usage to waste yesterday cause my reset was gonna happen. Ended up burning all my usage so had to finish the last part with Sol itself. Imo Remotion based videos fee…

  9. story
  10. story
    Claude weekly usage reset just now

    Was just checking my usage and I see its back to 0. So go nuts

  11. comment
    Comment #48804752

    As long as they have the best model they can afford to lose goodwill. People who don't wanna spend too much on LLMs and are trying to optimize whats subsidized even on the Max plan…

  12. comment
    Comment #48736775

    Ok thats a one month clock to the next Opus model at least, so thats a silver lining to a meh model.

  13. comment
    Comment #48709326

    Spoiler: more than half of it is just Meta ads spending cause its just addictive.

  14. story
  15. comment
    Comment #48576098

    Seems really good at frontend work, and as a result on remotion programmatic videos. Not the best yet, thats still Gemini 3.1 pro(trained on actual videos) or Fable, but often bett…

  16. comment
    Comment #48504101

    noticed it cause of ad manager, the main facebook site being down is kinda weird tho. I don't remember when was the last time that happened

  17. comment
    Comment #48468096

    Overall an improvement over Opus 4.8, but I'd still say Gemini 3.1 Pro has more of an artistic vision even tho it fails tool calls and writes buggy code sometimes. Ik almost everyo…

  18. story
  19. story
    Ask HN: Anyone else seeing serious degradation in DX with Opus 4.8?

    As an anthropic fan boy(check my prev. comments), this is the first opus release where I feel like the model is just not pleasant to talk to not to mention untrustworthy. The two e…

  20. comment
    Comment #48312346

    I think gpt 5.6 is coming out today so might wanna wait

  21. comment
    Comment #48312333

    /model claude-opus-4-8 seems to work but idk why they never set it so you can see it in the /model list. "what model are you I'm Claude Opus (claude-opus-4-8), running in Claude Co…

  22. comment
    Comment #48301358

    Yep sorry was just pulling it out my rear, not like a market trend that nearly every enterprise uses Anthropic or Openai models for coding or that Anthropic has had such ridiculous…

  23. comment
    Comment #48301262

    My point was that even openrouter, the one place people who are looking for open source SOTA models go to, doesn't definitively have opensource models at the top. Esp considering q…

  24. comment
    Comment #48301216

    As long as closed source is 6 months ahead in terms of current difference. Although this is hard to figure out using simple percent based coding benchmarks, you def. notice it when…

  25. comment
    Comment #48301174

    As long as closed models are 6 months ahead I won't be switching from them to prev. 6 month SOTA open source models. Maybe its just a different calculation if you're in a job, but …