Live data from Hacker News

Viewing profile — markab21

markab21

HN member
Joined
Mon, Mar 31, 2014, 1:54 PM UTC
HN karma
179
Public activity
62 items

About markab21

No profile information was provided.

Recent public activity

  1. comment
    Comment #49123005

    We use it at a moderate scale, self-hosted on B300 hardware. It's great :D QA analysis of voice transcriptions. Napkin math: we operate at 2-5% of the cost of running on Equiv Fron…

  2. comment
    Comment #49042475

    Yes - THIS! I can't even believe how exhausting it is to read. I'm not sure why or what changed in Fable. Did they do this writing-style output to give it more token compression du…

  3. comment
    Comment #48623970

    I'm mildly surprised that more people aren't using Nemo models for this reason. We've moved most of our processing to a combination of Nemo Ultra and Super, with some support for m…

  4. comment
    Comment #47620217

    I'll pipe in here as someone working on an agentic harness project using mastra as the harness. Nemotron3-super is, without question, my favorite model now for my agentic use cases…

  5. comment
    Comment #47491199

    I think the entire premise that the prompting is the surface area for optimizing the application is fundamentally the wrong framing, in the same way that in 1998 better cpam will s…

  6. comment
    Comment #47075920

    You just articulated why I struggle to personally connect with Gemini. It feels so unrelatable and exhausting to read its output. I prefer to read Opus/Deepseek/GLM over Gemini, Qw…

  7. comment
    Comment #46945989

    And I think you basically just described the OpenAI approach to building models and serving them.

  8. comment
    Comment #46903235

    Shaking fist at clouds!!

  9. comment
    Comment #46873238

    It's getting a lot easier to do this using sub-agents with tools in Claude. I have a fleet of Mastra agents (TypeScript). I use those agents inside my project as CLI tools to do re…

  10. comment
    Comment #46591283

    Basically looking for emergent behavior.

  11. comment
    Comment #46402397

    I love where you're going with this. In my experience it's not about a different persona, it's about constantly considering context that triggers, different activations enhance a d…

  12. comment
    Comment #45494830

    The skepticism is understandable given the trajectory of GPTs and custom instructions, but there's a meaningful technical difference here: the Apps SDK is built on the Model Contex…

  13. comment
    Comment #40996235

    It looks like it was built jointly with nvidia: https://huggingface.co/nvidia/Mistral-NeMo-12B-Instruct

  14. comment
    Comment #39685865

    I bet you have been waiting years to pull that one out of your pocket. Well played sir! Nice shot man! :D

  15. comment
    Comment #39361667

    I've found myself more and more using local models rather than ChatGPT; it was pretty trivial to set up Ollama+Ollama-WebUI, which is shockingly good. I'm so tired of arguing with …

  16. comment
    Comment #39208810

    For Llama-based progress - Reddit - /r/LocalLlama has been my top source of info, although it's been getting a little more noisy lately. I also hang out on a few Discord servers: -…

  17. comment
    Comment #37934035

    [flagged]

  18. comment
    Comment #37483859

    Yeah, slow news day.

  19. comment
    Comment #37162294

    Linux distributions, including Debian, offer a variety of desktop environments, each with its own design philosophy and user experience. If one environment doesn't suit your prefer…

  20. comment
    Comment #37162254

    Debian and Ubuntu have similarities, but keep in mind - Ubuntu is derived from Debian, not the other way around. However, they differ in areas like release cycles, package manageme…

  21. comment
    Comment #33954074

    I used it as a consultant on a development project to help me organize some of the milestones and design goals in some documentation. It wasn't that I didn't know the stuff, I do, …

  22. comment
    Comment #29947013

    I'd be surprised if they don't have async mechanisms.

  23. comment
    Comment #29397053

    Assume anything sent over a cellular network carrier via normal SMS can not only be retrieved, but intercepted.

  24. comment
    Comment #23067794

    Why anyone would use Oracle for anything other than supporting legacy systems is beyond me.

  25. comment
    Comment #20263976

    The claim is dead-right. I own a Tesla Model 3, my wife drives a BMW i3, my daughter has a leaf. The ONLY car we can effectively travel outside of the greater Tampa area without ma…