Live data from Hacker News

Viewing profile — diwank

diwank

HN member
Joined
Sun, Sep 11, 2011, 6:05 PM UTC
HN karma
2,117
Public activity
441 items

About diwank

Dropped out of Columbia (physics, philosophy) for a Thiel Fellowship. Built Julep AI — open-source AI agent framework, 7k stars on GitHub — because I wanted agents that could actually do things. Now building Memory Store (YC X26): persistent memory for AI agents. Turns out the hard problem isn’t making agents smart, it’s making them remember.

Still fascinated by what foundational models and cognitive architectures teach us about our own minds. Building beats theorizing, but sometimes they’re the same thing.

https://memory.store

https://diwank.space

hi@diwank.space

Recent public activity

  1. comment
    Comment #49231771

    you lose CoT monitorability which is a big issue since models have become quite powerful and also often deceptive but i do think that efficiency pressure will keep nudging us towar…

  2. story
    Ask HN: Which is the least sloppy and claudeism free model you have used?

    I feel like recent models have been consistently getting more sloppy and increasingly claude-ism heavy (load-bearing seams galore) with every new release. I was hoping this trend w…

  3. comment
    Comment #48984895

    [flagged]

  4. comment
    Comment #48897466

    this is surprisingly high delta. to make matters worse, reasoning tokens account for the majority of tokens and they are completely opaque so it's hard to tell how much of that is …

  5. comment
    Comment #48850328

    i'm not happy with how openai is trying to pit 5.6 sol as a cheaper equivalent to fable here for one thing, they said that on AA, sol is "within one point of fable" at 58.9 vs 59.9…

  6. comment
    Comment #48742277

    where did you find that? weird coz their post announcing this also mentioned Claude Code: > Fable 5 will be available starting tomorrow, Wednesday, July 1, to users globally on the…

  7. story
  8. comment
    Comment #48404890

    """ It remains unclear whether Anthropic’s engineers are assisting the NSA in active operations. However, one person close to the situation said Mythos would be useful for infiltra…

  9. story
  10. comment
    Comment #48090676

    in order for us to get there, i think we need a standardized api at the os layer for local models so that the os could optimize, batch and safely allocate resources. something like…

  11. comment
    Comment #47929758

    and device dependent. this is a very tricky thing to get rendered consistently

  12. comment
    Comment #47929756

    this is actually a surprisingly rich area of debate in philosophy of mind. see: https://plato.stanford.edu/entries/qualia-inverted/

  13. story
  14. story
  15. comment
    Comment #47596013

    had a bad experience with pg_search (paradedb) in the past

  16. comment
    Comment #47596006

    we have been using pg_textsearch in production for a few weeks now, and it's been fairly stable and super speedy. we used to use paradedb (aka pg_search -- it's quite annoying that…

  17. comment
    Comment #47539596

    this is so disingenuous on symbolica's part. these insincere announcements just make it harder for genuine attempts and novel ideas

  18. story
    Show HN: Datetime-bench: which datetime formats LLMs get right (and wrong)

    tl;dr * If you need an LLM to parse OR emit a timestamp, use: RFC 3339 ( e.g. 2024-03-26 10:30:00-05:00 ) * python date format also works well * Do NOT use unix epoch or javascript…

  19. comment
    Comment #47521609

    Angels & Demons anyone?

  20. story
  21. comment
    Comment #47031935

    opus 4.6 gets it right more than half the times

  22. story
    Show HN: IQT – Why space feels panoramic and time feels fleeting

    I've spent the past year building a theory of phenomenal experience (consciousness) that's designed to be falsifiable. It identifies phenomenal quality with a specific mathematical…

  23. comment
    Comment #46941107

    Working on Memory Store: persistent, shared memory for all your AI agents. https://memory.store The problem: if you use multiple AI tools (Claude, ChatGPT, Cursor, etc.), none of t…

  24. comment
    Comment #46879865

    I dont think this is Cerebras. Running on cerebras would change model behavior a bit and it could potentially get a ~10x speedup and it'd be more expensive. So most likely this is …

  25. story