Live data from Hacker News

Viewing profile — rchaves

rchaves

HN member
Joined
Tue, Feb 05, 2019, 1:54 PM UTC
HN karma
432
Public activity
142 items

About rchaves

No profile information was provided.

Recent public activity

  1. story
  2. story
  3. comment
    Comment #44367786

    Hello HN! tl;dr: We built Scenario, an open-source testing library for AI agents. It simulates real conversations with your agent, its code-driven, and lets you assert anything mid…

  4. comment
    Comment #43771619

    well I think hype is not bad per se, I'd do it even if not trying to make a buck, it's okay (up to a point) to hype up something so that eventually it finds a problem where it fits…

  5. comment
    Comment #43771568

    same here, but I would even avoid "strong arguments" because that's what we all have been doing so far what I want is real use cases, show me real-world production examples from es…

  6. comment
    Comment #43771517

    is this multi-agent collaboration though, or is it just a workflow? All examples you listed seem to have pretty deterministic control flows (write then validade, context exceeded, …

  7. story
  8. story
  9. story
  10. comment
    Comment #42492641

    Nah it's just a marketing problem, "GPT" and "ChatGPT" names is the biggest asset OpenAI has, people have expectations so high for GPT-5 that they cannot burn this name unless it's…

  11. story
    Discuss HN: Agents are the new object-oriented programming

    From @lateinteraction on twitter: In so many ways, good and bad, (LLM-based) Agents are the new object-oriented programming. Part of it is that: there's nothing that you can't do w…

  12. story
  13. story
  14. story
  15. story
  16. story
  17. story
  18. comment
    Comment #42353950

    Erm, he wrote the article with “you” to invoke the feeling of the reader thinking about their own use case, which I did Different because I ran without good practices before, got m…

  19. comment
    Comment #42353881

    Yes exactly, from the experience he had as Facebook massively scaling up, while all good practices were thrown down the window (except for foundation and critical parts) and extrem…

  20. comment
    Comment #42352319

    I’ve seen people spending 10 minutes to test things by hand, would have taken them less to write and run a test, specially with AI now When writing test actually makes it faster to…

  21. comment
    Comment #42352296

    Not true. Kent Back’s 3X is a much better take, test and good practices for what is high risk and hard to change, move fast for most of it on the rest to try to find that black swa…

  22. comment
    Comment #41526927

    yeah I guess base models without built-it CoT are not going away, exactly because you might want to tune it yourself. If DSPy (or similar) evolves to allow the same or similar than…

  23. story
  24. comment
    Comment #38884250

    Nope, the trains are not often late, this is just in Germany

  25. comment
    Comment #38642078

    If you look closely it actually does give multiple instructions per screenshot! However it cannot get too far, because the screen changes under it. For example when it starts typin…