Live data from Hacker News

Viewing profile — mattcollins

mattcollins

HN member
Joined
Thu, Jul 15, 2010, 10:51 AM UTC
HN karma
789
Public activity
51 items

About mattcollins

https://www.mattcollins.net/

Recent public activity

  1. story
  2. story
  3. comment
    Comment #47247304

    On the other hand, AI coding tools make it relatively easy to set and apply policies that can help with this sort of thing. I like to have something like the following in AGENTS.md…

  4. story
  5. comment
    Comment #45745563

    Results from some further tests here: https://www.improvingagents.com/blog/toon-benchmarks

  6. comment
    Comment #45730789

    FWIW, I ran a test comparing LLM accuracy with TOON versus JSON, CSV and a variety of other formats when using them to represent tabular data: https://www.improvingagents.com/blog/…

  7. comment
    Comment #45605955

    This is a follow-up to previous work looking at which format of TABULAR data LLMs understand best: https://www.improvingagents.com/blog/best-input-data-format-... (There was some g…

  8. story
  9. comment
    Comment #45581267

    Here you go: https://www.improvingagents.com/blog/best-input-data-format-...

  10. comment
    Comment #45489243

    Author here. This has made me chuckle several times - thanks!

  11. comment
    Comment #45489160

    I did a small test with just a couple of formats and something like 100 records, saw that the accuracy was higher than I wanted, then increased the number of records until the accu…

  12. comment
    Comment #45484385

    I'm the person who ran the test. To hopefully clarify a bit... I intentionally chose input data large enough that the LLM would be scoring in the region of 50% accuracy in order to…

  13. comment
    Comment #45484208

    I'm the person who ran the test. To explain the 60% a bit more... With small amounts of input data, the accuracy is near 100%. As you increase the size of the input data, the accur…

  14. comment
    Comment #45484033

    I'm the person who ran the test. The context I used in the test was pretty large. You'll see much better (near 100%) accuracy if you're using smaller amounts of context. [I chose t…

  15. comment
    Comment #44444995

    "This feature is available to all customers, meaning anyone can enable this today from the Cloudflare dashboard." https://blog.cloudflare.com/control-content-use-for-ai-train...

  16. comment
    Comment #44444931

    I wondered about this, too. Cloudflare have some recent data about traffic from bots ( https://blog.cloudflare.com/from-googlebot-to-gptbot-whos-cr... ) which indicates that, for t…

  17. story
  18. story
  19. story
  20. story
  21. story
  22. story
  23. comment
    Comment #41657579

    I noticed that, too. It does seem 'odd'.

  24. story
  25. story