Live data from Hacker News

Viewing profile — sethkim

sethkim

HN member
Joined
Sat, Jan 23, 2021, 10:25 PM UTC
HN karma
417
Public activity
98 items

About sethkim

Founder of Sutro (https://sutro.sh/).

seth@sutro.sh

Recent public activity

  1. story
  2. comment
    Comment #47491914

    This is extremely true. In fact, from what we see many/most of the problems to be solved with LLMs do not have ground-truth values; even hand-labeled data tends to be mostly subjec…

  3. comment
    Comment #47491524

    Feel free to shoot me a note at seth@sutro.sh if you want to check it out!

  4. comment
    Comment #47491489

    We build a product that's somewhat similar in spirit to DSPy, but people come to us for different reasons than the OP listed here. 1) It's slow: you first have to get acquainted wi…

  5. story
  6. comment
    Comment #45650018

    Under-discussed superpower of LLMs is open-set labeling, which I sort of consider to be inverse classification. Instead of using a static set of pre-determined labels, you're using…

  7. comment
    Comment #45389964

    The models you called out at the beginning were all released this year. What do you think is the difference between this generation of models and previous ones?

  8. comment
    Comment #44458100

    Yes! Both Llama 3 and Gemma 3 have 128k context windows.

  9. comment
    Comment #44458078

    Yes, we're a startup! And LLM inference is a major component of what we do - more importantly, we're working on making these models accessible as analytical processing tools, so we…

  10. comment
    Comment #44458059

    My two cents here is the classic answer - it depends. If you need general "reasoning" capabilities, I see this being a strong possibility. If you need specific, factual information…

  11. comment
    Comment #44457994

    No doubt prices will continue to drop! We just don't think it will be anything like the orders-of-magnitude YoY improvements we're used to seeing. Consequently, developers shouldn'…

  12. comment
    Comment #44457905

    Both great points, but more or less speak to the same root cause - customer usage patterns are becoming more of a driver for pricing than underlying technology improvements. If so,…

  13. story
  14. comment
    Comment #44302932

    I run a batch inference/LLM data processing service and we do a lot of work around cost and performance profiling of (open-weight) models. One odd disconnect that still exists in L…

  15. comment
    Comment #44164721

    Sutro.sh (fka Skysight) | Infrastructure/LLMs & Research Engineering | SF Bay Area | Full-time We are building batch inference infrastructure and a great/user developer experience …

  16. comment
    Comment #43864168

    Skysight | Infrastructure/LLMs & Research Engineering | SF Bay Area | Full-time We are building large-scale batch inference infrastructure and a great/user developer experience aro…

  17. comment
    Comment #43722614

    How "huge" are these datasets? Did you build your own tooling to accomplish this?

  18. comment
    Comment #43710694

    Thanks for the reply and the notes. On 4. specifically we've got some thoughts here as well. Will reach out!

  19. story
  20. story
  21. story
  22. story
  23. story
  24. comment
    Comment #43245243

    Skysight | Infrastructure/LLMs & Product Engineering | SF Bay Area | Full-time We are building large-scale, data-intensive inference tooling and a great/user developer experience a…

  25. comment
    Comment #42891163

    What's extremely confusing to me (as a private pilot) is that traffic is almost always routed directly over an airport (midfield), to safely avoid departing and landing traffic. Th…