Live data from Hacker News

Viewing profile — funfunfunction

funfunfunction

HN member
Joined
Sun, Dec 04, 2016, 6:02 PM UTC
HN karma
356
Public activity
129 items

About funfunfunction

80eafd64915ed77b50fd68efed69bfa51407b135a5f7b093006e175e4c23296a

Recent public activity

  1. comment
    Comment #48655053

    There's some benchmarks in the repo for AppWorld. Looks promising

  2. comment
    Comment #48653394

    Cool project. A team at work was building something similar to internal use. I'm curious how this compares to just using Claude Code directly and giving it a dump of the agent trac…

  3. story
  4. story
  5. comment
    Comment #47767850

    Hi all, we built a super easy way to train small language models from your production data - install a gateway to save your request/response data from a frontier provider, and use …

  6. story
  7. story
    Show HN: Project AELLA – Open LLMs for structuring 100M research papers

    We're releasing Project AELLA - an open-science initiative to make scientific knowledge more accessible through AI-generated structured summaries of research papers. Blog: https://…

  8. story
  9. comment
    Comment #45669534

    We'll release the full data explorer soon, with more info. At the core of this project is a structured-extraction task using a custom Qwen 14B model, which we distilled from larger…

  10. story
  11. comment
    Comment #45634099

    Creator of inference.net / schematron here. There is growing emphasis on efficiency as more companies adopt and scale with LLMs in their products. Developers might be fine paying G…

  12. story
  13. comment
    Comment #44925408

    OP here. I wanted a dead-simple way to quickly generate CLI commands without the overhead of Claude Code or Cursor, so I built it in an afternoon. The project uses some zsh magic t…

  14. story
  15. story
    Show HN: UwU – Generate CLI commands inline with GPT-5

    I wanted a dead-simple way to quickly generate CLI commands without the overhead of Claude Code or Cursor, so I built it in an afternoon. The project uses some zsh magic to allow f…

  16. story
  17. comment
    Comment #44677354

    This is a cheap marketing ploy for a GPU reseller with billboards on highway 101 into SF.

  18. story
  19. comment
    Comment #44671292

    There are even companies starting to offer distillation as a service https://inference.net/explore/model-training

  20. story
  21. story
  22. story
  23. comment
    Comment #41252851

    This is cool! I don’t see many people doing write ups on their tech stack as much any more. It’s nice to see the inside of a production-grade app like this. I’m curious, why comman…

  24. comment
    Comment #41241921

    It’s unlikely an individual would need this much capacity. Folks who need tokens at this level are apps with lots of users that don’t have their own GPUs. Think character.ai type a…

  25. story